Data Steward
novartis
Job Description
- Design scalable data ingestion and integration solutions to support data products for data science and reporting & analytics.
- Ensure data quality by applying business and technical rules throughout its lifecycle.
- Identify and implement automation opportunities to streamline data processes.
- Build data pipelines using Python and CI/CD workflows for seamless data integration.
- Collaborate with solution architects and vendors to align with best practices.
- Apply data management principles including modelling, harmonization, and ontology standards.
- Manage metadata effectively and leverage enterprise ontology tools.
- Ensure adherence to FAIR data principles across applicable projects.
- Conduct feasibility assessments and define project requirements with stakeholders.
- Support end-user training and promote self-service data capabilities.
Minimum Requirements:
- University degree in Informatics, Computer Sciences, Life Sciences, or a related field.
- 1-4 years of experience in data engineering with good understanding of healthcare or life sciences. Experience with Commercial is a plus.
- Proven expertise in Python, PySpark and R for ETL and BI data product development.
- Strong experience with DevOps, AWS cloud data integration and third-party ingestion tools.
- Proficiency in SQL for relational databases such as Oracle and MS SQL Server.
- Proficiency in Cloud based databases like Snowflake is a plus.
- Hands-on experience with ETL tools like Alteryx and BI platforms like Power BI.
- Solid understanding of data architecture, modelling, and analytics concepts.
- Familiarity with Agile methodologies in global project environments.
- Understanding of Data science concepts is a plus