freehire launches on Product Hunt on 26 August.

Follow →

Data Steward

Summary

Build and maintain cloud-based data pipelines for life-sciences analytics using Python, PySpark, and AWS, ensuring data quality and FAIR principles to power medical and sales insights.

Job Description Summary

#LI-Hybrid
Location: Hyderabad

Are you ready to shape the future of data engineering in life sciences? As a Data Engineer Manager, you’ll lead the data engineering team and support the data science and/or reporting & analytics team inbuilding scalable Commercial solutions to enhance the medical/sales representative actionable. You’ll be at the forefront of cloud-based data integration, driving automation, quality, and efficiency while collaborating with cross-functional teams to deliver impactful insights. This is your opportunity to make a real difference in global healthcare through cutting-edge data engineering.


Job Description

Major Accountabilities:

  • Design scalable data ingestion and integration solutions to support data products for data science and reporting & analytics.
  • Ensure data quality by applying business and technical rules throughout its lifecycle.
  • Identify and implement automation opportunities to streamline data processes.
  • Build data pipelines using Python and CI/CD workflows for seamless data integration.
  • Collaborate with solution architects and vendors to align with best practices.
  • Apply data management principles including modelling, harmonization, and ontology standards.
  • Manage metadata effectively and leverage enterprise ontology tools.
  • Ensure adherence to FAIR data principles across applicable projects.
  • Conduct feasibility assessments and define project requirements with stakeholders.
  • Support end-user training and promote self-service data capabilities.

Minimum Requirements:

  • University degree in Informatics, Computer Sciences, Life Sciences, or a related field.
  • 1-4 years of experience in data engineering with good understanding of healthcare or life sciences. Experience with Commercial is a plus.
  • Proven expertise in Python, PySpark and R for ETL and BI data product development.
  • Strong experience with DevOps, AWS cloud data integration and third-party ingestion tools.
  • Proficiency in SQL for relational databases such as Oracle and MS SQL Server.
  • Proficiency in Cloud based databases like Snowflake is a plus.
  • Hands-on experience with ETL tools like Alteryx and BI platforms like Power BI.
  • Solid understanding of data architecture, modelling, and analytics concepts.
  • Familiarity with Agile methodologies in global project environments.
  • Understanding of Data science concepts is a plus

Why Novartis: Helping people with disease and their families takes more than innovative science. It takes a community of smart, passionate people like you. Collaborating, supporting, and inspiring each other. Combining to achieve breakthroughs that change patients’ lives. Ready to create a brighter future together. https://www.novartis.com/about/roadmap/people-and-culture

Commitment to Diversity & Inclusion:

Novartis is committed to building an outstanding, inclusive work environment and diverse team’s representative of the patients and communities we serve.

Values and Behaviors: Demonstrates and upholds Novartis values and behaviors in all aspects of work and collaboration.

Location: Hyderabad NKC. Hybrid | 3 days a week in office is mandatory.


Skills Desired

Alteryx, Analytical Thinking, Brand Awareness, Business Networking, Digital Marketing, Media Campaigns, Process Documentation, Statistical Analysis

See also