Senior/Lead Data Engineer (AWS,Python, SQL, PySpark)
Summary
Senior/Lead Data Engineer in Hyderabad (hybrid) designing, building, and optimizing scalable ETL/ELT pipelines on AWS for healthcare/life-science analytics. Core stack: Python/Scala, PySpark, SQL, AWS services (S3, Glue, Athena, Redshift, EMR, Lambda), with Terraform IaC and Jenkins CI/CD.
Job Description
Role: Senior / Lead Data Engineer (Python, PySpark, SQL, AWS) Experience: 6–12 years Location: Hyderabad Work Mode: Hybrid (3 days/week in-office) Open Positions: 5 Join Time: Immediate Domain : Healthcare / Life Sciences Must-Have Technical Skills: - Strong programming skills in Python and/or Scala - Hands-on experience with Apache Spark for big data processing on AWS cloud - Proficiency with AWS services such as S3, AWS Glue, Athena, Redshift, EMR, Lambda - Strong SQL skills for data transformation and analytics - Experience in Infrastructure as Code (IaC) using Terraform - Expertise in setting up and managing CI/CD pipelines with Jenkins Responsibilities: - Design, build, and optimize scalable ETL/ELT data pipelines on AWS - Implement data ingestion, transformation, and integration solutions using Spark, AWS Glue, and SQL - Manage and optimize cloud storage and compute environments - Ensure robust, automated deployments with Terraform and Jenkins - Collaborate with cross-functional teams to deliver high-quality data products Nice to Have: - Prior experience in the Healthcare / Life Sciences domain - Familiarity with modern data lake and data mesh architectures Why Join Us? - Work on cutting-edge data engineering projects in healthcare analytics - Hybrid work model for flexibility and collaboration - Opportunity to grow in a fast-paced, innovation-driven environment Apply Now! Send your updated resume to careers@sidinformation.com