Data Engineer - Databricks & Informatica (IDMC)
Summary
The Data Engineer will modernize and scale cloud data platforms by building ETL/ELT pipelines using Databricks, Informatica (IDMC), and AWS. The role involves migrating legacy workflows to PySpark and SQL while managing data governance and infrastructure performance.
About the role
We are looking for a skilled Data Engineer to help modernize, scale, and optimize our cloud data platform. In this role, you will build robust data pipelines, migrate legacy workflows, and manage enterprise datasets using AWS, Databricks, and Informatica.
Key responsibilities
- Design and maintain scalable batch/real-time ETL/ELT pipelines using Informatica (IDMC) and Databricks.
- Implement Delta Lake medallion architecture (Bronze/Silver/Gold) and data governance via Unity Catalog on AWS.
- Convert legacy Informatica/on-prem ETL workflows into modern PySpark and SQL cloud solutions.
- Fine-tune Spark performance, manage AWS cloud infrastructure costs, and automate CI/CD deployments.
About you
- Proficiency in Databricks (PySpark, Delta Lake)
- Proficiency in Informatica (IDMC)
- Proficiency in AWS (S3, Glue, Lambda, Redshift)
- Strong proficiency in Python and SQL
- 3-6 years of hands‑on data engineering, data modeling, and cloud pipeline experience