Data Engineer
Summary
The Data Engineer will design, develop, and maintain ETL workflows and data pipelines using Informatica PowerCenter, Python, Spark, and Hive. The role involves optimizing SQL queries, performing production support, and collaborating within an Agile environment.
Key Responsibilities:
. Design, develop, and maintain ETL workflows using Informatica PowerCenter.
. Develop data pipelines using Python, Spark, Hive, and HDFS.
. Perform data extraction, transformation, and loading (ETL/ELT) from multiple source systems.
. Write and optimize SQL queries for Oracle, Teradata, or SQL Server databases.
. Prepare technical design documents, mapping specifications, and data flow documentation.
. Perform performance tuning, trouble shooting, and production support.
. Coordinate with Business Analysts, QA, and business stakeholders throughout the project lifecycle.
. Participate in Agile/Scrumceremonies and contribute to sprint planning and delivery.
. Support production deployments and incident resolution.
RequiredSkills:
Must Have
. Informatica PowerCenter
. SQL (Oracle/Teradata/SQL Server)
. Python
. Apache Spark
. Hive & HDFS
. ETL/Data Warehousing concepts
. Unix Shell Scripting
. Production Support
. Agile/Scrum
Good to Have
. Apache Sqoop
. Apache NiFi
. Control-M or AutoSys
. Git/Bitbucket
. Jenkins
. Banking or Financial Services domain experience
Preferred Qualifications
. Bachelor's degree in Engineering, Computer Science, Information Technology, or a related field.
. 8-12 years of experience in ETL and Data Engineering.
. Strong analytical, problem-solving, and communication skills.
. Experience working with large-scale enterprise data platforms.