Data Engineer
Summary
Data Engineer responsible for designing and maintaining ETL workflows using Informatica PowerCenter, developing data pipelines with Python, Spark, Hive, and HDFS, and optimizing SQL queries for various databases. Works with enterprise data platforms in an Agile environment.
Key Responsibilities:
. Design, develop, and maintain ETL workflows using Informatica PowerCenter.
. Develop data pipelines using Python, Spark, Hive, and HDFS.
. Perform data extraction, transformation, and loading (ETL/ELT) from multiple source systems.
. Write and optimize SQL queries for Oracle, Teradata, or SQL Server databases.
. Prepare technical design documents, mapping specifications, and data flow documentation.
. Perform performance tuning, trouble shooting, and production support.
. Coordinate with Business Analysts, QA, and business stakeholders throughout the project lifecycle.
. Participate in Agile/Scrumceremonies and contribute to sprint planning and delivery.
. Support production deployments and incident resolution.
RequiredSkills:
Must Have
. Informatica PowerCenter
. SQL (Oracle/Teradata/SQL Server)
. Python
. Apache Spark
. Hive & HDFS
. ETL/Data Warehousing concepts
. Unix Shell Scripting
. Production Support
. Agile/Scrum
Good to Have
. Apache Sqoop
. Apache NiFi
. Control-M or AutoSys
. Git/Bitbucket
. Jenkins
. Banking or Financial Services domain experience
Preferred Qualifications
. Bachelor's degree in Engineering, Computer Science, Information Technology, or a related field.
. 8-12 years of experience in ETL and Data Engineering.
. Strong analytical, problem-solving, and communication skills.
. Experience working with large-scale enterprise data platforms.