Senior Data Engineer
Summary
Design and build scalable ETL pipelines on AWS using PySpark, Glue, and Step Functions to transform and move data for enterprise clients across finance, healthcare, retail, and travel.
ITC Infotech a wholly owned ITC ltd company is a leading global technology services and solutions provider, led by Business and Technology Consulting. ITC Infotech provides business-friendly solutions to help clients succeed and be future-ready, by seamlessly bringing together digital expertise, strong industry specific alliances and the unique ability to leverage deep domain expertise from ITC Group businesses. The company provides technology solutions and services to enterprises across industries such as Banking & Financial Services, Healthcare, Manufacturing, Consumer Goods, Retail, Travel and Hospitality, through a combination of traditional and newer business models, as a long-term sustainable partner.
Your X-Factor
- Work ethic - You are a consummate professional
- Aptitude - You have an innate capacity to transition from project to project without skipping a beat.
- Communication - You have excellent written and verbal communication skills for coordination across projects and teams.
- Impact - You are a critical thinker with an emphasis on creativity and innovation.
- Passion - You have the drive to succeed paired with a continuous hunger to learn.
- Leadership - You are trusted, empathetic, accountable, and empower others around you
Requirement Analysis Solution Design
Collaborate with business and data stakeholders to understand data requirements, transformation rules, and use cases
Translate functional requirements into scalable data engineering and ETL design solutions
Define approach for data ingestion, transformation, orchestration, and access control
Design and develop robust ETL pipelines using AWS Glue PySpark
Implement complex data transformations including joins, aggregations, conditional logic, and business rule processing
Build reusable, metadata-driven ETL frameworks where applicable
Optimize data pipelines for performance, scalability, and reliability
Perform unit testing and end-to-end validation of data pipelines
Implement data quality checks including integrity, completeness, and reconciliation
Validate transformation logic and ensure accuracy of aggregated and processed data
3. Orchestration Integration
Develop and manage workflow orchestration using AWS Step Functions and EventBridge
Implement event-driven and schedule-based pipeline execution strategies
Ensure seamless integration between data pipelines, control frameworks, and downstream systems
4. Infrastructure Deployment IaC
Provision and manage cloud infrastructure using Terraform Infrastructure as Code
Deploy and configure AWS services including Glue, Lambda, DynamoDB, and orchestration components
Ensure consistent, repeatable, and scalable deployments aligned with DevOps practices
6. Security Governance
Implement secure data access controls using IAM and Lake Formation
Ensure compliance with data governance policies and role-based access requirements
Manage data security, encryption, and access auditing
Set up monitoring, logging, and alerting mechanisms for data pipelines e.g., SNS, audit logs
Troubleshoot data and pipeline issues, ensuring timely resolution
Continuously enhance pipeline performance, reliability, and maintainability
Collaborate in an agile setup to drive continuous delivery and improvements
Role Summary :
Must-Have Skills
- 7-10 years of AWS Data Engineering experience.
- Strong experience in AWS Glue PySpark and data processing on S3
- Proven expertise in ETL development and complex data transformations
- Hands-on experience with AWS Step Functions, EventBridge, and orchestration patterns
- Proficiency in Terraform Infrastructure as Code
- Strong SQL and data modeling skills
- Experience with data validation, reconciliation, and quality frameworks
- Solid understanding of IAM, security, and cloud best practices
Nice-to-Have Skills
- Familiarity with Lake Formation and data governance frameworks
- Experience with DynamoDB or metadata-driven ETL architectures
- Exposure to event-driven architectures
- Knowledge of CICD tools
ITC Infotech is an Equal Opportunity Employer. We believe that no one should be discriminated against because of their differences, such as age, disability, ethnicity, gender, gender identity and expression, religion, or sexual orientation. All employment decisions shall be made without regard to age, race, creed, color, religion, sex, national origin, ancestry, disability status, veteran status, sexual orientation, gender identity or expression, genetic information, marital status, citizenship status or any other basis as protected by federal, state, or local law. ITC infotech is committed to providing veteran employment opportunities to our service men and women