Software Data Engineer (AWS)
Summary
Design and implement data ingestion and processing pipelines on AWS to transform manufacturing data. You will build streaming (Kafka/Flink) and batch (Spark/Iceberg) pipelines, manage AWS infrastructure via Terraform, and move prototypes to production while ensuring observability and quality.
In this role you will design and implement data ingestion and processing pipelines to enable a data-driven transformation of manufacturing data. You will join the AMS team to support global incident handling and enhancement requests for ingestion processes, collaborating with the Milan team to align requirements and loading pipelines. You will deploy and operate data and ML workloads in AWS, moving from prototype to production while ensuring quality and observability. This position offers hands-on work with cutting-edge data tech in a globally connected, innovation-driven environment.
Retribuzione / Benefits- Ticket restaurant
- Flexible working hours
- Hybrid Remote Working
- Employee benefits
- Design and build streaming data pipelines using Kafka/MSK and stream processing (Flink/Spark Streaming)
- Develop batch transformations into a medallion lakehouse on S3 (Bronze/Silver/Gold) with Iceberg/Parquet
- Manage data storage across relational (Aurora/RDS), NoSQL (DynamoDB), and lakehouse layers
- Design, deploy, and manage AWS infrastructure for data/ML workloads via IaC (EKS, Fargate, Lambda, VPC)
- Migrate prototypes to production with refactoring, optimization, CI/CD, and best practices
- Ensure observability, data quality, and maintenance of pipelines in production
- Master's in Computer Science/Engineering
- Strong SQL and Python
- Streaming pipelines with Kafka/MSK and Flink or Spark Streaming
- AWS hands-on: EKS, Fargate, Lambda, S3, Aurora/RDS, DynamoDB, VPC
- Big-data stack (Spark, Hive, HDFS) and lakehouse formats (Iceberg/Parquet)
- Git, CI/CD, Docker, IaC (Terraform), REST APIs; agile/scrum
- Strong communication
- Curiosity about learning new technologies
- Good organizational and time management skills
- Kafka/MSK
- Flink
- Spark Streaming