Data Engineer: Python, Spark on AWS EMR + ML Ops
Summary
Michael Page International (HK) is recruiting a Data Engineer in Hong Kong to design, build, and maintain scalable cloud data pipelines, optimizing Spark workloads on AWS EMR for batch and real-time processing. The core stack is Python, SQL, and AWS services including S3, Glue, Athena, Redshift, and Lambda.
Michael Page International (HK) Ltd in Hong Kong is seeking a Data Engineer to design, develop, and maintain scalable data pipelines on a cloud-based platform. You will optimize Spark workloads on AWS EMR for both batch and real-time processing and enable self-service analytics with high-quality, governed datasets.
The role requires 3–5 years in data engineering, strong Python and SQL skills, plus experience with S3, Glue, Athena, Redshift, and Lambda.