Data Engineer - AWS Ecosystem (office OR remote)
Summary
Designs and maintains scalable AWS-based data pipelines and cloud architectures to transform raw data into reliable datasets for analytics and business intelligence, using AWS Glue, EMR, Redshift, and related tools.
We are looking for a skilled and proactive Data Engineer with deep expertise in theAmazon Web Services (AWS) ecosystemto join our data team. You will be responsible for designing, building, and optimizing scalable data pipelines and cloud architectures that power analytics and business intelligence across the organization. Your role is central to transforming raw data into reliable, high-quality datasets using modern tools likeAWS Glue,Amazon EMR, andAmazon Redshift.
LOCATION:
POLAND - Flexibility to work either from our Krakow/ Wroclaw offices or from a home office - the choice is up to you!
RESPONSIBILITIES:
- Develop and maintain complex ETL/ELT pipelines usingAWS GlueandAWS Step Functionsto ingest and transform data from diverse sources.
- Design and implement scalable data storage solutions usingAmazon S3(Data Lake) and relational databases such asAmazon RDSorAmazon Aurora.
- UtilizeAmazon EMR(with Spark/Hadoop),AWS Glue Studio, andAmazon Redshiftto process large-scale structured and unstructured datasets.
- Optimize performance of data pipelines, ensuring cost-efficiency and high availability.
- Troubleshoot and monitor workflows to ensure data integrity and system reliability (usingAmazon CloudWatch,AWS X-Ray, andAWS CloudTrail).
- Collaborate with Data Architects and Business Analysts to translate requirements into technical solutions.
- Implement CI/CD practices for automating the deployment of data infrastructure and pipeline logic.
REQUIREMENTS:
- 4+ years of experience as a Data Engineer or in a similar cloud-focused role.
- Knowledge ofAWS Services, specificallyAWS Glue,Amazon EMR,Amazon Redshift,AWS Lambda,Amazon S3, andAWS DMS(Database Migration Service).
- Strong proficiency inSQLand experience withPython,Scala, orPySparkfor data transformation.
- Solid understanding of Database Management, including relational data modeling,Amazon RDS,Aurora.
- Hands-on experience withGit (or equivalent version control systems) for version control and CI/CD.
- Proven ability to work with Big Data processing and building robust ETL/ELT architectures.
- Experience with cloud storage (S3) and partitioning strategies.
- AWS Certified Data Engineer - Associate(or higher) certification is required. If you do not currently hold this certification, we expect you to obtain it within your first 1 month with us.
NICE TO HAVE:
- Experience with migration from on-premises or other cloud systems to Cloud.
- Knowledge of Data Governance principles and tools within the AWS environment (e.g.,AWS Lake Formation,Glue Data Catalog).
- Understanding of Infrastructure as Code (IaC) usingAWS CloudFormationorTerraform.
- Experience working with other cloud platforms, such asMicrosoft Azure,Google Cloud Platform (GCP), orOracle Cloud, demonstrating the ability to adapt quickly to different cloud ecosystems and transfer data engineering best practices across environments.
WE OFFER YOU:
- Available cooperation models: UoP or B2B
- Office space or 100% remote - whichever suits you best!
- Communication in English - only foreign customers, and international Teams
- Simple structure and 'open door' way of communication
- Full-time English teachers
- Medical insurance for employees
- HiQo University- internal education and training programs
- HiQo Coins - system of rewarding employees for extracurricular activities