Data Engineer
Ready to reshape a data platform from the ground up? Join Us!
As a leading AWS Partner in Europe, we support organizations at every stage of their cloud journey. Our strength comes from deep experience in cloud-native architectures. By joining our team, you'll have the opportunity to design and build scalable data platforms on AWS, work hands-on with modern data stack technologies, and directly shape how businesses in regulated industries process, store, and make sense of their data.
Join our Data & AI team and work on a project where the decisions you make today become the foundation others build on tomorrow. As a Data Engineer, you'll join an insurance industry data platform project - an AWS-native solution that recently shipped its MVP and is now entering a phase of serious architectural evolution. You'll be rebuilding it the right way: from the ground up, with the right tools, the right patterns, and full ownership over what you ship.
What we're building right now:
Migrating from direct-to-Redshift streaming to a proper data lake architecture
Introducing batch loading alongside existing streaming pipelines
Replacing Step Functions with Airflow (AWS MWAA) as the orchestration layer
Moving workloads from ECS to EKS for better scalability and control
Working with Kafka + Debezium for CDC, dltHub for ingestion, dbt + Cosmos for transformations, and OpenMetadata for data governance
We work in a "we build it, we run it" model - every engineer on the team can deploy their solutions to production without handing off to anyone else.
Your Responsibilities:
Design and implement the new data lake architecture on AWS
Build and maintain ELT pipelines using dbt, dltHub, and Airflow
Migrate existing orchestration from Step Functions to Airflow (MWAA)
Ensure data quality, observability, and lineage across the platform
Deploy your own solutions to production - no handoffs, full ownership
Collaborate closely with a small, senior engineering team and contribute to technical decisions
Requirements:
2+ years of experience in data engineering, with production-grade delivery
Hands-on experience with dbt and at least one modern orchestration tool (Airflow preferred)
Solid Python skills and experience building and operating data pipelines on AWS
Familiarity with data lake concepts and batch/streaming architectures
Comfort with Infrastructure as Code (Terraform or similar)
Self-driven and able to operate with high autonomy - you don't wait to be told what to do next
Fluency in both English (C1) and Polish (C1), written and spoken
Nice to have:
Experience with Kafka, Debezium, or other CDC tooling
Familiarity with dltHub or similar ingestion frameworks
Hands-on experience with Amazon Redshift and data warehouse optimization
Background with OpenMetadata or other data catalog/governance tools
Experience with healthcare datasets (FHIR, OMOP, DICOM)
Benefits:
Continuous learning and growth – we invest in your development through internal training, knowledge-sharing sessions, and hands-on learning from real projects. We also support growth beyond the organization by actively engaging in the AWS Community, including conference talks and industry events.
Training budget – dedicated funds to support certifications, courses, and professional development aligned with your career goals.
Medical care package – comprehensive private healthcare to help you stay healthy and focused.
Multisport card co-financing – because staying active matters, both in and outside of work.
Language learning platform access – improve your language skills at your own pace, whenever it suits you.
Flexible working hours & remote work – work when and where you’re most productive.
Company events – time for integration and some fun.
Stay up to date with us: Follow our blog regularly and participate in the events we organize.