freehire launches on Product Hunt on 26 August.

Follow →

Data Engineer (Databricks)

Summary

Build and maintain scalable data pipelines on Azure and Databricks for a client project, processing large data volumes to support analytics and reporting.

In Cyclad we work with top international IT companies in order to boost their potential in delivering outstanding, cutting-edge technologies that shape the world of the future. We are looking for a Data Engineer to join a client project focused on building and optimizing modern data platforms on Azure and Databricks. The role involves designing, developing, and maintaining scalable data pipelines, processing large volumes of data, and supporting business-critical analytics solutions.

Project information:

  • Type of project: IT Services

  • Office location: Wrocław - Swojczyce

  • Work model: Hybrid mode – 2 days per week in the office in Wrocław

  • Budget: 125 - 145 PLN net/ h - b2b

  • Project length: till the end of 2026, possible to extend it

  • Only candidates with citizenship in the European Union and residence in Poland

  • Start date: ASAP

Project scope:

  • Build and optimize modern data platforms on Azure and Databricks

  • Design, develop, and maintain scalable data pipelines

  • Build and optimize data lakehouse solutions supporting analytics, reporting, and regulatory needs

  • Implement data ingestion from multiple sources (databases, APIs, files, messaging/ event streams)

  • Perform data transformation, validation, and quality checks across end‑to‑end data flows

  • Collaborate with Business Analysts and stakeholders to translate requirements into data solutions

  • Support production environments, handle incident analysis, and contribute to performance optimization

  • Ensure compliance with data governance, security, and regulatory standards

Competence demands:

  • Min 4 YoE in Data Engineering

  • 2+ years of hands-on experience with Databricks and Delta Lake in production environments

  • 2+ years of experience developing and maintaining Azure Data Factory (ADF) pipelines

  • Very strong Python, Spark SQL, PySpark programming skills

  • Experience with Azure Data Lake Storage Gen2 (ADLS Gen2)

  • Knowledge of Medallion architecture, dimensional models

  • REST APIs, batch and streaming ingestion

  • GitHub Copilot, CI/CD pipelines, Azure DevOps or similar

  • Strong understanding of data quality, performance tuning, and scalability

  • Strong understanding of data modeling, ETL/ELT processes, and modern data architecture principles

  • Very good English skills

Nice to have:

  • PostgreSQL integration

  • Time-series data processing (GPS/parameter data)

  • Slowly Changing Dimensions (SCD) patterns

  • Spark performance tuning (SLA-driven optimization)

  • Adaptive Query Execution (AQE) configuration

  • CDC (Change Data Capture) and incremental loading

  • SSIS migration patterns

We offer:

  • Hybrid working model

  • Dynamic and innovation-driven engineering environment

  • Full-time job agreement based on b2b

  • Private medical care with dental care (covering 70% of costs)

  • Multisport card (also for an accompanying person)

  • Life insurance