Data Engineer (Databricks)
Summary
Build and maintain scalable data pipelines on Azure and Databricks for a client project, processing large data volumes to support analytics and reporting.
In Cyclad we work with top international IT companies in order to boost their potential in delivering outstanding, cutting-edge technologies that shape the world of the future. We are looking for a Data Engineer to join a client project focused on building and optimizing modern data platforms on Azure and Databricks. The role involves designing, developing, and maintaining scalable data pipelines, processing large volumes of data, and supporting business-critical analytics solutions.
Project information:
Type of project: IT Services
Office location: Wrocław - Swojczyce
Work model: Hybrid mode – 2 days per week in the office in Wrocław
Budget: 125 - 145 PLN net/ h - b2b
Project length: till the end of 2026, possible to extend it
Only candidates with citizenship in the European Union and residence in Poland
Start date: ASAP
Project scope:
Build and optimize modern data platforms on Azure and Databricks
Design, develop, and maintain scalable data pipelines
Build and optimize data lakehouse solutions supporting analytics, reporting, and regulatory needs
Implement data ingestion from multiple sources (databases, APIs, files, messaging/ event streams)
Perform data transformation, validation, and quality checks across end‑to‑end data flows
Collaborate with Business Analysts and stakeholders to translate requirements into data solutions
Support production environments, handle incident analysis, and contribute to performance optimization
Ensure compliance with data governance, security, and regulatory standards
Competence demands:
Min 4 YoE in Data Engineering
2+ years of hands-on experience with Databricks and Delta Lake in production environments
2+ years of experience developing and maintaining Azure Data Factory (ADF) pipelines
Very strong Python, Spark SQL, PySpark programming skills
Experience with Azure Data Lake Storage Gen2 (ADLS Gen2)
Knowledge of Medallion architecture, dimensional models
REST APIs, batch and streaming ingestion
GitHub Copilot, CI/CD pipelines, Azure DevOps or similar
Strong understanding of data quality, performance tuning, and scalability
Strong understanding of data modeling, ETL/ELT processes, and modern data architecture principles
Very good English skills
Nice to have:
PostgreSQL integration
Time-series data processing (GPS/parameter data)
Slowly Changing Dimensions (SCD) patterns
Spark performance tuning (SLA-driven optimization)
Adaptive Query Execution (AQE) configuration
CDC (Change Data Capture) and incremental loading
SSIS migration patterns
We offer:
Hybrid working model
Dynamic and innovation-driven engineering environment
Full-time job agreement based on b2b
Private medical care with dental care (covering 70% of costs)
Multisport card (also for an accompanying person)
Life insurance