Senior Data Engineer (Databricks)
The Role
We're looking for a hands-on Senior Data Engineer to own the technical design, reliability and evolution of our Databricks data foundation. You'll work alongside Clinical Data Insights Analysts — colleagues with strong business, clinical and BI expertise who also write SQL and Python — complementing their skills with deeper engineering, automation and production-grade rigour.
Responsibilities
- Own the architecture, reliability and governance of the team's Databricks lakehouse (Delta Lake, medallion architecture, Unity Catalog).
- Design and operate batch and streaming ingestion pipelines from the CluePoints platform and other business systems (e.g. Zendesk).
- Co-develop SQL, Python and PySpark transformations with the analyst team; review and optimize for performance, scalability and cost.
- Productionize analytical and AI prototypes — adding testing, orchestration, monitoring, error handling and deployment controls.
- Build automated data-quality controls, monitoring and alerting; investigate and resolve pipeline incidents.
- Implement technical governance: naming conventions, metadata, lineage, tagging and access controls via Unity Catalog.
- Apply AI and automation across the data lifecycle — code development, documentation, metadata classification and inconsistency detection.
- Collaborate with Engineering on source-system access and with Product when Databricks insights are candidates for platform integration.
Requirements
- 5+ years in data engineering or data-platform engineering.
- Strong hands-on Databricks experience (or comparable cloud lakehouse platform).
- Advanced SQL and strong Python / PySpark skills.
- Production-grade pipeline experience — design, build, test, deploy and maintain.
- Solid understanding of Delta Lake, medallion architecture, and OLAP/OLTP transformation.
- Experience with data governance, Unity Catalog (or equivalent), automated quality controls and CI/CD.
- At least one major cloud platform: Azure, AWS or GCP.
- Practical use of generative AI in engineering workflows, with solid critical judgement on its outputs.
- High autonomy — you take ownership from design through deployment, monitoring and maintenance.
- Collaborative by nature; comfortable working directly alongside analysts who also write code.
Nice to Have:
- Lakeflow Jobs / Lakeflow Declarative Pipelines; Databricks Asset Bundles.
- Databricks AI/BI, Genie, conversational analytics or AI agent experience.
- Databricks Data Engineer certification.
- Experience in a regulated environment (clinical trials, life sciences or similar).
Tech Stack:
Databricks · Delta Lake · Unity Catalog · Lakeflow Jobs & Declarative Pipelines · Medallion architecture · SQL / Python / PySpark · Microsoft Azure · Git / CI/CD · Databricks Asset Bundles · Databricks AI/BI · Genie
Benefits
🇧🇪 What We Offer – Belgium- Health Insurance through Alan (100% hospitalisation cover, 80% ambulatory and dental)
- Mobility Budget for eco-transport, housing, or car allowance (flexible 3-pillar system)
- Group Insurance Plan with 6–12% employer pension contribution based on seniority
- Meal Vouchers (€8/day) and Eco Vouchers for sustainable purchases
- A hub-based hybrid model that blends flexibility with purpose — connecting teams through collaboration, learning, and a vibrant social culture.
Equal Opportunities & GDPR Notice
CluePoints is an equal opportunities employer. We value and respect diversity in our workforce and do not tolerate discrimination based on gender, age, disability, ethnic origin, religion, sexual orientation, or any other protected ground under Belgian law.
Personal data collected as part of your application will be processed in compliance with the EU GDPR and Belgian data protection legislation.
You have the right to access, correct, or delete your personal data at any time by contacting privacy@cluepoints.com.