Platform Data Engineer: Databricks, RAG & Feature Store

Summary

The Platform Data Engineer will stream clinical data into Databricks to curate feature tables for AI training and manage RAG infrastructure, including vector databases and retrieval pipelines.

Geisinger is seeking a senior data engineer to stream data from Epic SDE, ADT feeds, and clinical sources into Databricks, and curate shared clinical feature tables for AI model training and monitoring.

You will own RAG infrastructure, including ingestion, chunking, and embedding pipelines, administer vector databases, and build retrieval pipelines with hybrid search and reranking. Strong Databricks and real-time data skills are required.

See also

Data Engineering jobs by country — openings, pay and top skills →

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available