freehire launches on Product Hunt on 26 August.

Follow →

Kubernetes ML Inference Engineer: Model Serving

Open 17d reposted 2× · 2 open copies

Summary

Builds and optimizes Kubernetes-based infrastructure to deploy and serve AI models in real-time for healthcare applications.

Abridge is seeking an ML Infrastructure Engineer, Model Inference in San Francisco to build and optimize the core inference infrastructure powering our AI‑driven healthcare solutions. You will collaborate across Infrastructure and Research to deploy, optimize, and orchestrate AI models at scale.

Ideal candidates have 2+ years of production ML infrastructure experience, strong Kubernetes know‑how, and a track record of engineering scalable APIs and distributed systems for real-time workloads.

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available