Build and operate internal AI platforms for research and engineering teams — backends, dashboards, CLIs/SDKs, GPU cluster scheduling, and model serving on managed inference and self-operated GPUs. Core stack is production Python, Kubernetes, Slurm, CI/CD, and MLOps/LLMOps tooling, based on-site in San Francisco.
Sign in to see your match