Senior ML engineer building secure, on-premise AI infrastructure for autonomous agents used by government and enterprise clients. Day to day: LLM gateway/routing, RAG pipelines, guardrails, observability, inference optimization, and GPU scaling using Python/FastAPI, LangChain/LlamaIndex, vLLM/Ollama, Qdrant, Docker/Kubernetes and CUDA.
Sign in to see your match