AI Backend Engineer
Summary
Build and run the backend systems that turn AI models into fast, reliable APIs for users, focusing on inference pipelines, orchestration, and production reliability.
You will build and operate production systems that turn model capability into fast, stable, observable APIs used across mobile and desktop clients.
Job Responsibilities:
- Build and operate backend systems that serve AI-powered features in production.
- Design inference pipelines, orchestration layers, and service boundaries around models.
- Own production concerns: monitoring, logging, alerting, and incident response.
- Optimize latency and throughput across inference, caching, batching, and streaming.
Job Requirements:
- Strong backend engineering fundamentals in production environments.
- Experience running high-throughput, low-latency services.
- Familiarity with AI inference patterns (LLMs, embeddings, multimodal).
- Comfortable debugging distributed systems under load.
- Bias toward shipping and learning from production behaviour.
Required Tech Stacks:
- Python
- Node.js
- Pytorch
- OpenAI / Anthropic / open-source LLMs
- SQL & noSQL
- Kubernetes
- Docker