AI Backend Engineer: Scalable Inference & Orchestration
Summary
Build and maintain the inference and orchestration layer for AI models, ensuring fast, reliable APIs for mobile and desktop clients with a focus on latency, correctness, and cost.
A1 is seeking a Backend Engineer, AI, to own the inference and orchestration layer powering all AI interactions in the product. You will sit between models and users, focusing on latency, correctness, reliability and cost to deliver fast, stable APIs across mobile and desktop clients.
You will build and operate production systems that turn model capability into highly available APIs used across platforms, ensuring scalable, observable services with robust monitoring and incident response.