freehire launches on Product Hunt on 26 August.

Follow →

Full Stack Engineer, AI systems

Summary

Builds and operates backend systems that power AI-driven workflows, focusing on inference pipelines, orchestration, and production reliability for a proactive smart assistant.

About A1

There are over 5 billion users using basic applications today such email, notes, tasks that are not AI-native. Our mission is to build a proactive smart assistant for everyday users to bring intelligence to conversations, errands, organising and workflows, with minimal prompting.

Our product focuses on achieving high reliability for long-running workflows, persistent context, and real-world task completion. The system must handle multi-step reasoning, interact with external tools, and remain reliable despite non-deterministic model behavior. Our objective is to help users complete tasks daily enjoyable with over ~90%* reduced time.

Role

We are looking for a Full Stack Engineer - AI Systems to build the product layer that turns these capabilities into usable, production-grade workflows. This includes designing how agents operate, fail, recover, and deliver consistent value to users.

Focus

  • Build and operate backend systems that serve AI-powered features in production.

  • Design inference pipelines, orchestration layers, and service boundaries around models.

  • Own production concerns: monitoring, logging, alerting, and incident response.

  • Optimize latency and throughput across inference, caching, batching, and streaming.

Ideal Experiences

  • Strong backend engineering fundamentals in production environments.

  • Experience running high-throughput, low-latency services.

  • Familiarity with AI inference patterns (LLMs, embeddings, multimodal).

  • Comfortable debugging distributed systems under load.

  • Bias toward shipping and learning from production behavior.

Outcomes

  • Backend systems run reliably at scale, handling production AI traffic with low latency and high throughput.

  • APIs are stable, clear, and support seamless integration with frontend and ML systems.

  • Production incidents are quickly detected, diagnosed, and resolved, minimizing user impact.

  • Iterative improvements based on real usage continuously increase system performance and reliability.

Tech Stack

  • Python

  • NodeJs

  • Pytorch

  • OpenAI / Anthropic / open-source LLMs

  • SQl & noSQL

  • Kubernetes

  • Docker

How We Work

The best products today in the world were built by small, world class teams. We are a high talent density and hands-on team. We make decisions collectively, move at rapid speed, striking a balance between shipping high quality work and learning. Joining our team requires the ability to bring structure, exercise judgment, and execute independently. Our goal is to put in hands of our users a truly magical product

Interview process

If there appears to be a fit, we'll reach to schedule 3, but no more than 4 interviews.

Applications are evaluated by our technical team members. Interviews will be conducted via virtual meetings and/or onsite.

We value transparency and efficiency, so expect a prompt decision. If you've demonstrated the exceptional skills and mindset we're looking for, we'll extend an offer to join us. This isn't just a job offer; it's an invitation to be part of a team that's bringing AI to have practical benefits to billions globally.

What this application asks

ashby

Name, Email, Where are you currently based?, Resume

  • Phone number
  • Linkedin URL optional
  • GitHub URL optional
  • Do you need a Work VISA to work in the country where this job is located? yes / no
  • Consent choose any
  • Which best describes your experience building backend systems in production? choose one
  • What is your hands-on experience with deploying and executing AI systems in a live production environment? choose one
  • Which production coWhat is your level of English proficiency for professional and technical communication?ncerns have you personally handled?  choose one

See also