freehire launches on Product Hunt on 26 August.

Follow →

LLM Systems / AI Agent Engineer (Remote, Full-Time) [AS311]

Summary

LLM Systems/AI Agent Engineer builds and evolves production AI agents on foundation models (AWS Bedrock), focusing on orchestration, context engineering, evaluation pipelines, and production observability. Requires 2+ years of experience with production LLM agents and backend engineering proficiency with TypeScript/Node, Postgres, and serverless AWS.

About Smart Working

At Smart Working, we believe your job should not only look right on paper but also feel right every day. We're one of the highest-rated workplaces on Glassdoor, connecting exceptional professionals with outstanding global teams and products through long-term remote opportunities.

Our mission is to break down geographical barriers and create meaningful career opportunities where talented individuals can thrive, grow, and make a genuine impact. When you join Smart Working, you become part of a supportive and collaborative community that values integrity, excellence, continuous learning, and professional growth. We provide the tools, support, and environment needed to help you succeed while enjoying the flexibility of a truly remote-first workplace.

About the Role

As an LLM Systems / AI Agent Engineer, you will focus on building and evolving production AI agents on foundation models, covering orchestration, context engineering, evaluation pipelines, and production observability and monitoring. This is an LLM systems and AI agent engineering position rather than a traditional ML model-training role. You will be the first dedicated AI-agenting hire, working directly alongside the engineer who currently leads this function. The product currently uses a custom orchestration layer, and your knowledge of agent architecture patterns will help inform whether to continue with this approach or adopt a production framework such as LangGraph or LangChain. Roadmap timings are somewhat tentative, but you will be assigned measurable deliverables immediately upon onboarding.

Responsibilities

  • Build and evolve production AI agents on foundation models, currently AWS Bedrock.
  • Develop and evolve orchestration for production AI agents.
  • Apply context engineering to production LLM and agent systems.
  • Build and maintain evaluation datasets and pipelines, including tool-selection, trajectory and LLM-as-judge evaluations.
  • Build and maintain production observability and monitoring for LLM and agent systems.
  • Implement and work with tracing and instrumentation for production LLM systems.
  • Work directly alongside the engineer currently leading the AI-agenting function as the first dedicated hire in this area.
  • Apply strong agent-architecture fundamentals to help inform whether the existing custom orchestration layer should be retained or a production framework adopted.
  • Work as part of a new three-person product team alongside the AI function lead and a Full Stack Engineer.
  • Operate as an individual contributor working alongside the AI function lead rather than managing others.
  • Take ownership of measurable deliverables immediately upon onboarding.
  • Deliver similar AI/agent engineering work against roadmap timelines.

Requirements

  • 2+ years of experience building production LLM agents, including tool-calling agent loops, streaming, context management, structured outputs and orchestration frameworks.
  • Production experience with an agentic framework such as LangGraph, LangChain or custom orchestration. There is no fixed orchestration framework requirement; strong agent-architecture fundamentals and production experience with any agentic framework are in scope. - 1.5+ years of experience with evaluation-driven development, including building and maintaining evaluation datasets and pipelines covering tool selection, trajectory evaluation and LLM-as-judge.
  • 1+ year of experience with LLM observability, tracing and instrumentation using Langfuse, OpenTelemetry or similar tooling.
  • Genuine production agent-observability exposure. Direct, hands-on Langfuse experience is strongly preferred because this is a confirmed skill gap within the team; OpenTelemetry or other tracing tools are acceptable only as a secondary signal alongside real agent-observability exposure. - 1+ year of experience with LLM cost optimisation, including prompt caching, model selection and routing, and LLM FinOps.
  • 5+ years of backend engineering proficiency, including TypeScript/Node, Postgres and serverless AWS.
  • Proven experience delivering similar work on similar timelines.
  • Experience shipping agentic AI systems to production, with the ability to speak to concrete failure modes and mitigations and operate end-to-end across the AI stack.

Nice to Have

  • 1+ year of AWS Bedrock experience. Equivalent production experience with other foundation-model providers, including OpenAI, Anthropic API, Azure OpenAI or Vertex AI, is fully transferable.
  • 1+ year of experience with AI safety and guardrails, including prompt-injection screening, output validation and handling untrusted input.
  • 6+ months of familiarity with MCP and multi-agent patterns.
  • Familiarity with geospatial data.

Benefits

  • Fixed Shifts: 12:00 PM - 9:30 PM IST (Summer) | 1:00 PM - 10:30 PM IST (Winter)
  • No Weekend Work: Real work-life balance, not just words
  • Day 1 Benefits: Laptop and full medical insurance provided
  • Support That Matters:Mentorship, community, and forums where ideas are shared
  • True Belonging: A long-term career where your contributions are valued

What this application asks

lever

In which location did you find the job?, Resume/CV, Full name, Email, Phone, Current location, Current company, LinkedIn URL, Twitter URL, GitHub URL, Portfolio URL

  • Please respond truthfully. How much professional experience do you have building production LLM agents, including tool-calling agent loops, streaming, context management, structured outputs and orchestration frameworks? choose one
  • Please respond truthfully. How much professional experience do you have with evaluation-driven development, including building and maintaining evaluation datasets and pipelines covering tool selection, trajectory evaluation and LLM-as-judge? choose one
  • Please respond truthfully. How much professional experience do you have with LLM observability, tracing and instrumentation using Langfuse, OpenTelemetry or similar tooling? choose one
  • Please respond truthfully. How much professional backend engineering experience do you have, including TypeScript/Node, Postgres and serverless AWS? choose one
  • Please respond truthfully. What level of professional experience do you have designing and evolving orchestration for production AI agents using frameworks such as LangGraph, LangChain or custom orchestration? choose one
  • Please respond truthfully. What level of professional experience do you have applying context engineering to production LLM and AI agent systems? choose one
  • Please respond truthfully. What is your level of spoken and written English? choose one
  • Are you comfortable working the fixed shift hours of 12 PM – 9:30 PM IST (Summer) and 1 PM – 10:30 PM IST (Winter), Monday to Friday? choose one
  • We are looking for candidates who can start within the next 30 days. ⚠️ Providing false or misleading information will result in disqualification. Candidates progressing to the next stage will be required to upload proof of their notice period or last working day before meeting the client. When is your earliest realistic start date? choose one
  • What is your current salary in LPA? We will cross-check references later in the process. choose one · optional
  • What is your minimum expected salary (LPA, CTC)? choose one · optional
  • Are you open to negotiation? choose one
  • Do you acknowledge that submitting more than one application for the same role, or providing any false or misleading information, will automatically and permanently disqualify you from the recruitment process? choose one

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available