freehire launches on Product Hunt on 26 August.

Follow →

Senior AI Software Engineer (Java | Spring Boot | Python)

Summary

Senior AI Software Engineer building production-grade AI/LLM applications connecting enterprise backend systems using Java, Spring Boot, and Python, with a focus on RAG, vector search, and API integrations.

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior AI Software Engineer (Java | Spring Boot | Python) based in Mexico.

This role offers the opportunity to build production-grade AI applications that connect enterprise systems, backend services, and large language models.
You will combine strong backend engineering expertise with hands-on development of modern AI and LLM-powered solutions.
The position covers APIs, integrations, Retrieval-Augmented Generation, prompt-driven functionality, and intelligent enterprise architectures.
You will help transform structured and unstructured data into reliable inputs for scalable AI applications.
A strong focus on quality, security, observability, performance, and production readiness will be central to your work.
You will collaborate closely with Product, Architecture, and Engineering teams to turn complex business requirements into robust technical solutions.
This is an opportunity to work at the intersection of enterprise software engineering and rapidly evolving AI technologies.

Accountabilities

  • Design, develop, maintain, and optimize backend services using Java, Spring Boot, Python, and FastAPI or similar frameworks.
  • Build production-ready REST APIs and integrations connecting AI services with enterprise applications and internal systems.
  • Develop and integrate LLM-powered solutions using platforms and models such as OpenAI, Claude, Gemini, Llama, Azure OpenAI, or comparable technologies.
  • Design and implement Retrieval-Augmented Generation (RAG) solutions, including embeddings, vector search, document retrieval, and knowledge-base integrations.
  • Develop backend services supporting prompt management, context windows, token limits, response validation, large-input handling, and hallucination mitigation.
  • Integrate AI services through REST APIs, JSON, asynchronous processing, and event-driven architectures.
  • Design processes for transforming, normalizing, and preparing structured and unstructured data for AI applications.
  • Implement guardrails, validation mechanisms, error handling, logging, monitoring, and other controls required for reliable AI systems.
  • Evaluate the quality and reliability of model-generated responses through validation criteria, test cases, and structured review processes.
  • Optimize backend and AI services for performance, concurrency, scalability, reliability, and production readiness.
  • Use AI-assisted development tools to accelerate coding, refactoring, testing, debugging, and documentation.
  • Participate in code reviews, automated testing, CI/CD processes, and production deployments.
  • Collaborate with Product, Architecture, and Engineering teams to design scalable solutions aligned with enterprise requirements.
  • Contribute to technical standards and best practices for maintainable, secure, and high-quality software and AI development.
  • Requirements

    • Strong professional experience developing backend applications with Java and Spring Boot.
    • Practical experience with Python, particularly for backend services, AI integrations, automation, or data processing.
    • Proven experience developing and maintaining REST APIs and production-grade backend services.
    • Hands-on experience developing or integrating applications powered by Large Language Models (LLMs).
    • Strong understanding of RAG, embeddings, vector databases, prompt engineering, context windows, token management, chunking, large-input processing, output validation, and hallucination handling.
    • Experience working with one or more LLM platforms or models, such as OpenAI, Claude, Gemini, Llama, or Azure OpenAI.
    • Solid knowledge of API design, JSON, data structures, concurrency, asynchronous processing, and enterprise system integration.
    • Ability to write clean, maintainable, scalable, secure, and production-ready code.
    • Experience using AI-assisted development tools such as Cursor, Claude Code, GitHub Copilot, or similar solutions.
    • Experience with Git, Agile methodologies, CI/CD, automated testing, and software development best practices.
    • Experience with FastAPI or Flask is desirable.
    • Experience with vector technologies such as Pinecone, Weaviate, FAISS, Chroma, Azure AI Search, or similar platforms is a plus.
    • Familiarity with semantic search, document parsing, retrieval tuning, and knowledge-base integration is advantageous.
    • Experience building AI agents, tool-calling capabilities, or LLM orchestration workflows is a plus.
    • Experience with Docker, Azure, or other cloud platforms is desirable.
    • Familiarity with PostgreSQL, Redis, or comparable technologies is beneficial.
    • Experience with testing frameworks such as pytest, JUnit, Playwright, or API testing tools is advantageous.
    • Experience with event-driven architectures, processing queues, and background jobs is a plus.
    • Previous experience in enterprise, financial, regulated, or security-focused environments is desirable.
    • Strong problem-solving, communication, collaboration, and analytical skills.
    • Ability to work effectively with cross-functional teams while taking ownership of technical decisions and deliverables.
    • Benefits

      • Fully remote position within Mexico.
      • Direct employment with an indefinite-term contract.
      • Full-time position.
      • Monday–Friday schedule, 9:00 AM–6:00 PM Mexico time.
      • Compensation paid in Mexican pesos.
      • Opportunity to work on production-grade AI and enterprise software solutions.
      • Hands-on exposure to LLMs, RAG, AI agents, vector databases, and modern AI development tools.
      • Opportunity to contribute to scalable backend architectures and AI-powered business applications.
How Jobgether works:
We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
Why Apply Through Jobgether?
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
#LI-CL1

What this application asks

lever

Resume/CV, Full name, Email, Phone, Current location, Current company

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available