freehire launches on Product Hunt on 26 August.

Follow →

AI / ML Engineer

Summary

Build and scale multi-agent LLM systems and pipelines to power conversational user experiences across devices, using AWS Bedrock, vector search, and Kubernetes.

Overview

US Mobile is a venture‑backed company on a mission to revolutionize connectivity with a software platform for 5G and IoT. We are expanding rapidly and seek an AI/ML Engineer to develop and scale machine learning models that power user experiences across devices.

Key Responsibilities

  • Design & Deploy Conversational / Multi‑Agent LLM Solutions
    • Craft multi‑agent conversational flows capable of handling a wide range of user requests—both purely informational and action‑oriented.
    • Employ advanced LLM techniques (prompt engineering, context retrieval, multi‑step reasoning) to ensure robust, context‑aware dialogues.
  • Multi‑Modal & Multi‑Model Integration
    • Explore different input/output formats (e.g., text, voice, image) to enrich user interactions.
    • Evaluate models based on technical capabilities and cost efficiency.
  • Platform & Pipeline Building
    • Design data pipelines that feed models with real‑time or near real‑time data.
    • Implement best practices around model lifecycle management—versioning, containerisation, deployment orchestration, etc.
  • Optimization & Scale
    • Ensure the chat system can handle thousands, eventually millions, of concurrent interactions with low latency and high availability.
    • Monitor performance, define metrics (latency, user success rate, fallback rate) and iteratively improve.
  • Ongoing Innovation & Experimentation
    • Stay current on AI/ML advancements, especially generative models, multi‑agent orchestration and knowledge retrieval.
    • Propose new ways to extend AI across the platform—advanced personalization, proactive customer engagements, etc.

Qualifications

  • Core AI/ML Expertise
    • 3+ years of hands‑on experience building and deploying machine learning solutions at scale.
    • Solid understanding of NLP techniques, including transformer models and embeddings, and experience with Hugging Face, AWS Bedrock, and OpenAI’s API.
    • Experience with vector search solutions (e.g., Pinecone, Weaviate, Elasticsearch with vector plugins).
    • Experience building or deploying large language models in the AWS Bedrock ecosystem.
    • Familiarity with multi‑agent LLM frameworks or orchestrations.
  • Backend & Data Infrastructure
    • Proficient in Python or a similar language for data pipelines and model development.
    • Experience with cloud platforms (AWS preferred), containerisation (Docker, Kubernetes) and microservices.
  • Research & Problem‑Solving Mindset
    • Up‑to‑date on AI/ML trends—especially multi‑agent systems, generative modeling, or multi‑modal approaches.
    • Skilled at diagnosing bottlenecks, scaling solutions, and balancing innovation against real‑world constraints.
  • Collaboration & Communication
    • Comfortable presenting complex ML concepts to non‑technical stakeholders.
    • Passion for iterative development—able to pivot based on user feedback and product metrics.

Benefits

  • Competitive salary 130k‑220k CAD (based on experience/location)
  • Flexible working hours
  • Supplemental health insurance
  • Professional development stipend
  • 500 CAD in WFH tech set‑up reimbursement

Think you could be a fit? Apply to learn more!

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available