freehire launches on Product Hunt on 26 August.

Follow →

Senior Backend Software Engineer (Infrastructure MLOps)

Summary

Senior backend engineer builds and scales AWS/Kubernetes infrastructure for AI-powered media platforms, operationalizing ML workloads like recommendation systems and LLM assistants.

Job Description

Insight Global is seeking a Senior Backend Software Engineer (Infrastructure MLOps) for a leading streaming and media technology client. This engineer will serve as a critical bridge between the Infrastructure and Machine Learning organizations, helping operationalize and scale AI-powered applications that support customer-facing recommendation systems, internal AI assistants, and advanced classification platforms. The ideal candidate brings a strong software engineering foundation combined with deep expertise in AWS, Kubernetes, Infrastructure-as-Code, and production systems. This role offers the opportunity to own highly visible infrastructure initiatives, partner closely with ML engineers, influence MLOps strategy, and help build reliable, secure, and scalable AI platforms within a lean, high-impact engineering organization.

Day-to-Day:

  • Build and maintain infrastructure supporting AWS and Kubernetes environments

  • Partner with ML engineers to operationalize and scale machine learning workloads

  • Develop and improve MLOps processes, tooling, and deployment practices

  • Manage infrastructure-as-code and automation initiatives

  • Support CI/CD pipelines and developer tooling platforms

  • Monitor system health, reliability, and production performance

  • Troubleshoot networking, infrastructure, and application issues

  • Improve observability through monitoring, logging, alerting, and dashboards

  • Design secure and scalable solutions for AI/LLM workloads

  • Participate in load testing, change management, incident response, and on-call activities

  • Support recommendation systems, internal AI tools, and ad-classification platforms in production

  • Drive reliability, scalability, and operational excellence across critical services

We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to HR@insightglobal.com.To learn more about how we collect, keep, and process your private information, please review Insight Global's Workforce Privacy Policy: https://insightglobal.com/workforce-privacy-policy/.

Skills and Requirements

  • 8+ years of experience in infrastructure / backend / platform engineering

    - Strong AWS experience

    - Strong Kubernetes administation and production support expeience

    - Infrastructre as Code experience (Terraform preferred)

    - Linux adminstration proficiency

    - Experience supporting ML/LLM deployments in production

    - Deep understanding of networking, debugging, and production systems

    - Experience with observability, monitoring, logging, and alerting tools

  • Strong CI/CD knowledge - Experience supporting recommendation engines or ML-driven customer-facing products

  • Experience securing AI/LLM-powered applications and internal tools

  • Background supporting inference workloads at scale

  • Experience with model deployment, orchestration, and ML infrastructure automation

  • Experience supporting internal AI assistants or developer productivity tools

  • Video streaming or media platform infrastructure experience