DevOps Engineer

Summary

The DevOps Engineer will deploy and maintain the DataRobot platform within customer-managed Kubernetes environments using infrastructure-as-code tools like Terraform and Helm. This role involves troubleshooting complex networking and platform issues while acting as a technical subject matter expert for customer deployments.

Join Our Team at Lean Solutions Group (LSG)! Lean Solutions Group (LSG) is a next-generation solutions provider combining AI-driven automation, industry expertise, and tech-powered talent. Built in the demanding Supply Chain sector, our model now supports 600+ clients across multiple industries, powered by 10,000+ employees in five countries. We help businesses achieve immediate efficiency, long-term resilience, and scalable growth by integrating intelligent technology, optimized processes, and high-performance teams. At LSG, we believe in your talent and your potential. Join a multicultural, people-first environment where you can grow, sharpen your skills, and unlock new career opportunities. Here, every day brings fresh challenges, collaboration, and purpose.
Our Mission: Transform business challenges into lasting success through purpose-built teams, technology, and expertise.
Our Vision: A world where people, empowered by technology, turn any challenge into a catalyst for growth.

What You’ll Do (Key Responsibilities)

  • Customer Deployments: Deploy the DataRobot platform into customer-managed, CNCF-compliant Kubernetes environments (AWS, Azure, Google Cloud, hybrid, or true on-premises).

  • Infrastructure Automation: Write and maintain automated deployment code using Terraform, Helm, and Ansible.

  • Technical Problem Solving: Partner directly with customer teams to troubleshoot complex Kubernetes, networking, and platform integration issues.

  • Escalation Support: Act as the Level 3 subject matter expert to resolve escalated technical cases from internal support teams.

  • Product & Engineering Feedback Loop: Work closely with product and engineering teams to feed deployment insights back into the core platform.

  • Knowledge Sharing: Draft technical documentation and troubleshooting guides for customers and internal team members.

What We’re Looking For

  • Experience: 3+ years in a customer-facing DevOps, Site Reliability, or Systems Engineering role.

  • Kubernetes Expertise: Hands-on mastery of CNCF-compliant Kubernetes deployment and log troubleshooting.

  • Cloud & Networking: Deep knowledge of AWS, Azure, or GCP, plus strong Linux administration and enterprise networking basics.

  • IaC & Automation: Advanced skills with Terraform, Helm, and Ansible.

  • Scripting: Solid Python scripting abilities for debugging and automation tasks.

  • Communication: Ability to explain complex infrastructure concepts clearly to both engineers and non-technical stakeholders.

Nice to Have

  • Certifications: Certified Kubernetes Administrator (CKA) (Required within the first 6 months if not currently held).

  • Domain Knowledge: Prior experience with AI, Machine Learning, or Big Data platforms.

  • Databases: Basic operational experience with MongoDB, PostgreSQL, and Redis.

See also

DevOps jobs by country — openings, pay and top skills →

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available