DevOps Engineer OpenStack - EPAM SYSTEMS PTE. LTD.
Summary
DevOps engineer at EPAM Systems in Singapore who designs, automates and operates Red Hat OpenStack cloud infrastructure, integrating it with Kubernetes/OpenShift and building AI-assisted operational workflows. Core stack: OpenStack, Kubernetes/OpenShift, RHEL, Python/Ansible, Terraform, Prometheus and Elastic Stack.
Why EPAM?
- By choosing EPAM, you're getting a job at one of the most loved workplaces according to Newsweek 2022 & 2023 .
- Employee ideas are the main driver of our business. We have a very supportive environment where your voice matters.
- You will be challenged while working side-by-side with the best talent globally. We work with top-notch technologies, constantly seeking new industry trends and best practices.
- We offer a transparent career path and an individual roadmap to engineer your future & accelerate your journey.
- At EPAM, you can find vast opportunities for self-development: online courses and libraries, mentoring programs, partial grants of certification, and experience exchange with colleagues around the world. You will learn, contribute, and grow with us.
What You'll Do
- Design, implement and maintain Red Hat OpenStack cloud infrastructure
- Automate deployment, configuration and day-two operations using scripting and automation tools
- Integrate OpenStack with Kubernetes and OpenShift to support container platforms
- Troubleshoot platform and infrastructure issues, lead root cause analysis and drive preventative fixes
- Improve availability, security and performance through monitoring, hardening and capacity planning
- Plan and deliver platform changes including upgrades, hotfixes and operational maintenance
- Build AI-assisted operational workflows for monitoring, incident analysis and remediation using OpenStack APIs
- Create and maintain runbooks, playbooks and system documentation and mentor teammates on procedures
What Will Make You Shine
- Expertise in OpenStack administration and operations in production environments
- Expertise with Kubernetes and container platform integrations
- Strong background in Red Hat Enterprise Linux (RHEL) administration and troubleshooting
- Hands-on experience with automation using Python, Ansible or similar tooling
- Experience with infrastructure as code using Terraform or equivalent tools
- Knowledge of monitoring and observability tools such as Prometheus and Elastic Stack
- Familiarity with virtualization platforms such as VMware ESXi
- Ability to communicate clearly with technical and non-technical stakeholders
Nice to Have
- Red Hat or OpenStack certifications
- Familiarity with Generative AI, large language models (LLMs), agentic AI and AIOps concepts in infrastructure operations
- Experience with continuous integration and continuous delivery (CI/CD) for platform lifecycle management