Senior DevOps Operation engineer
Summary
Senior DevOps engineer responsible for deploying and maintaining Kubernetes clusters, writing automation scripts in Python/Java/Go, managing databases, and building observability systems while optimizing CI/CD workflows.
Join Tata Consultancy Services, Asia Pacific and be part of an organization committed to sustainable development for our future. TCS follows the Tata group philosophy of building sustainable businesses that are rooted in the community and demonstrate care for the environment. Our unique values position us to combine a purpose-driven worldview with digital innovation, collaborating with customers, communities and governments to lead and shape innovative solutions for a sustainable future. TCS has been carbon neutral in its operations across 11 countries, 12 delivery centres and 18 offices in Asia Pacific since 2022. This is only the initial stage in TCS’ journey as we strive to achieve long-term net zero emissions by 2030.
Corporate sustainability is embedded in our triple-bottom-line, focusing on people, the planet, and our purpose. Our offices are designed with eco-friendly features that significantly reduce our carbon footprint and enhance energy efficiency. We actively champion green initiatives, such as promoting paperless operations, implementing energy-efficient practices, and fostering employee engagement in sustainability efforts. When you become part of the TCS family, you will play an essential role dedicated to innovation, excellence, and crafting a brighter, greener future together. Join us and be a part of our mission to drive sustainability through technology and talent at Tata Consultancy Services, APAC today.
Location: Petaling Jaya, Malaysia
Key Responsibilities
- Deploy, operate, monitor and troubleshoot Kubernetes (K8s) clusters and containerized workloads to ensure stable production environments.
- Develop automation scripts and internal tools using Python, Java or Go to reduce manual workload.
- Manage and maintain databases, including routine maintenance, performance tuning, backup and recovery, and fault resolution.
- Build and maintain observability systems, including log collection, metrics monitoring and performance tracking.
- Collaborate with development teams to optimize CI/CD workflows and improve delivery efficiency.
- Perform daily system checks, incident handling, root cause analysis, and implement optimization plans.