Senior DevOps Engineer (AWS | Kubernetes | Terraform)
Lead AWS cloud infrastructure, Kubernetes, Terraform, CI/CD, and observability for Musafir’s cloud-native travel platform, with ownership of reliability, security, automation, and operational excellence.
Company Details
Musafir is a technology-driven travel platform building scalable digital travel experiences. Website: Musafir
Requirements
8–12 years of experience in Cloud/DevOps engineering
Strong hands-on AWS experience across core cloud services
Strong Kubernetes experience with EKS, ECS, Docker, Helm, ingress, and service mesh
Strong Terraform experience; CloudFormation or CDK experience
Experience with CI/CD using Azure Pipelines or GitHub Actions
Strong scripting skills in Python, Bash, or PowerShell
Experience with OpenTelemetry, Prometheus, Grafana, CloudWatch, and ELK
Strong understanding of AWS security, IAM, KMS, secrets management, and security scanning
Experience with Git and Agile/Scrum
Strong incident management, troubleshooting, RCA, and post-mortem skills
Understanding of high availability, disaster recovery, RTO/RPO, and cost optimization
Responsibilities
Design and manage scalable, secure, and cost-optimized AWS infrastructure
Build and operate Kubernetes workloads using EKS, ECS, Docker, Helm, and related tooling
Develop and maintain Infrastructure as Code using Terraform
Build and manage CI/CD pipelines using Azure Pipelines and GitHub Actions
Manage AWS serverless and event-driven services including Lambda, API Gateway, SQS/SNS, and EventBridge
Implement observability using OpenTelemetry, Prometheus, Grafana, CloudWatch, and ELK
Define and monitor SLIs, SLOs, and SLAs
Drive cloud security, compliance, secrets management, and vulnerability scanning
Design backup and disaster recovery strategies aligned with RTO/RPO requirements
Lead production incident response, root-cause analysis, and post-mortems
Improve infrastructure reliability, automation, performance, and cloud costs
Use AI/Copilot to generate and review Terraform, scripts, and CI/CD pipelines
Apply AIOps for anomaly detection, log summarization, and AI-assisted incident RCA
Support infrastructure for AI workloads including GPU/inference nodes, vector databases, and model-serving platforms
Job Details
Pune — Work from Office
Interview Process
Technical Screening
Technical Interview
System Design Interview
Hiring Manager Round
HR Round
Important Note
ClanX is a recruitment partner, helping Musafir hire Senior DevOps Engineer.
Skills
- Agile
- AI
- Anomaly Detection
- API
- Automation
- AWS
- Azure
- Bash
- CDK
- CI/CD
- Cloud
- Cloud Native
- Cloud Security
- CloudFormation
- CloudWatch
- DevOps
- Docker
- ECS
- EKS
- ELK
- Event Driven Architecture
- Eventbridge
- Git
- GitHub
- GitHub Actions
- Grafana
- Helm
- IAM
- Infrastructure as Code
- Kubernetes
- Lambda
- Observability
- OpenTelemetry
- PowerShell
- Prometheus
- Python
- Scrum
- Secrets Management
- Serverless
- SNS
- SQS
- Technical Screening
- Terraform
- Vector Databases
- Vulnerability Scanning
As published by recruitee · 18 questions · 6 written answers
Basics
Full name, Email, CV, Phone
Short answers (3)
- Please share your current CTC breakup, including Fixed, Variable, ESOPs, and other components, if any. Please stick to this format: Fixed: 15L, Variable: 3L, ESOPs: 5L with 4-year vesting
- What is your expected annual CTC? (INR)
- Help us with your LinkedIn profile link
Pick from a list (9)
- How many years of experience do you have as a DevOps Engineer primarily with AWS?
- Do you have hands-on experience with AWS Cloud?
- Are you willing to relocate to Pune?
- What is your notice period? (We are looking for candidates who can join within 30 days)
- Is your notice period negotiable?
- Have you managed production AWS infrastructure using Kubernetes/EKS and Terraform?
- Do you have experience working at a product-based company?
- Are you a citizen of India and currently based in India?
- What is your current annual CTC range? (INR)
Written answers (6)
- Can you briefly describe the AWS infrastructure you currently manage, including the scale, services, and your ownership?
- Tell us about a production Kubernetes/EKS setup you have worked on. What challenges did you face and how did you solve them?
- Describe a major production incident you handled. How did you troubleshoot it, identify the root cause, and prevent it from happening again?
- What interests you about the company Musafir (https://in.musafir.com/)? Tip: Research the company and share a specific reason. Generic or copy-pasted answers will be rejected.
- Why do you want to leave your current role?
- Where did you find this job opportunity? (e.g., LinkedIn, Twitter, WhatsApp, etc). If someone referred you, please mention their name and contact number so we can thank them. (Write NA if not applicable).

