DevOps Engineer
Summary
Build and maintain Kubernetes clusters, CI/CD pipelines, and GitOps workflows using Terraform, Ansible, and FluxCD to deploy and scale containerized applications.
DevOps & Onboarding
Role Overview:
1. Kubernetes & Container Orchestration
Deploy and manage containerized applications on a container orchestration platform
Independently triage and troubleshoot workload issues (e.g., pods not starting, crash loops)
Understand and work with: Workloads (stateless and stateful), Service exposure (internal/external), Configuration and secrets.
Diagnose and resolve: Scheduling issues & Resource constraints (CPU/memory) & Networking issues between services
Manage application scaling and resource tuning
Understand cluster architecture and core components
2. CI/CD & Automation
Design, build, and maintain CI/CD pipelines
Understand full pipeline lifecycle:
Build → Test → Scan → Release → Deploy
Integrate pipelines with: Source control systems, Artifact repositories & deployment platforms
Automate Application builds, Testing processes & Deployment workflows
Troubleshoot pipeline failures independently
Implement pipeline optimization (speed, reliability)
Support secure pipeline practices (e.g., secrets handling, access control)
3. GitOps
Understand and apply GitOps principles (Git as source of truth)
Manage deployments using declarative configuration stored in version control
FluxCD
Structure repositories for:
Multi-environment deployments
Multi-tenancy setups
4. Infrastructure as Code (IaC) & ConfigurationManagement
Write and maintain Terraform
Provision and manage infrastructure using automation tools
Manage services configuration through Ansible
Implement Reusable modules/templates & Parameterization for environments
Maintain version-controlled infrastructure
Automate system configuration and setup
Troubleshoot provisioning and configuration failures
5. Incident Management & Troubleshooting
Perform independent incident triage and resolution based on defined procedures and SLAs
Follow structured troubleshooting approach:
Identify → Analyze → Plan of action - Mitigate → Resolve
Diagnose issues across Services & K8s Use logs, metrics, and system behavior to identify root cause Perform service restoration within defined SLAs Conduct post-incident analysis (RCA) Document: Incident findings & Preventive measure Escalate issues appropriately when required 6. Service Request Handling Fulfill service requests based on defined procedures Handle requests such as Onboarding & Bringing in of packages Validate and review request requirements before execution Maintain proper documentation and tracking Communicate status and updates clearly to stakeholders Identify opportunities to automate repetitive requests
Interested candidates, please send your CVs on (HIDDEN TEXT) Regret to inform that only shortlisted candidates will be notified.
CEI: R1988671 EA License: 14C7275
Diagnose issues across Services & K8s Use logs, metrics, and system behavior to identify root cause Perform service restoration within defined SLAs Conduct post-incident analysis (RCA) Document: Incident findings & Preventive measure Escalate issues appropriately when required 6. Service Request Handling Fulfill service requests based on defined procedures Handle requests such as Onboarding & Bringing in of packages Validate and review request requirements before execution Maintain proper documentation and tracking Communicate status and updates clearly to stakeholders Identify opportunities to automate repetitive requests
Interested candidates, please send your CVs on (HIDDEN TEXT) Regret to inform that only shortlisted candidates will be notified.
CEI: R1988671 EA License: 14C7275