Site Reliability Engineer (SRE) – DevOps & Cloud Engineering
Summary
Hands-on Site Reliability Engineer in Abu Dhabi running and improving AWS-based production systems — Amazon EKS/Kubernetes, Lambda, CloudWatch observability — using Terraform, GitOps and Python automation, plus incident management and production support. Requires 5+ years' experience; only UAE-based candidates may apply.
Experience: 5+ Years
Availability: UAE-based candidates only
We are looking for a hands-on Site Reliability Engineer (SRE) with strong expertise in AWS, Kubernetes/EKS, Python, DevOps, production support, and observability.
Key Requirements:
- 1. 5+ years of experience in SRE, Production Engineering, DevOps, or Software Engineering
- 2. Strong hands-on AWS experience, especially Amazon EKS / KubernetesStrong Python skills for automation and operational tooling
- 3. Experience with AWS Lambda, IAM, VPC, S3, ECR, SQS/SNS/EventBridge
- 4. Strong knowledge of AWS CloudWatch, X-Ray, CloudTrail and observability
- 5. Experience with Kubernetes, Helm/Kustomize and GitOps
- 6. Hands-on experience with Terraform / CloudFormation / CDK
- 7. Strong understanding of CI/CD, incident management, troubleshooting and production support
- 8. Experience with autoscaling, networking, storage, monitoring and high-availability environments
- 9. Strong Linux, containers, networking and distributed-systems fundamentals
P.S : Only candidates who are currently available in the UAE and meet the 5+ years experience requirement should apply.