Senior Site Reliability Engineer - London
Salary: £60,000 - 65,000 per year
Requirements:- Experience in a Site Reliability Engineering, Production Engineering, Cloud Operations, or NOC environment
- Exposure to Linux systems administration
- Exposure to AWS cloud infrastructure
- Exposure to Kubernetes and Docker
- Exposure to production support and incident management
- Exposure to Python, Bash, or Go scripting
- Exposure to monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk, or CloudWatch
- Exposure to networking fundamentals including DNS, TCP/IP, and load balancing
- A passion for automation, continuous improvement, and operational excellence
- Experience with Infrastructure as Code such as Terraform would be beneficial
- Experience with SRE principles such as SLIs and SLOs would be beneficial
- Experience in regulated environments would be beneficial
- Monitor and maintain highly available production platforms running in AWS
- Respond to and manage production incidents across a 24/7 service
- Investigate complex technical issues and restore services quickly and effectively
- Develop automation to reduce manual operational tasks and improve platform resilience
- Build and improve monitoring, alerting, and observability across cloud environments
- Work alongside Software, Platform, Cloud, and Security Engineers to improve reliability and operational excellence
- Contribute to post-incident reviews and drive continuous service improvements
- Support containerised workloads using Kubernetes and Docker
- AI
- AWS
- Bash
- Cloud
- CloudWatch
- Datadog
- Docker
- Grafana
- Support
- Kubernetes
- Linux
- Load Balancing
- Prometheus
- Python
- Security
- Splunk
- TCP/IP
- Terraform
- DevOps
More:
We are a global leader in AI-powered customer experience and cloud technology, and we are expanding our engineering teams following the award of a major government programme. We are building and supporting highly secure, cloud-native platforms that deliver sensitive communication services. This is a fully remote role in the UK on a 24/7 shift pattern with a 28-day rota including days and nights. We offer a competitive salary, bonus, and excellent benefits, and you will join an engineering-led organisation where reliability, automation, and continuous improvement are central to our platform.
last updated 36 week of 2026