SRE Technical Lead
Summary
Leads a team to design and maintain resilient cloud infrastructure using Kubernetes/OpenShift, SRE practices, and observability tools like Prometheus and Grafana.
Salary: £70,000 - 90,000 per year
Requirements:- Deep expertise in Kubernetes and/or OpenShift
- Experience working in multi-cloud or hybrid cloud environments
- Strong understanding of SRE principles, including SLAs, SLOs, error budgets, and reliability engineering
- Hands-on experience with observability tooling such as Prometheus, Grafana, OpenTelemetry, Loki, and Tempo
- Strong knowledge of Infrastructure as Code and GitOps tools such as Helm, Kustomize, ArgoCD, and Tekton
- Experience with CI/CD pipelines and automation
- Proven ability to operate as a technical leader in complex, multi-team environments
- Must be eligible for SC Clearance
- Define and implement our SRE strategy, standards, and best practices, including SLAs, SLOs, and error budgets
- Embed reliability principles into our platform and service design from the outset
- Lead key SRE practices such as reliability reviews, operational readiness, and toil reduction
- Drive automation across our monitoring, incident response, and remediation processes
- Act as our technical escalation point for major incidents and high-risk releases
- Lead blameless post-incident reviews and ensure continuous improvement
- Establish observability and capacity management practices using modern tooling
- Identify and eliminate systemic reliability risks and operational inefficiencies
- Collaborate with our engineering, platform, security, and operations teams across multiple vendors
- Provide coaching and mentorship to engineers, raising SRE capability across our organisation
- ArgoCD
- CI/CD
- Cloud
- GitOps
- Grafana
- Helm
- Kubernetes
- Kustomize
- OpenTelemetry
- OpenShift
- Prometheus
- Security
- DevOps
- Embedded
More:
We are hiring an experienced SRE Technical Lead for a senior, client-facing leadership role based in Reading on a hybrid UK-based working pattern, combining home, office, and client site work. We operate across complex, large-scale platforms and are looking for someone who will act as the technical authority for Site Reliability Engineering, driving reliability, availability, and operational excellence across multi-team and multi-vendor environments. This role combines hands-on engineering with strategic leadership, and requires eligibility for SC Clearance.
last updated 32 week of 2026