Senior DevOps & Site Reliability Engineer (12-months contract)
About the job Senior DevOps & Site Reliability Engineer (12-months contract)
We are looking for a Senior DevOps & Site Reliability Engineer to take a leading role in designing, implementing and continuously improving highly available technology platforms supporting mission-critical business applications.
This is an opportunity to work across cloud, DevOps, platform engineering, automation, observability and SRE , while helping engineering teams deliver software faster, safer and more reliably.
What You’ll Be Doing
Platform Engineering & Automation
- Build and maintain cloud-native platforms and infrastructure.
- Develop IaC and automation using Terraform, Bicep, ARM and Ansible.
- Automate provisioning, deployments and operational processes.
DevOps & CI/CD
- Design and maintain CI/CD pipelines.
- Automate testing, security, deployments and rollbacks.
- Improve release speed, quality and reliability.
Site Reliability Engineering
- Implement SRE practices, SLIs, SLOs and SLAs.
- Improve system availability, resilience, performance and scalability.
- Lead incident response, root-cause analysis and reliability improvements.
Cloud Operations
- Manage and optimize enterprise cloud environments.
- Ensure high availability, security, scalability and cost efficiency.
- Support disaster recovery and hybrid/multi-cloud environments.
Monitoring & Observability
- Implement monitoring, logging, metrics, tracing and alerting.
- Build dashboards and improve MTTD and MTTR.
- Drive proactive issue detection.
Security & Compliance
- Embed DevSecOps and security controls into delivery pipelines.
- Support vulnerability management and regulatory compliance.
- Strengthen overall platform security.
Technical Leadership
- Provide technical leadership across DevOps, cloud and platform teams.
- Mentor engineers and promote engineering best practices.
- Contribute to architecture, technology roadmaps and continuous improvement.
Experience
- 8+ years of experience across software engineering, infrastructure, cloud, DevOps or platform engineering.
- 5+ years of hands-on DevOps engineering experience.
- 3+ years of experience in SRE or production operations environments.
- Proven experience supporting mission-critical production systems.
- Experience operating and supporting large-scale enterprise platforms.
Technical Skills
- Cloud: Azure, AKS, Azure App Services, Networking, Monitor, Storage & Identity; AWS/GCP advantageous.
- DevOps: Azure DevOps, GitHub Enterprise, Git, Jenkins, SonarQube, Artifactory/Nexus.
- IaC & Automation: Terraform, Bicep, ARM Templates, Ansible.
- Containers: Docker, Kubernetes, Helm; OpenShift advantageous.
- Observability: Dynatrace, Grafana, Prometheus, Elastic, Splunk, Azure Monitor, OpenTelemetry.
- Programming & Scripting: Python, PowerShell, Bash, C#, Java; Go advantageous.
- Core Expertise: Cloud & Platform Engineering, SRE, Infrastructure Automation, CI/CD, DevSecOps, Systems Integration, Performance Optimisation and Incident Management.
Advantageous Certifications
- Azure DevOps Engineer Expert / Azure Solutions Architect Expert
- Certified Kubernetes Administrator (CKA) / CKAD
- HashiCorp Terraform Associate
- AWS Certified DevOps Engineer
- ITIL Foundation / SRE Foundation
If you're passionate about DevOps, cloud, automation and reliability engineering and want to make a meaningful impact in a complex enterprise technology environment, we'd love to hear from you.
#J-18808-LjbffrSkills
- AKS
- Ansible
- Artifactory
- Automation
- AWS
- Azure
- Azure DevOps
- Bash
- Bicep
- CI/CD
- Cloud
- Cloud Native
- C#
- DevOps
- DevSecOps
- Docker
- Dynatrace
- GCP
- Git
- GitHub
- Grafana
- Helm
- Infrastructure as Code
- ITIL
- Java
- Jenkins
- Kubernetes
- Networking
- Observability
- OpenShift
- OpenTelemetry
- PowerShell
- Prometheus
- Python
- Regulatory Compliance
- SonarQube
- Splunk
- Terraform