Software Engineer (DevOps) - Site Reliability
Summary
Build and scale Revolut’s multi-cloud and on-premise infrastructure using IaC and CI/CD, ensuring reliability, security, and global failover for a high-growth fintech platform.
About The Role
Our Technology team builds the systems and experiences that keep Revolut moving. We’re looking for a DevOps Engineer to join the Site Reliability team and build scalable infrastructure from scratch, delivering IaC and CI/CD solutions across multi‑cloud and on‑premise environments.
What You’ll Be Doing
- Building infrastructure and processes to enable entry into new markets and scale 10x
- Delivering easy‑to‑use solutions for provisioning (IaC) and CI/CD across multi‑cloud and on‑premise environments
- Architecting multi‑cloud regional failover, disaster recovery planning, and self‑service BCP to ensure mission‑critical reliability
- Designing and implementing bespoke third‑party connectivity from scratch to handle a high volume of diverse and complex integrations
- Creating and maintaining security‑related infrastructure, automating patching, ensuring audit readiness, and supporting regulatory reporting
- Implementing robust managed services, including Postgres and Redis, across cloud and hybrid footprints
What You’ll Need
- Hands‑on software engineering skills with experience as a backend developer
- Experience with Linux, Docker, and cloud providers (GCP and AWS)
- Experience with distributed systems, including scaling, fault‑tolerance, load‑balancing, networking, and security
- Knowledge of continuous delivery systems such as Jenkins and TeamCity
- Familiarity with configuration management systems and deployment tools such as Ansible and Terraform
- Familiarity with relational databases such as Postgres and MySQL
- Experience with continuous delivery using Kubernetes
- The ability to work independently across internal teams and third parties to design and deliver scalable, end‑to‑end solutions from scratch
Nice to have
- Experience with monitoring solutions such as NewRelic, StackDriver, and Prometheus
- Expertise in on‑premise and hybrid environments, leveraging tools like KubeVirt to orchestrate virtual machines alongside containerised workloads