Senior Site Reliability Engineer
About the role
We're a fast-moving scale-up and our infrastructure needs to keep pace. That’s why we're looking for a Senior Site Reliability Engineer to own the reliability, automation, and security of our GCP platform, and to set the bar for how we run production as we grow.
You'll spend most of your time building and operating systems, but you'll also mentor more junior engineers and shape the practices the rest of the team works by. If you like having real ownership, moving quickly without breaking things, and turning manual toil into automation, you'll fit right in.
Responsibilities:
- Own the reliability of our production systems on GCP: define SLOs/SLIs, drive down incidents, and lead blameless postmortems.
- Build the observability stack (Prometheus, Grafana, and related tooling) so we catch problems before customers do.
- Design and maintain infrastructure as code with Terraform, no click-ops.
- Operate and scale our Kubernetes / GKE workloads.
- Build and harden CI/CD pipelines so teams can ship safely and often.
- Relentlessly automate manual work; if it's done twice by hand, it's a candidate for automation.
- Bake security into the platform: IAM, secrets management, network policy, and least-privilege by default.
- Help us meet compliance requirements as we mature, and make the secure path the easy path for other engineers.
- Own cloud cost visibility and efficiency: right-sizing, capacity planning, and scaling strategy as the company grows.
Requirements:
- 5+ years in infrastructure, SRE, DevOps, or platform engineering, including production ownership of cloud systems.
- Strong hands-on experience with GCP.
- Deep experience with Terraform.
- Production experience running Kubernetes / GKE.
- Proven track record building and maintaining CI/CD pipelines.
- Solid grasp of observability practices and tooling (Prometheus, Grafana, etc.).
- A bias toward automation and a security-conscious mindset.
- The communication skills to mentor others and influence how the team works.
- Experience scaling infrastructure in a startup or high-growth environment.
- Scripting/programming beyond config (Go, Python, etc.).
- Service mesh, GitOps (e.g., ArgoCD/Flux), or progressive delivery experience.
Recruitment Process:
- Screening call with Doriane (30 min)
- Hiring Manager interview (45 min)
- Technical onsite Interview - System Design (60 min)
- Technical Interview - Problem Solving (60 min)
- Leadership Interview (30 min)
- Fit Interview (30 min)
Skills
As published by lever · 6 questions · 1 written answer
Basics
Resume/CV, Full name, Pronouns, Email, Phone, Current location, Current company, LinkedIn URL, Twitter URL, GitHub URL, Portfolio URL, Other website
Short answers (3)
- Do you currently hold a valid work permit for France?
- Are you currently located in the Paris region or seeking relocation?
- Are you seeking a full-time role?
Pick from a list (2)
- Do you speak Fluent French? optional
- Do you speak Fluent English? optional
Written answers (1)
- What is your expected salary range for this role (annual gross in EUR)?
