Senior Platform Engineer
Summary
The Senior Platform Engineer will design, secure, and operate scalable cloud infrastructure using AWS, Terraform, and Puppet. The role focuses on automating environments, driving reliability through SLO governance, and integrating AI-assisted workflows into the development lifecycle.
Senior Platform Engineer
Join a team passionate about leveraging cloud engineering practices to deliver superior services. Our dedication to quality and continuous improvement is the cornerstone of our culture. We are doing great things which will revolutionise how our customers interact with us and our products.
What role will you play
As a Senior Platform Engineer, you will be instrumental in shaping the engineering foundations that every development team builds upon. You will own the design, security, and operation of our infrastructure, driving reliability and scalability through automation, robust configuration management, and the strategic use of cloud services.
In this role, you will collaborate with a team of experts to stand up and automate environments end-to-end, build the tooling that makes our processes repeatable, and own the frameworks that ensure every release is trustworthy. From how we deploy and operate services to how we detect and respond to incidents; your work will directly contribute to our commitment to excellence and innovation in our services. You will be the engineer other teams rely on when something foundational needs to just work.
What you offer
Join a team passionate about leveraging cloud engineering practices to deliver superior services. Our dedication to quality and continuous improvement is the cornerstone of our culture. We are doing great things which will revolutionise how our customers interact with us and our products.
What role will you play
As a Senior Platform Engineer, you will be instrumental in shaping the engineering foundations that every development team builds upon. You will own the design, security, and operation of our infrastructure, driving reliability and scalability through automation, robust configuration management, and the strategic use of cloud services.
In this role, you will collaborate with a team of experts to stand up and automate environments end-to-end, build the tooling that makes our processes repeatable, and own the frameworks that ensure every release is trustworthy. From how we deploy and operate services to how we detect and respond to incidents; your work will directly contribute to our commitment to excellence and innovation in our services. You will be the engineer other teams rely on when something foundational needs to just work.
What you offer
- Technical Leadership & Mentorship: Proven ability to drive ambiguous, cross-team technical initiatives end-to-end as a technical lead. You actively coach and mentor other senior engineers to elevate overall capability and architectural mastery, going far beyond standard pull request reviews.
- Production Reliability, SLO Governance & Incident Management: Hands-on ownership of production platform health, ensuring high-availability environments that meet our 99.98% SLO targets. You will lead incident response, establish robust observability/alerting, and actively guide other engineering and product squads on reliability best practices.
- Deep AWS Platform Expertise: Deep, multi-region expertise in AWS at platform scale. You possess broad and deep mastery across a wide array of AWS native services—including serverless compute (Lambda), relational databases (Aurora RDS), enterprise security design (KMS/CMK, IAM), and complex networking (VPC, ALB, Route53)—coupled with the architectural judgment to make design trade-offs that hold up across dozens of highly active customer environments.
- Infrastructure as Code (IaC) Mastery: Track record owning Terraform/Terragrunt at a module-author level—designing, versioning, and maintaining reusable, semver-tagged modules consumed across environments, rather than just writing per-environment configs.
- AI-Assisted Engineering & Agent Development: Hands-on proficiency with AI tooling (e.g., Claude, GitHub Copilot) and integrating AI-assisted workflows into the SDLC. Experience crafting custom system instructions, defining tool-calling schemas, and implementing custom model tool definitions to orchestrate DevOps agents and autonomous, agent-based tools.
- Next-Generation Cloud Networking: Experience designing secure transit routing, with the ability to help deliver on our upcoming network roadmap—including routing egress via AWS Transit Gateways, configuring AWS Network Firewalls, and managing complex VPC routing tables.
- Zero-Risk Security & Vulnerability Posture: Commitment to a zero-risk production environment by implementing robust vulnerability management, pipeline-gated compliance scanning, and endpoint protection.
- Modern CI/CD & Secure Delivery: Multi-platform pipeline expertise (spanning GitHub Actions and Bitbucket Pipelines with OIDC-federated AWS authentication) with a track record of continuously improving pipeline reliability and building automated security gates.
- Advanced Configuration Management: Advanced Puppet expertise, including owning a deep, multi-dimensional Hiera hierarchy and roles/profiles architecture serving per-customer configuration at scale without introducing fleet regressions.
- Systems Administration & Scripting: Strong Linux administration background (performance tuning, security hardening, multi-system diagnosis) combined with solid Python and Bash scripting to build tooling that other engineers rely on.