Senior SRE Engineer
Summary
The Senior SRE Engineer will manage AWS cloud infrastructure and Kubernetes environments, focusing on automation, Infrastructure as Code, and building AI workloads. The role involves solution design, platform strategy, and driving engineering excellence using tools like Terraform, Python, and Go.
Are you an experienced SRE who thrives on automation, Infrastructure as Code, and building highly reliable cloud platforms? We're looking for a Senior Site Reliability Engineer to take ownership of the AWS cloud infrastructure and Kubernetes environment
This is a highly hands-on role, with 70-90% of your time focused on Infrastructure as Code, automation, coding and solution design. You'll work closely with technical leadership, contribute to platform strategy, and help drive engineering excellence across the business.
What You'll Bring
This is a highly hands-on role, with 70-90% of your time focused on Infrastructure as Code, automation, coding and solution design. You'll work closely with technical leadership, contribute to platform strategy, and help drive engineering excellence across the business.
What You'll Bring
- Strong Site Reliability Engineering (SRE) background
- Proven Solid experience designing, maintaining and supporting Kubernetes clusters
- Experience building AI workloads and agents on Kubernetes EKS
- Deep experience with Infrastructure as Code (Terraform)
- Security mindset
- Strong automation mindset
- Development capability using Python and/or Go
- Experience designing scalable, secure and resilient cloud solutions
- Strong understanding of observability, reliability, CI/CD and operational excellence
- Ability to provide technical thought leadership
- AWS Cloud experience with agnostic mindset, with experience applying best practices across cloud platforms