Site Reliability Engineer II
You will manage multiple AWS accounts using tools like AWS Control Tower and Terraform, deploy and debug cloud stacks, and educate teams on new cloud projects while ensuring the security of the cloud infrastructure. As an SRE, you'll identify optimal cloud-based solutions for internal users and maintain cloud infrastructure following industry-leading practices and company security policies. This is a 24/7 on-call position with a rotation schedule.
Responsibilities
- Manage multiple AWS accounts with tools like AWS Control Tower and Terraform
- Deploy and debug cloud stacks
- Educate teams on new cloud projects
- Ensure the security of the cloud infrastructure
- Identify optimal cloud-based solutions for internal users
- Maintain cloud infrastructure following industry leading practices and company security policies
- Participate in a 24/7 on-call rotation schedule
Requirements
- 3+ years of experience with Linux/UNIX systems administration
- 3+ years of experience with Amazon Web Services (AWS)
- 2+ years experience in scripting skills in Python / Ruby / Bash / Go
- Experience provisioning public cloud resources using frameworks such as CloudFormation and Terraform
- Solid experience with server configuration with Ansible / Puppet / Chef / Salt
- Experience using Git in a team environment (merge requests, branching, push, and pulls)
- Experience with Docker and container orchestration (Kubernetes) is a plus
- Experience with DevOps automation platforms such as Jenkins and Artifactory is a plus
- 2+ years of experience with AWS Control Tower and Terraform is a plus
- Developed monitoring solutions and analysis across multiple data centers is a plus
Benefits
- Generous PTO & Holiday Schedule
- Parental Leave
- Progressive Healthcare Options
- Retirement Programs
- Opportunity for Education Reimbursement
- Commuter Offset (Specific locations)