Site Reliability Manager
Staples Site Reliability Manager
Requirements
- Bachelor’s degree in Computer Science or related field with continuous and progressive experience
- Minimum of 8 years of related experience working with these technologies:
- Application Performance Management and Monitoring tools such as New Relic, AppDynamics, SiteSpect, and Datadog
- Content Delivery: Akamai
- Infrastructure monitoring tools like Zabbix, and Prometheus
- Databases eg: MongoDB, Oracle, Couchbase, Redis, MySQL
- Frameworks such as Dust/Angular, Nodejs, Springboot
- Log Analytics tools like Splunk, and ELK/Elastic
- Digital experience tools like Fullstory
- Performance Testing tools such as JMeter, Loadrunner, etc.
- Performance tuning experience with Tomcat, Node.js and Spring Boot.
- Strong understanding of non-functional requirements, performance testing processes, and defect tracking.
- 8+ years of experience with Cloud Technologies, at least half of which should be on the Microsoft Azure platform
- Strong hands-on experience with infrastructure and services (systems, network, cloud technology, provisioning, storage, etc)
- Must have strong experience with programming in one or more scripting languages (Python, Azure CLI, or Powershell)
- Hands-on experience with tool sets related to automation, orchestration, and managing infrastructure (Terraform, Puppet, Ansible, or Jenkins)
- Experience with configuring, deploying, and administering infrastructure and application monitoring tools that assist in troubleshooting performance and stability issues in a cloud environment.
- As SRE and EIRE are global operational functions providing 24x7 support, weekend and public holiday coverage is an inherent expectation of these roles.
- Eligible coverage will be offset through compensatory time off, aligned with company policy.


