DevOps Engineer/Site Reliability Engineer
Summary
DevOps/Site Reliability Engineer improving deployment, testing, automation, and monitoring for a Boston marketing and analytics software platform processing billions of data points daily. Core stack includes AWS, Ruby/Python scripting, Docker, Chef/Puppet/Ansible, and ELK-style logging.
We’re looking for engineers with:
- A track record managing large AWS deployments, including deep knowledge of AWS security and networking best practices
- The ability to script AWS operations using Ruby or Python
- Knowledge of Docker, experience managing Docker-based deployments a plus
- Experience with infrastructure automation tools such as Chef, Puppet, or Ansible
- Strong background in Unix and networking
- Experience with metrics and log aggregation systems (ELK, collectd and similar)
- Experience with SQL, Hadoop a plus
As a DevOps/Site Reliability Engineer you’ll help us:
- Make it easy to deploy new code and difficult to make mistakes in production
- Monitor, measure and optimize our systems and processes
- Automate everything
- Build or implement resource sharing frameworks like Mesos