DevOps / SRE Cloud Engineer Site Reliability Engineer - WatersEdge Solutions
Summary
A DevOps/SRE engineer at WatersEdge Solutions who defines SLOs and error budgets, builds observability stacks (metrics, logging, tracing), automates toil to improve reliability, and runs blameless post-mortems. Core stack includes Azure and Kubernetes, with monitoring tools like Prometheus, Grafana, and Datadog.
DevOps / SRE Cloud Engineer Site Reliability Engineer at WatersEdge Solutions.
Key technologies: Azure, Kubernetes.
Key Responsibilities
- Define and track SLOs, SLIs and error budgets
- Design and implement observability stacks (metrics, logging, tracing)
- Automate toil and improve system reliability through engineering
- Conduct post-mortems and drive blameless incident retrospectives
Requirements
- 3+ years of relevant experience in site reliability engineer
- Proficiency with monitoring tools (Prometheus, Grafana, Datadog)
- Strong programming skills for automation and tooling