Point your AI agent at freehire and let it find you a job.

Get the CLI →

Infinity Quest UK

NewBe an early applicant

SRE

Posted Updated
Discussion

Summary

Site Reliability Engineer modernizing IT operations: building observability platforms, driving AIOps and toil automation, and enforcing SRE practices (SLOs, SLIs, error budgets) with automated incident response. Core stack is Dynatrace and Datadog, Python and Ansible, AWS and Azure, with Docker/Kubernetes and microservices.

Primary Responsibilities: Work closely with Product Engineering team and implement strategies for modernizing IT operations enhancing observability and toil reduction. Architect and deploy observability platforms to monitor system health, performance, and reliability effectively. Propose & drive strategies for AI-driven alerting and proactive anomaly detection to reduce MTTD & MTTR. Develop and enforce SRE best practices, including Service Level Objectives (SLOs), Service Level Indicators (SLIs…

Skills

Apply

See also

SRE jobs by country — openings, pay and top skills →

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available