SRE
Summary
Site Reliability Engineer modernizing IT operations: building observability platforms, driving AIOps and toil automation, and enforcing SRE practices (SLOs, SLIs, error budgets) with automated incident response. Core stack is Dynatrace and Datadog, Python and Ansible, AWS and Azure, with Docker/Kubernetes and microservices.