Observability SME/Architect/Consultant
NewBe an early applicantOur key national client is looking for various observability candidates for upcoming projects.
Observability Architect / SME / Consultant6 months+ contracts
Inside IR35
Hybrid working
Role PurposeThe Observability Architect / SME / Consultant will play a key role within the client's Operational Resilience services, helping customers establish, enhance, and optimise observability capabilities across critical business services.The role is responsible for aligning telemetry, monitoring, service health, dashboards, Service Level Objectives (SLOs), and resilience reporting with Operational Resilience outcomes. Working closely with technology, operations, service management, and business stakeholders, the successful candidate will ensure organisations have the visibility, insight, and governance required to proactively manage service performance, reliability, and resilience.This role combines architecture, consulting, and hands-on subject matter expertise, requiring the ability to design strategic solutions while supporting implementation and continuous improvement initiatives.
Essential Skills & ExperienceObservability & Monitoring
- Enterprise Observability Architecture
- Monitoring Strategy & Design
- Telemetry, Metrics, Logs, and Distributed Tracing
- Service Health Monitoring
- Dashboard Design & Reporting
- Alerting & Event Correlation
- AIOps and Operational Analytics
- Performance & Availability Monitoring
Operational Resilience
- Important Business Services (IBS)
- Service Mapping and Dependency Mapping
- Impact Tolerances
- Scenario Testing
- Resilience Risk Assessment
- Operational Resilience Frameworks
- Business Continuity Principles
Service Management & Operations
- Site Reliability Engineering (SRE)
- Incident Management
- Problem Management
- Major Incident Support
- IT Operations
- Service Reliability & Availability Management
- ITIL Frameworks and Best Practice
Consulting & Leadership
- Stakeholder Management
- Communication & Presentation Skills
- Workshop Facilitation
- Business Analysis
- Change Management
- Governance & Reporting
- Problem Solving & Decision Making
- Cross-functional Collaboration
Technical ExpertiseCandidates should demonstrate experience with one or more of the following technologies and platforms:Observability Platforms
- Dynatrace
- Splunk
- Grafana
- Prometheus
- Elastic / ELK Stack
- AppDynamics
- New Relic
Service Management & IT Operations
- ServiceNow
- ServiceNow Event Management
- ServiceNow ITOM
- CMDB and Dependency Mapping Solutions
Cloud Monitoring
- Azure Monitor
- AWS CloudWatch
- Google Cloud Operations Suite
AIOps & Analytics
- Dynatrace Davis AI
- Splunk ITSI
- Moogsoft
- BigPanda
- Other Operational Analytics and AIOps Platforms
QualificationsEssential
- Degree-level qualification or equivalent experience in Information Technology, Computer Science, Engineering, or a related discipline.
- Demonstrable experience in Observability, Monitoring, Service Reliability Engineering, IT Operations, Service Management, Operational Resilience, or Platform Engineering environments.
Desirable
- ITIL Foundation or Advanced ITIL Certifications
- TOGAF
- Dynatrace Certification
- Splunk Certification
- SRE Foundation
- PRINCE2, Agile, Scrum, or PMP
- AWS, Azure, or GCP Certifications
Hays Specialist Recruitment Limited acts as an employment agency for permanent recruitment and employment business for the supply of temporary workers. By applying for this job you accept the T&C's, Privacy Policy and Disclaimers which can be found at hays.co.uk