Senior SRE
Summary
Senior SRE responsible for the reliability of global Markets operations platforms, using Prometheus, Grafana, and Elasticsearch while supporting on-premises environments and exploring AI-driven automation.
About the Role
Join a newly formed, international SRE team responsible for the reliability and stability of technology platforms supporting global Markets operations. The team works closely with a central observability platform team, offering a unique opportunity to combine classic SRE work with the development of AI-driven solutions.
This role is ideal for someone who enjoys having a real impact on the stability of critical systems, while also wanting to grow into automation and AI-driven operations.
What We're Looking For
Must have:
- Strong SRE fundamentals: incident management, production support, capacity planning.
- Experience diagnosing and fixing production issues in high-availability environments.
- Python (preferred) plus knowledge of Java or shell scripting.
- Hands-on experience with Prometheus, Grafana, Elasticsearch.
- Experience in on-premises environments (public cloud exposure is limited).
- Strong English language skills.
- High level of self-motivation and ability to work autonomously.
Nice to have:
- Familiarity with Azure DevOps (ADO).
- Experience building or working with AI agents / prompt engineering.
- Background in banking or another environment of similar complexity and criticality (fintech, telco, large-scale e-commerce).