freehire launches on Product Hunt on 26 August.

Follow →

Site Reliability Engineer - Selby Jennings

Summary

Maintains and scales the cloud and container platforms that underpin a top-tier quant hedge fund’s research and trading systems, ensuring high availability and rapid incident resolution.



Site Reliability Engineer

Our client is a world-renowned quantitative hedge fund where technology sits at the core of the business. They are seeking a Site Reliability Engineer to join a highly visible platform team responsible for the reliability, performance, and scalability of the infrastructure that powers cutting-edge research and trading. This is a unique opportunity for a Site Reliability Engineer to work directly alongside traders, quantitative researchers, and engineering teams in an environment where your work has immediate business impact.

As a Site Reliability Engineer, you'll gain exceptional ownership from day one, helping to support and evolve a sophisticated technology estate built on modern cloud and containerisation technologies. The role combines troubleshooting, automation, platform engineering, and stakeholder interaction, making it ideal for someone looking to accelerate their career within a top 1% investment firm. The fund offers top-of-market compensation, strong bonus potential, and annual compensation increases, alongside the opportunity to work with some of the industry's brightest engineers and researchers.



Key Responsibilities

  • Ensure the reliability, availability, and performance of critical research and trading platforms
  • Act as the first point of contact for infrastructure and platform-related issues
  • Investigate and resolve production incidents, minimising business impact
  • Partner closely with traders, quants, and engineers to support business-critical systems
  • Drive improvements to monitoring, automation, operational processes, and platform tooling
  • Manage onboarding, permissions, access requests, and platform support activities


Key Skills & Experience

  • 3-8 years' experience in Site Reliability Engineering, Platform Engineering, DevOps, Infrastructure Engineering, or Production Engineering
  • Strong experience with AWS and Kubernetes in production environments
  • Excellent Linux administration and troubleshooting skills
  • Experience with Docker, CI/CD pipelines, infrastructure automation, and cloud-native technologies
  • Knowledge of observability and monitoring tools such as Datadog, Prometheus, Grafana, CloudWatch, or ELK
  • Scripting experience with Python, Bash, or similar languages
  • Strong communication skills, a proactive mindset, and a genuine sense of ownership
  • Experience within financial services, trading, or other high-performance environments is advantageous


See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available