Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Associate Platform Engineer (SRE) monitors and automates cloud infrastructure on AWS, responds to incidents, and learns reliability engineering with mentorship.
Lead Site Reliability Engineer on Mastercard's Business Operations team in Ashbourne, keeping the payments platform stable, scalable, and production-ready. Day to day: building observability and automation (Python/Go/Bash, Linux, AWS/Azure/GCP), leading incident triage and blameless post-mortems, and mentoring junior engineers.
Lead overnight incident response for a global dental-lab SaaS platform, owning triage, fixes, and seamless shift handoffs while training a backup support engineer.
First senior SRE hire at Omnicell, tasked with building the Site Reliability Engineering practice from the ground up as the company shifts from on-premise, hardware-centric products to a cloud-native SaaS platform that hospitals depend on 24/7. Day to day: design reliability standards, run operations hands-on, and grow the team.
Site Reliability Engineer at a consumer lending company providing 24x7 production support for Java-based, cloud-native applications on AWS (EKS). Day-to-day work centers on monitoring with Datadog and Splunk, incident response and RCA, CI/CD deployments, Kubernetes troubleshooting, and building automation in Python/Bash.
Senior Site Reliability Engineer maintains production Kubernetes clusters, builds observability with Prometheus/Grafana/ELK, and automates monitoring and incident response.
Site Reliability Engineer for Bithumb's blockchain services: defining SLI/SLO policies and error budgets, running on-call, escalation and postmortem processes, improving Datadog observability, and designing highly available blockchain node and RPC infrastructure on AWS and Kubernetes.
Senior Site Reliability Engineer in Bristol keeping critical services available through monitoring and alerting. Day-to-day work spans AWS, Prometheus and Grafana, plus Infrastructure as Code, CI/CD pipelines and automation; an active UK Government Security Clearance (or willingness to be sponsored for one) is required.
Executes and governs end-to-end production and non-production releases at DBS Bank, reviewing change requests, performing deployments (app, middleware, database, privileged access via CyberArk/PAM), and providing L1 incident support in a 24x7 regulated banking environment. Core stack includes Linux, middleware, Oracle/PostgreSQL, and CI/CD/OpenShift tooling with ITIL-based change governance.
Senior Lead SRE at Lumen (CenturyLink) owning production support, incident management, and performance optimization for its customer-facing portal (Lumen Connect) on AWS. Day-to-day spans observability (Datadog, CloudWatch), Terraform IaC, automation, and AI-assisted engineering. Fully remote within the US.
Voleon, an AI/ML-driven asset manager, is hiring a Site Reliability Engineer in London to improve, manage, and monitor production-critical infrastructure and data pipelines for its trading systems. Day-to-day involves debugging Python code, leading deployments, automating workflows, and sharing an on-call rotation, with a stack including Linux, SQL, Prometheus, Grafana, Kubernetes, and Airflow.
Hands-on engineering manager leading Wayve's Japan Fleet Reliability SRE team in Yokohama, accountable for the reliability, safety and availability of the autonomous vehicle fleet. The role blends people leadership with incident management, SLI/SLO-driven reliability engineering, release safety, and building automation and tooling for vehicle systems.
Hands-on engineering manager leading Wayve's Japan Fleet Reliability SRE team in Yokohama, keeping the autonomous vehicle fleet safe, available and reliable. The role blends people leadership with incident command and reliability engineering — SLIs/SLOs, observability, automation and fleet tooling — across vehicle software, cloud infrastructure and fleet operations.
Senior Site Reliability Engineer builds and maintains cloud infrastructure for an AI-driven enterprise revenue platform, focusing on availability, resilience, and operational maturity.
Senior Site Reliability Engineer at Worldline designing and maintaining high-performance, resilient card issuing systems handling 1.16B monthly transactions using Java, Python, and cloud platforms.
Maintains and improves the reliability, performance, and security of Worldline's payment infrastructure platforms, automating deployments and monitoring systems while ensuring compliance with SLAs and regulations.
Maintains and automates infrastructure platforms to ensure high service availability, deploys changes safely, and troubleshoots incidents for a payment processing company.
BT Group is the UK’s leading communications group and the holding company behind some of the country’s most recognised brands – including BT, EE, Openreach and Plusnet. Our purpose is as simple as it is ambitious: we…
Maintains and improves the reliability, scalability, and performance of Worldline's payment platforms by managing monitoring, automation, incident response, and cloud infrastructure in a hybrid work environment.
Maintains and improves the reliability, security, and performance of Worldline's ATM France payment platform through monitoring, automation, and incident response.
We couldn't check your fit for this role — add a CV to your profile to see it next time.