Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Lead overnight incident response for a global dental-lab SaaS platform, owning triage, fixes, and seamless shift handoffs while training a backup support engineer.
First senior SRE hire at Omnicell, tasked with building the Site Reliability Engineering practice from the ground up as the company shifts from on-premise, hardware-centric products to a cloud-native SaaS platform that hospitals depend on 24/7. Day to day: design reliability standards, run operations hands-on, and grow the team.
Site Reliability Engineer at a consumer lending company providing 24x7 production support for Java-based, cloud-native applications on AWS (EKS). Day-to-day work centers on monitoring with Datadog and Splunk, incident response and RCA, CI/CD deployments, Kubernetes troubleshooting, and building automation in Python/Bash.
Senior Site Reliability Engineer maintains production Kubernetes clusters, builds observability with Prometheus/Grafana/ELK, and automates monitoring and incident response.
Site Reliability Engineer for Bithumb's blockchain services: defining SLI/SLO policies and error budgets, running on-call, escalation and postmortem processes, improving Datadog observability, and designing highly available blockchain node and RPC infrastructure on AWS and Kubernetes.
Senior Site Reliability Engineer in Bristol keeping critical services available through monitoring and alerting. Day-to-day work spans AWS, Prometheus and Grafana, plus Infrastructure as Code, CI/CD pipelines and automation; an active UK Government Security Clearance (or willingness to be sponsored for one) is required.
Executes and governs end-to-end production and non-production releases at DBS Bank, reviewing change requests, performing deployments (app, middleware, database, privileged access via CyberArk/PAM), and providing L1 incident support in a 24x7 regulated banking environment. Core stack includes Linux, middleware, Oracle/PostgreSQL, and CI/CD/OpenShift tooling with ITIL-based change governance.
Voleon, an AI/ML-driven asset manager, is hiring a Site Reliability Engineer in London to improve, manage, and monitor production-critical infrastructure and data pipelines for its trading systems. Day-to-day involves debugging Python code, leading deployments, automating workflows, and sharing an on-call rotation, with a stack including Linux, SQL, Prometheus, Grafana, Kubernetes, and Airflow.
Hands-on engineering manager leading Wayve's Japan Fleet Reliability SRE team in Yokohama, accountable for the reliability, safety and availability of the autonomous vehicle fleet. The role blends people leadership with incident management, SLI/SLO-driven reliability engineering, release safety, and building automation and tooling for vehicle systems.
Hands-on engineering manager leading Wayve's Japan Fleet Reliability SRE team in Yokohama, keeping the autonomous vehicle fleet safe, available and reliable. The role blends people leadership with incident command and reliability engineering — SLIs/SLOs, observability, automation and fleet tooling — across vehicle software, cloud infrastructure and fleet operations.
Senior Site Reliability Engineer builds and maintains cloud infrastructure for an AI-driven enterprise revenue platform, focusing on availability, resilience, and operational maturity.
Senior Site Reliability Engineer at Worldline designing and maintaining high-performance, resilient card issuing systems handling 1.16B monthly transactions using Java, Python, and cloud platforms.
Maintains and improves the reliability, performance, and security of Worldline's payment infrastructure platforms, automating deployments and monitoring systems while ensuring compliance with SLAs and regulations.
Maintains and automates infrastructure platforms to ensure high service availability, deploys changes safely, and troubleshoots incidents for a payment processing company.
Maintains and improves the reliability, scalability, and performance of Worldline's payment platforms by managing monitoring, automation, incident response, and cloud infrastructure in a hybrid work environment.
Maintains and improves the reliability, security, and performance of Worldline's ATM France payment platform through monitoring, automation, and incident response.
Maintain and improve fraud-processing applications for European financial clients, using Java, Linux, cloud tools (Puppet, Terraform), and monitoring stacks (Zabbix, Grafana) while on-call 24/7.
Maintain and improve the reliability of global payment systems and customer-facing portals, troubleshooting incidents and ensuring 24/7 availability.
The Site Reliability Engineer ensures 24/7 availability of fintech payment systems by managing incidents, monitoring production environments, and optimizing ITIL processes. Core tech: databases (Oracle/PostgreSQL), monitoring tools (Grafana/Dynatrace), scripting, cloud platforms, and SEPA payment infrastructure.
The Site Reliability Engineer will manage production incidents, monitor applications, and coordinate service requests within the SEPA and Instant Payment processing chains. The role requires expertise in SQL, Linux, monitoring tools like Dynatrace and Grafana, and experience with ITIL processes and CI/CD pipelines.
We couldn't check your fit for this role — add a CV to your profile to see it next time.