Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Build and maintain the reliable, scalable infrastructure that powers Fivetran’s data-pipeline platform using Kubernetes, cloud providers, and Go/Python, while owning incident response and reliability improvements.
JOB DESCRIPTION SUMMARY The Site Reliability & Engineering Manager is the senior technical leader for the Minong facility, responsible for maintenance, reliability, facilities, utilities, and site engineering…
Maintains AI datacenter hardware by analyzing firmware, diagnosing failures, managing vendor RMAs, and building monitoring tools to keep systems reliable.
The Senior Site Reliability Engineer will lead the design, architecture, and observability of Seekr's AI platform to ensure system reliability and performance. The role involves collaborating with engineering teams to manage hybrid cloud infrastructure, automate operational tasks, and handle incident response.
A four-year degree apprenticeship in Newcastle upon Tyne combining study at the University of Warwick with hands-on site reliability work: supporting critical business applications in a financial services environment using Kubernetes, Windows/Linux servers, observability tools, and Python/PowerShell automation, plus incident triage and release management.
Leads the multi-site Collaboration SRE organization for Google Drive and Docs across Sunnyvale and Zurich, owning end-to-end availability, architectural resilience, and operational excellence of Google's collaboration platform. Day to day: managing engineers, running follow-the-sun on-call, and building automation for large-scale distributed systems.
The Senior Systems Reliability Engineer will build and operate Cloudflare's global Edge platform, focusing on automation, scalability, and operational excellence. The role involves using Go or Python to develop tools that improve service availability and performance across a large-scale distributed network.
SRE engineer at БЮРО 1440 keeps company services reliable and predictable: building observability tooling, designing and rolling out monitoring metrics, investigating incidents, and raising engineering standards across teams. Core stack: Linux, Go/Python/Java, Docker/Kubernetes, Prometheus/Grafana, ELK/Splunk/Graylog, CI/CD and IaC.
Senior Site Reliability Engineer at PrizePicks, a fast-growing daily fantasy sports platform. You’ll design, monitor, and scale cloud infrastructure, lead incident response, and mentor engineers using AWS, Kubernetes, Terraform, and observability tools.
About the Role ServiceNow is seeking a Staff Software Engineer – SRE & AIOps to drive infrastructure automation, operational resilience, and toil elimination across our hybrid cloud and data center operations.…
A people-manager role leading a distributed SRE team that owns end-to-end reliability of Google's F1 Query federated query engine and CDQ services, including Tier-1 on-call, incident response, and automation of manual toil. Day-to-day work spans mentoring engineers, design reviews, and technical guidance across C++, Java, and Go at planet scale.
Contract DevOps/SRE role on a client's Information Security team: build and run an LLM-based code vulnerability scanner on AWS Bedrock, with Jenkins/GitHub CI/CD, Terraform-provisioned ECS deployments, and Splunk telemetry, then support it via a 24/7 rotation. Requires 4+ years engineering experience with AWS, Terraform, CI/CD, and appsec; pays $75-95/hr W2.
Own and scale the reliability of Orkes’ cloud-native orchestration platform, ensuring high availability and performance for distributed workflows across fintech, e-commerce, logistics, and healthcare.
SRE/support engineer at Sber IT maintaining mission-critical, high-availability services in Moscow: monitoring, incident management and post-mortems, automating deployment and operations, and maintaining CI/CD pipelines. Core stack includes Linux, Python/Bash scripting, Docker, Kubernetes, Git and monitoring/alerting systems.
A contract SRE II role at VDart in Atlanta, GA (Northlake, DeKalb County), focused on ensuring the reliability, scalability, performance, and security of data platforms, analytics systems, AI/ML services, and BI applications. Day-to-day work blends software and systems engineering with automation and operational excellence.
Lead Site Reliability Engineer responsible for keeping large-scale systems reliable across multiple clouds (AWS, GCP, Azure, PCF/VMware), automating with Java/Python/Go and scripting, working with databases like Postgres, Cassandra and Redis, and using observability tools such as Splunk, Dynatrace and AppDynamics.
Senior Site Reliability Engineer at Anadea will help build the Infrastructure and DevOps department, working day to day with GCP infrastructure (Cloud Run, BigQuery, Cloud Storage, Networking), Terraform-based IaC, Docker, monitoring/logging, and CI/CD pipelines supporting mobile backends. Programming experience (TypeScript, JS, Go, or Python) is also expected.
On-site Site Reliability Engineer at Chile's Digital Government Secretariat in downtown Santiago, responsible for SLOs/error budgets, observability, infrastructure-as-code, CI/CD pipelines, incident management and on-call for critical public-sector platforms. Stack centers on Linux, AWS (or Azure/GCP), Docker/Kubernetes, Terraform, Prometheus/Grafana and GitLab CI.
Senior Site Reliability Engineer at Claranet in Lisbon (hybrid), splitting time 50/50 between operational excellence (incident response, troubleshooting, reliability) and engineering projects (automation, observability, platform improvements). Core stack is Azure and Kubernetes (AKS), with Terraform, CI/CD, and Datadog.
Design and implement monitoring, alerting, and reliability frameworks for Solaris’s embedded finance platform, ensuring high availability and resilience across microservices.
We couldn't check your fit for this role — add a CV to your profile to see it next time.