Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Join us to build the next generation of cloud-native reliability, release, and test platforms that enable engineering excellence, developer productivity, and high-confidence ServiceNow releases through automation,…
Lead reliability engineering for Sumo Logic’s observability and security products, optimizing cloud-native systems, automating operations, and mentoring teams to ensure high availability and security.
About Kinaxis Are you looking to join an innovative, market-leading company where you can truly elevate your career? At Kinaxis we are serious about culture, we are serious about technology, we are serious about…
TransUnion's Job Applicant Privacy Notice Personal Information We Collect Your Privacy Choices Team Overview At TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives…
For over 25 years, NVIDIA has been at the forefront of transforming computer graphics, PC gaming, and accelerated computing, driven by a legacy of continuous innovation and exceptional talent! We are now bringing to…
A senior SRE role at ServiceNow in Dublin focused on designing, building, and operating cloud-native engineering platforms for release validation, testing, and production readiness. Day-to-day work centers on Kubernetes platforms, CI/CD and GitOps automation, observability, progressive delivery, and developer productivity tooling, using languages like Python, Go, Java, or Ruby.
Staff-level Site Reliability Engineer at ServiceNow who designs automation-first infrastructure: enterprise-scale Kubernetes across hybrid/multi-cloud, closed-loop auto-remediation using AI/ML, SRE tooling, SLO frameworks, and IaC/GitOps pipelines. The role combines hands-on engineering (Kubernetes, AWS/Azure/GCP, Terraform, Python/Go/Bash) with mentoring and technical leadership across global 24/
Leads Google's AlphaNet Core SRE team in Waterloo, running and improving Google's SDN-enabled campus and WAN production networks. Day-to-day involves leading system designs, incident response and blameless postmortems, building automation, and mentoring engineers, using Go, Python, or C++ and deep distributed-systems and networking expertise.
Staff Site Reliability Engineer on Crunchyroll's Center for Data & Insights team, designing and operating reliable, secure, cloud-native data platforms on GCP and Kubernetes. Day-to-day spans SLIs/SLOs, observability, incident management, automation, capacity planning, disaster recovery, and SecOps work like vulnerability management and Kubernetes security.
Staff-level SRE on Google's Ads Quality Infrastructure (AQI) Data team, keeping Search Ads and SAGE/SPARK data processing reliable. Day to day: set reliability strategy, lead system design reviews, build data-pipeline resilience and observability, mentor engineers, and lead incident response with tier 2 oncall.
As a founding Staff SRE, you will build and scale the reliability foundations for Wayve's AI model development and GPU compute infrastructure. You will define operational standards, manage large-scale Kubernetes clusters, and ensure the performance of distributed systems supporting autonomous driving technology.
● Design, write, and deliver software to improve the availability, latency, and efficiency of Freshworks’ Products & Platforms. ● Develop scalable, cloud-native architectures that support business growth. ● Design…
The Staff Site Reliability Engineer will lead the reliability and availability of Waymo's autonomous fleet services and infrastructure. This role involves architecting mission-critical systems, driving operational standards, and mentoring teams using C++, Java, or Python.
Site Reliability Engineering (SRE) at NVIDIA is an engineering discipline to design, build and maintain large scale production systems with high efficiency and availability using the combination of software and systems…
Staff Engineer in Site Reliability at Nextiva, Bengaluru, responsible for ensuring the reliability and scalability of middleware and cloud infrastructure, primarily using Kafka, Kubernetes, GCP, and observability tools.
We Are Synopsys is the leader in engineering solutions from silicon to systems, enabling customers to rapidly innovate AI-powered products. We deliver industry-leading silicon design, IP, simulation and analysis…
Staff SRE on Altruist's Platform team, focused on the performance and resilience of backend systems across all product teams at a fintech serving financial advisors. Day to day: observability (OpenTelemetry, SLOs), incident response ownership, mentoring secure/performant coding, and shared infrastructure built on Java/Spring Boot, Kubernetes, and AWS.
A Staff Site Reliability Engineer at Airwallex in Singapore partners with Spend product teams to architect and run scalable, reliable cloud infrastructure for high-risk projects like new service launches and data centre migrations. Core work spans AWS/GCP, Kubernetes, observability, incident response, and SLO ownership.
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology—and amazing people. Today,…
We couldn't check your fit for this role — add a CV to your profile to see it next time.