Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Own and harden Fivetran’s cloud infrastructure, Kubernetes clusters, and deployment pipelines to keep data pipelines reliable and secure at scale.
Build and scale the reliability and observability platforms for OneTrust’s AI-Ready Governance Platform using Java, Spring Boot, and cloud-native tools.
Staff SRE responsible for optimizing cloud costs and efficiency across Kubernetes clusters and AWS infrastructure, partnering with FinOps and engineering teams to reduce spend and improve resource management.
Fully remote Staff SRE at StackBlitz (makers of Bolt.new and WebContainers) who embeds with product and platform teams from design through launch, defines SLIs/SLOs and production-readiness standards, and builds paved-road tooling across AWS, GCP, and Azure with Terraform — working daily with TypeScript and Ruby on Rails services and sharing a weekly-per-month on-call rotation.
Senior Staff Site Reliability Engineer at Fivetran will own the performance, reliability, and incident response of the production data platform infrastructure — monitoring availability and capacity, automating deployments, and hardening security. Core stack includes managed Kubernetes (EKS/AKS/GKE), AWS/Azure/GCP, Terraform, ArgoCD, Python, and Linux, working hybrid from the Oakland office.
Fingerprint (device intelligence/fraud detection) is hiring its first dedicated Site Reliability Engineer. This remote role defines SLIs/SLOs and error budgets, strengthens incident response and alerting, and coaches engineering teams in reliability practices while staying hands-on with Kubernetes, AWS, Datadog, and Go/TypeScript.
Build and maintain GitLab's production infrastructure, automate operations with Kubernetes and Go, and ensure reliable, scalable services using observability and SLOs.
Staff Backend Engineer builds and operates observability, anomaly detection, and reconciliation tooling for GitLab’s monetization stack, using Ruby on Rails, Prometheus, Grafana, and ML to prevent billing and data issues.
Build and run the core network infrastructure for a cloud platform powering AI workloads, ensuring high availability, observability, and safe automation at scale.
A senior/staff Java engineer role on a global cryptocurrency exchange's Core Infrastructure & Site Reliability team in Singapore, focused on proactively eliminating systemic risks in critical trading systems. Core stack: Java/Spring Boot, plus Nginx, Redis, ElasticSearch, Kafka, MySQL, Docker, and Kubernetes.
Designs, builds, and operates highly scalable, secure cloud infrastructure across AWS and GCP, leading Kubernetes migrations and SRE initiatives while mentoring teams.
The Staff Site Reliability Engineer will design, build, and operate secure, scalable cloud network infrastructure for Okta's Federal products. The role requires deep expertise in AWS networking, automation with Terraform, and navigating complex federal compliance frameworks like FedRAMP and IL6.
Staff Site Reliability Engineer specializing in Splunk to design, automate, and optimize Okta's observability platform, ensuring high reliability and low latency for distributed systems using Go, Python, or Ruby and infrastructure-as-code.
Staff Site Reliability Engineer designs and automates security-hardening for Okta’s GCP/AWS infrastructure, builds IAM policies, PKI, and container security, and drives incident response and prevention at scale.
Designs and maintains secure, air-gapped cloud infrastructure for government clients, ensuring high reliability and compliance with strict security standards.
Designs and maintains secure, highly available cloud infrastructure for Okta’s identity platform using AWS, Terraform, and automation tools, ensuring reliability and scalability.
Lead reliability engineering for Okta’s Federal SRE team, designing and operating secure, scalable cloud services for government customers using Kubernetes, Terraform, and Go/Python.
Build and manage Okta’s Kubernetes platforms on AWS, automating deployments with Helm, Karpenter, and Istio to ensure scalable, secure, and high-availability cloud-native infrastructure.
Staff SRE who designs, builds, and operates Okta's Kubernetes platforms on AWS: creating highly available clusters, automating deployments with Helm and Terraform, dynamically scaling with Karpenter, managing the Istio service mesh, and handling incident response, security, and cost optimization for cloud-native services.
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally…
We couldn't check your fit for this role — add a CV to your profile to see it next time.