Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Senior Site Reliability Engineer at Anduril Industries to maintain and secure AI-powered defense systems and Lattice OS infrastructure.
Embed with AI labs to onboard, tune, and debug large-scale training/inference workloads on GPU clusters, owning reliability and automation for distributed systems.
Leads a team to ensure the reliability and security of Okta’s federal infrastructure, maintaining high availability and compliance with security standards.
Maintain and automate cloud infrastructure for a premium ride-hailing service, ensuring high availability and fast incident response using AWS, Kubernetes, and observability tools.
Senior engineer on Upstart’s Site Reliability team builds monitoring, automation, and incident-response tooling to keep the company’s AI-powered lending platform reliable and observable for millions of users.
Leads a team to improve the reliability, observability, and operational resilience of Upstart’s AI-powered lending platform, using cloud-native tools and incident management practices.
Own uptime and incident response for a freight-terminal automation platform using GCP, Postgres, and containerized services; build monitoring, automation, and remediation to keep trucks moving.
Senior SRE role at BeReal to design and operate scalable, reliable GCP infrastructure, enforce SRE practices, and automate Kubernetes-based systems for a fast-growing social network.
Lead the observability and reliability engineering strategy for Citi’s global financial infrastructure, ensuring high availability and performance of systems handling trillions in daily transactions.
Build and maintain the infrastructure that keeps OneTrust’s AI-Ready Governance Platform™ secure, reliable, and scalable for enterprise customers.
Build and lead the reliability of ClickHouse’s cloud infrastructure, ensuring high availability and performance while collaborating with engineering teams on scalable, fault-tolerant systems.
Senior backend engineer/SRE building and operating a cloud platform for EV charging, bidirectional energy, and smart-grid services using Python/Go, Azure, and observability tools.
Lead a team to ensure the reliability and availability of critical platforms in complex, multi-vendor environments as a Site Reliability Engineering authority.
Maintains and scales Kong’s cloud infrastructure to ensure high availability and performance, using tools like Terraform, Kubernetes, and Prometheus.
Lead Site Reliability Engineering (SRE) role managing a team of 3-5 engineers, ensuring high availability and scalability of SberAds' AI-driven ad platform using Kubernetes, Kafka, Terraform, and other tech.
Job Title: Senior Site Reliability Engineer Salary: $130 - $225k Location: Remote, USA Reports to: Chief Technology Officer Closing Date: N/A We reserve the right to close this vacancy early if we receive sufficient…
Oversee site reliability engineering for a global financial-data platform, ensuring high availability and performance of trading and charting services.
Oversee mechanical and electrical engineering works on large construction projects in Singapore, ensuring systems integration, testing, and compliance with safety and design standards.
Designs, operates, and improves highly available production systems using Linux, Python, cloud/DevOps tools, and monitoring/automation for a global consulting firm serving financial services and technology clients.
Designs, deploys, and maintains Kubernetes-based microservices for Cisco’s Webex collaboration platform using GitOps, Helm, and observability tools to ensure high availability and performance.
We couldn't check your fit for this role — add a CV to your profile to see it next time.