Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Lead reliability engineering for a European-scale healthcare platform, driving SLOs, observability, and incident response while mentoring teams on Kubernetes, cloud-native tooling, and best practices.
Lead reliability engineering for a European healthcare platform, driving SLOs, observability, and incident response across 170+ apps to keep 90M patients and 520K health professionals online.
Build and maintain the infrastructure that keeps Anthropic’s AI systems safe and reliable, focusing on ML safeguards and operational stability.
Senior Staff Site Reliability Engineer (SRE) for AI/ML infrastructure in San Jose, CA. Designs, scales, and maintains resilient production systems.
Staff SRE to design and maintain highly available GCP infrastructure for voice AI services handling millions of interactions, focusing on reliability, automation, and observability.
Staff Site Reliability Engineer at SHEIN, maintaining 24/7 global ecommerce systems using Kubernetes, Kafka, and observability tools while automating operations and mentoring teams.
Talonic | Staff Site Reliability Engineer (founding) | Berlin, Germany | HYBRID | €80–90k + VSOP | jobs@talonic.ai We do enterprise document ingestion — capture a document once into a canonical field registry, then…
Senior engineer leading reliability for Google’s infrastructure, designing scalable systems and ensuring uptime for global services.
Own and scale the reliability, performance, and cost of ManyChat's AI infrastructure, including LLM inference services and AI Gateway, while shaping standards for the company's AI platform.
WHO WE ARE 🌍 Creating content that resonates is great — turning that attention into growth is even better. That's what Manychat does. Our AI-powered automations help creators and brands engage with audiences…
Staff SRE leads the reliability, performance, and security of a petabyte-scale genomics data platform (ClickHouse/OLAP) while shaping its Data-as-a-Service and AI-agent integrations, mentoring engineers, and defining technical direction for multi-tenant data infrastructure.
Design and run VMware-based private clouds, automate Linux/Windows infrastructure, and maintain F5 load balancers to keep a global SaaS platform highly available and secure.
Senior SRE manages and optimizes a high-scale AWS infrastructure (6 regions, 3M+ req/sec) for an AdTech company, focusing on automation, reliability, performance tuning, and cross-team collaboration to ensure seamless ad-platform operations.
Lead reliability engineering for Develocity, Gradle’s AI-native observability and build toolchain platform, ensuring high availability and performance for customers like Netflix and SAP.
Build and run the Kubernetes-based infrastructure that powers Circle’s blockchain platform, including full nodes for Arc, Ethereum, Solana, and Base, while automating deployments and reliability with Terraform, CI/CD, and AI-driven tooling.
Senior SRE role defining reliability strategy for Filevine’s Legal Operating Intelligence platform, using AI-driven observability and Kubernetes to ensure scalable, secure production systems.
We Are Synopsys is the leader in engineering solutions from silicon to systems, enabling customers to rapidly innovate AI-powered products. We deliver industry-leading silicon design, IP, simulation and analysis…
Build and maintain Trimble’s observability platform using AI-driven automation, OpenTelemetry, and cloud-native tooling to keep services reliable and cost-efficient.
Build and secure Okta’s global network fabric and multi-cloud infrastructure, automating edge routing, enforcing zero-trust mTLS, and defending against DDoS attacks using AI-driven SRE practices.
Lead a small team of site-reliability engineers, owning end-to-end production reliability for cloud, vehicle pipelines, and security while coordinating major incidents and driving systemic fixes.
We couldn't check your fit for this role — add a CV to your profile to see it next time.