Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Senior SRE engineer owning reliability, scalability, and operational excellence of SonicWall's cloud-based services on AWS — defining SLOs, building observability platforms, automating infrastructure with Terraform/Kubernetes, and driving toil reduction and chaos engineering.
Position Title: SRE Production Support with Crypto Experience: 6-8 years Mode of interview: (F2F Discussion) Work location-Alpharetta, GA Duration:6Months Must Have Requirements: 1. Production Support experience 2.…
Nosso Modo de Fazer no Time: Transforme sua carreira com o iFood! Somos uma empresa brasileira de tecnologia referência na América Latina. Por meio de soluções inovadoras, conectamos milhares de restaurantes a milhões…
Senior SRE/Incident Manager for iFood's platform, coordinating crisis response, managing end-to-end incident lifecycles, and performing root cause analysis to maintain reliability for Brazil's largest food delivery platform with 100M+ monthly orders. Role requires 24/7 availability, expertise in observability tools (Datadog, Grafana), Databricks/SQL for data analysis, and ITIL/Google SRE practices
Role: SRE / Observability Engineer (JD Attached) Remote to EST and CST Consultants Only We are conducting the first-round interview Contract Details Initial contract expected to last at least 12 months High likelihood…
Senior Site Reliability Engineer builds and runs a secure, scalable AWS/Kubernetes platform for home-health software, owns Postgres reliability, and matures observability and incident response.
The Site Reliability Engineer ensures the reliability and performance of technology platforms by managing infrastructure, monitoring systems, and leading disaster recovery efforts. The role involves using PowerShell, Python, and Bash to automate operations and improve system resiliency within a hybrid cloud environment.
Maintain and automate the reliability of a Brazilian neobank’s infrastructure, migrating workloads to Kubernetes, optimizing costs, and ensuring uptime during incidents.
About the Company If you have ever watched television or enjoyed a movie on your phone or tablet, this experience was likely brought to you through an Ateme solution created by our award-winning engineering teams.…
Site Reliability Engineer on a DevOps team building and operating reliable, scalable AWS cloud services and large-scale data delivery solutions for the Royal Society of Chemistry's digital platforms.
The Senior Site Reliability Engineer will manage and scale Maze's cloud infrastructure, focusing on reliability, security, and observability. The role involves automating workflows, supporting compliance initiatives, and collaborating with engineering teams using tools like AWS, Kubernetes, Terraform, and Datadog.
Designs and operates cloud-native infrastructure for satellite imaging data processing, ensuring reliability, scalability, and availability across customer environments (cloud/on-prem).
We are looking for a Senior SRE / Production Reliability Engineer to improve the reliability, performance, and availability of critical production systems. The ideal candidate should have strong experience in SRE/DevOps, Incident Management, Observability, Kubernetes, Google Cloud Platform, and production monitoring, along with hands-on experience with time-series anomaly detection.
Site Reliability Engineer maintaining legacy platforms and supporting cloud migration for a system modernization initiative, with monitoring and troubleshooting responsibilities.
The Staff SRE for Google Home will lead reliability, performance, and scalability efforts for Google's smart home infrastructure, focusing on distributed systems, ML model serving, and automation. This role involves architectural oversight and cross-functional collaboration to ensure high-availability services for millions of connected devices.
The Staff Software Engineer for Health SRE will design, scale, and maintain reliable infrastructure for Google Health products, including AI-powered tools and wearable tech. The role involves optimizing distributed systems, automating service lifecycles, and leading cross-functional technical initiatives.
Own and harden Sumo Logic’s planet-scale observability platform by writing automation, optimizing cloud resources, and improving reliability and security for microservices.
Owns observability/security reliability for Sumo Logic’s cloud-native products, optimizing cloud resources, reducing toil, and improving developer velocity via automation, SLOs, and incident analysis.
Lead reliability engineering for Sumo Logic’s observability and security products, optimizing cloud-native systems, automating operations, and mentoring teams to ensure high availability and security.
Lead a new Site Reliability and Security Engineering team in Dublin, ensuring the reliability, security, and scalability of Klaviyo’s global SaaS marketing platform and AI services.
We couldn't check your fit for this role — add a CV to your profile to see it next time.