Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Designs and maintains cloud infrastructure for a fintech startup, focusing on scalability, reliability, and cost optimization using AWS, Terraform, and DevOps practices. Leads observability, load testing, and chaos engineering initiatives while mentoring development teams.
SRE/Incident Operations Specialist maintaining production stability for a financial-sector client — handling incidents, analyzing logs and metrics, and automating operations using Datadog, Chronosphere, PagerDuty, Python, and Kafka.
Grow with us Are you detail-oriented with experience in software, appliances, or system engineering? Are you passionate about designing and operating cloud-native production services at scale? Do you enjoy…
Build and maintain Kubernetes-based infrastructure for managed cloud services like Nextcloud and IONOS GPT, focusing on reliability, automation, and observability.
Site Reliability Engineer maintaining and evolving a Kubernetes-based product platform (Managed Nextcloud, IONOS GPT, web services) at a European cloud/hosting provider, using Linux, Terraform, GitLab CI/CD, ArgoCD, Prometheus, and Grafana.
The Site Reliability Engineer will manage and optimize the containerized infrastructure for IONOS products like Nextcloud and IONOS GPT using Kubernetes, Terraform, and CI/CD pipelines. The role involves ensuring system stability, automating infrastructure tasks, and maintaining monitoring and logging solutions.
Lead observability and monitoring for enterprise-scale platforms, defining strategy, standards, and reusable tools to improve system reliability and developer productivity across Kubernetes/Azure environments.
Site Reliability Engineer (SRE) DevSecOps | Cloud Engineering | Observability | Production Environments | London SR2 is supporting a major long term programme and looking for an experienced Site Reliability Engineer…
The Lead Site Reliability Engineer will ensure the stability, scalability, and performance of Mastercard's global payment platforms. The role involves driving operational readiness, automating infrastructure, performing root cause analysis, and mentoring junior engineers using Java, Spring, and DevOps practices.
Senior SRE ensures high availability, security, and performance of healthcare platform systems by designing observability tools, leading incident response, optimizing Kubernetes/AWS environments, and automating infrastructure via Terraform.
Senior SRE ensuring reliability and health of ServiceTitan's cloud platform — designing SLO-based observability, operating Kubernetes at scale across AWS/Azure, automating incident response, and leveraging AI-assisted tooling.
The Site Reliability Engineering Manager oversees production readiness, system stability, and performance for Mastercard's payment platforms. The role involves driving automation, managing incident response, and collaborating with cross-functional teams to ensure scalable and fault-tolerant services.
Lead SRE building a greenfield reliability platform for Imunify360, a Linux server security suite running on hundreds of thousands of servers. You'll define SLIs/SLOs for ~70 components, build telemetry pipelines, and establish alerting and on-call practices using Prometheus, Grafana, ClickHouse, and Python/Go/Rust.
Senior SRE maintaining reliable, secure production systems for a cybersecurity startup, using Kubernetes, cloud infrastructure, IaC (Terraform), CI/CD, and AI-assisted observability/automation tooling across the APAC region.
Lead SRE responsible for the reliability, scalability, and operational excellence of Bank of America's Internal Kubernetes Container Platform (IKCP), driving automation, observability, and incident management across OpenShift, Kubernetes, and Rancher environments.
Lead site reliability, stability, and modernization of corporate banking payment platforms (domestic, international, instant, digital asset) using Java/J2EE, IBM MQ, OpenShift, Kafka, and MongoDB in an on-site role at PNC.
Intermediate SRE maintaining Infrastructure as Code with Terraform, managing Kubernetes microservices on GCP/AWS, building observability with Datadog, and participating in a 24/7 follow-the-sun incident rotation at Equifax in Pune.
Lead SRE responsible for the reliability, scalability, and operational excellence of Bank of America's Internal Kubernetes Container Platform (IKCP), driving automation, observability, incident management, and platform governance across OpenShift, Kubernetes, Rancher, and VKS.
The Site Reliability Engineer will manage and optimize production-grade Kubernetes clusters and automated deployment pipelines using GitOps practices. The role involves infrastructure as code, observability, and collaborating with engineering teams to improve system reliability and developer experience.
The Senior Site Reliability Engineer will design, build, and operate scalable payment systems on Microsoft Azure using GitOps, Kubernetes, and modern CI/CD practices. The role focuses on automation, observability, and maintaining high system reliability through infrastructure-as-code and incident management.
We couldn't check your fit for this role — add a CV to your profile to see it next time.