Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Senior SRE on Google Cloud's Technical Infrastructure team in Warsaw, responsible for the full lifecycle of large-scale distributed systems—designing, deploying, monitoring, and automating to ensure reliability and performance at Google scale.
Site Reliability Engineer owning application-level infrastructure—Kubernetes deployments, Terraform configs, CI/CD pipelines, SLO/SLI observability via Datadog/OpenTelemetry, capacity planning, and incident response—on GCP/GKE for a video-game commerce and payments platform.
The Site Reliability Engineer will manage application infrastructure, CI/CD pipelines, and observability tools to ensure system reliability. The role involves capacity planning, incident response, and automating operational tasks using technologies like Kubernetes, Terraform, and GCP.
O time de Engenharia da squad de Infraestrutura está precisando de alguém, e queremos você com a gente! Sua missão? Você será a pessoa chave para garantir a estabilidade e a performance de nossos sistemas,…
The Site Reliability Engineer will manage and optimize production eFX and crypto platforms, focusing on Kubernetes, automation, and observability. The role involves incident response, infrastructure design, and close collaboration with development teams to ensure high availability and performance.
Location: Newbury Hybrid Salary: Excellent basic salary plus bonus and Vodafone benefits Working hours: Full time 37.5 hours per week - Monday to Friday Hybrid We believe that through collaboration and connection with…
Site Reliability Engineer maintaining AWS/EKS cloud infrastructure, CI/CD pipelines (GitHub Actions), and observability tooling (Datadog, Prometheus) for a cyber risk analytics provider serving the insurance industry.
Site Reliability Engineer on the Platform Infrastructure team building and scaling the control plane and data plane that orchestrate Ray clusters across cloud environments, using Kubernetes, Go, and Python.
Senior SRE applying software engineering practices to operations at scale — monitoring SLOs, designing automation, and improving resiliency — using cloud infrastructure (Azure preferred), containers (Docker/Kubernetes), IaC tools (Terraform/Chef/Ansible), and languages like Python, Go, or Java.
Senior SRE owning production services end-to-end: design, deploy, monitor with Prometheus/Grafana, and lead incident response while mentoring junior engineers and integrating AI-ops tools.
The SRE ensures production systems on AWS are reliable, available, and performant by leading incident response, defining SLIs/SLOs, and improving observability. Daily work includes troubleshooting ECS Fargate, AWS networking, messaging (RabbitMQ/AmazonMQ), and mentoring teams on best practices.
Senior SRE maintaining production infrastructure for cloud-native core banking and payments platforms, working with Kubernetes, Terraform, and GCP/AWS in London.
The Site Reliability Engineer will manage and automate operational tasks for identity and access services using Linux and various scripting languages like Python or Java. The role focuses on maintaining system reliability and network infrastructure within AWS.
Build and secure IMC’s trading platforms by embedding DevSecOps practices, secret management, and identity architecture into CI/CD pipelines and Kubernetes infrastructure.
The Associate Architect will design and maintain scalable, secure cloud infrastructure while driving reliability through automation, monitoring, and incident management. The role requires extensive experience with cloud platforms, Kubernetes, IaC tools, and observability frameworks to ensure system stability.
Hands-on SRE Architect / Principal Engineer defining infrastructure and reliability architecture for a modernization journey centered on AWS, Terraform, GitHub Actions, and ECS/Fargate, bridging enterprise architecture and engineering teams.
Maintains and scales microservices-based delivery and e-commerce platforms, ensuring high availability and performance for millions of users in Thailand.
The Site Reliability Engineer will design, build, and maintain scalable physical and virtual infrastructure while automating processes to ensure system reliability. The role involves working across the full stack, from bare metal to cloud-based applications, using tools like Terraform and Ansible.
About the Role VESSL AI의 Senior Site Reliability Engineer는 VESSL GPU 클라우드 플랫폼의 가용성과 성능을 책임지며, 개별 장애 대응을 넘어 시스템 설계와 아키텍처 결정을 리드합니다. Observability와 자동화 체계를 구축하는 데 그치지 않고, 장애가 발생하기 전에 구조적 리스크를 발견해 제거하며, 팀 전반의 안정성 관행을…
About the Role VESSL AI의 Senior Site Reliability Engineer 는 VESSL GPU 클라우드 플랫폼의 가용성과 성능을 책임집니다. 메트릭·로그·트레이스 기반의 Observability 체계를 구축해 이상 징후를 조기에 탐지하고, 장애 발생 시 On-call 대응과 Root Cause Analysis를 통해 신속하게 복구하며, 반복적인 운영 작업은…
We couldn't check your fit for this role — add a CV to your profile to see it next time.