Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Senior Site Reliability Engineer at Blackpoint Cyber designs, implements, and maintains cloud and on-premise infrastructure and CI/CD pipelines, focusing on automation, scalability, and performance. Works with AWS, Kubernetes, data streaming, observability, and incident response to ensure system reliability and security.
Lead SRE and platform engineering for large-scale customer-facing web applications on AWS, Kubernetes/EKS, and Terraform, driving reliability, automation, incident management, and operational standards while mentoring engineers.
The SRE ensures high availability and reliability for Feeld’s consumer mobile dating app by designing observability, monitoring, and incident response systems. Core tech: Node.js, TypeScript, AWS, Cloudflare, Sentry.
Applications: candidates may apply from outside Japan. Visa sponsorship: available. Japanese: Not required. English: Business level. Japan Dev tags: Engineering, Go, Python, Kubernetes, SRE. Are you looking To make…
Build and maintain the reliability infrastructure for Affirm’s backend systems, ensuring high availability and performance of fintech services using AWS, Kubernetes, and observability tools.
Builds and operates AI-driven agentic systems to automate SRE workflows (incident response, root cause analysis, SLO tracking) while maintaining hands-on production reliability for an AI enterprise company.
The Senior Site Reliability Engineer will manage the reliability, scalability, and observability of cloud-based infrastructure using Kubernetes, cloud networking, and AI-assisted engineering tools. The role involves participating in on-call rotations, defining SLIs/SLOs, and collaborating with product teams to optimize system performance.
Builds and maintains scalable, automated cloud infrastructure using SRE principles, leveraging AWS, Terraform, Kubernetes, and Python to ensure reliability, availability, and operational excellence while integrating AI tools for coding and troubleshooting.
Who You'll Work With We are seeking an experienced and analytically-minded Site Reliability Engineer to join our organisation on a permanent, remote basis from Ireland. In this role, you will be instrumental in…
The Sr. Site Reliability Engineer III will work within a collaborative team to deliver technical solutions for federal government clients. The role requires an active security clearance and expertise in site reliability engineering practices.
Design and maintain GitGuardian’s self-hosted Kubernetes infrastructure, ensuring reliability, security, and smooth monthly releases while automating CI/CD and observability for large enterprise customers.
SRE-инженер поддерживает и развивает инфраструктуру на базе Linux, Kubernetes, PostgreSQL и других инструментов, внедряет автоматизацию и мониторинг для обеспечения отказоустойчивости систем.
Senior SRE role maintaining a high-load sports betting platform: debugging production issues, tuning PostgreSQL/Redis/Kafka, and operating Kubernetes clusters to ensure uptime and performance.
Senior SRE managing multi-cloud AWS/Azure infrastructure, Kubernetes clusters, and PostgreSQL databases. Handles P1/P2 production incidents, builds IaC with Terraform/Ansible, and maintains observability and reliability frameworks.
Senior SRE specializing in Dynatrace observability for core banking systems, ensuring service reliability, performance monitoring, and automation in hybrid IT operations.
Senior SRE role building reliability, automation, and observability for Sysco’s commercial tech platforms using software engineering and cloud-native tools.
FinOps Senior SRE focused on embedding cost optimization into cloud infrastructure design, automation, and monitoring across Azure and AWS, including AI/GPU workloads, using Kubernetes, Terraform, and observability tooling.
Senior SRE maintaining and scaling large-scale Kubernetes clusters for NVIDIA's DGX Cloud AI platform, working with GPU workloads across major cloud providers using infrastructure automation and observability tools.
Senior SRE designing, building, and operating highly reliable hybrid (on-prem and AWS) platforms, with a focus on automation, resiliency, observability, and FinOps using Terraform, Python/Java, and a broad AWS service stack.
Leads reliability engineering for ultra-high-availability 9-1-1 call-routing SaaS systems, focusing on observability, incident response, and SLO management in a public safety tech environment.
We couldn't check your fit for this role — add a CV to your profile to see it next time.