Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
The Site Reliability Engineer will manage and maintain highly available infrastructure, focusing on AWS and Azure environments. The role involves automating workflows, deploying CI/CD pipelines, and ensuring system performance through monitoring and infrastructure-as-code practices.
Lead Site Reliability Engineer at JPMorgan Chase's AIML Platform team, leading resiliency design reviews, breaking down complex problems for engineers, mentoring, using AI for incident triage and post-incident analysis, and improving reliability with observability and CI/CD.
About the Team The team builds and operates large-scale, massively distributed infrastructures, applying Site Reliability Engineering (SRE) principles of software and systems engineering to ensure our traffic services…
Senior SRE role building reliability, automation, and observability for Sysco’s commercial tech platforms using software engineering and cloud-native tools.
FinOps Senior SRE focused on embedding cost optimization into cloud infrastructure design, automation, and monitoring across Azure and AWS, including AI/GPU workloads, using Kubernetes, Terraform, and observability tooling.
Lead a team to design, scale, and harden Zoom’s DevOps platform—Kubernetes, cloud, and FedRAMP-rated government environments—while driving SRE best practices and cross-team reliability improvements.
Senior SRE maintaining and scaling large-scale Kubernetes clusters for NVIDIA's DGX Cloud AI platform, working with GPU workloads across major cloud providers using infrastructure automation and observability tools.
The Site Reliability Engineer will manage cloud infrastructure, CI/CD pipelines, and database reliability while leveraging AI-assisted coding tools to improve system performance and scalability. This remote-first role involves collaborating with engineering teams to maintain observability, security, and disaster recovery protocols for Haven's digital platforms.
Senior SRE designing, building, and operating highly reliable hybrid (on-prem and AWS) platforms, with a focus on automation, resiliency, observability, and FinOps using Terraform, Python/Java, and a broad AWS service stack.
Ensures reliability and stability of Bank of America’s Capital Markets and Investment Banking platforms by monitoring, troubleshooting, and resolving production incidents, automating operational tasks, and improving system resilience using Java, Spring Boot, SQL, and SRE best practices.
Full Stack Developer on the SRE team building and maintaining scalable web applications with .NET (C#), Angular, and SQL Server, focusing on platform reliability, automation, and incident response.
The Site Reliability and Security Operations Engineer will support Core Speech products by managing deployments, automating security patching, and maintaining system performance and monitoring. The role involves collaborating on architecture improvements and participating in an on-call rotation for production support.
The Site Reliability Engineer III will work with development and platform teams to migrate and maintain applications in Google Cloud. The role focuses on ensuring operational resiliency, managing incident response, and implementing automation to improve system health and performance.
Leads reliability engineering for ultra-high-availability 9-1-1 call-routing SaaS systems, focusing on observability, incident response, and SLO management in a public safety tech environment.
The FinOps Senior Site Reliability Engineer II optimizes cloud infrastructure costs across AWS and Azure while supporting AI and machine learning platforms. The role focuses on implementing automation, observability, and financial governance to improve resource efficiency and platform reliability.
The Senior Director of SRE will lead and grow a team of software engineers to ensure the scalability and reliability of retail and digital consumer platforms. This role involves managing both internal and outsourced teams while driving continuous improvement and innovation in an agile environment.
Leads AI-driven Autonomous SRE transformation at Palo Alto Networks, architecting self-healing, scalable cloud-native platforms and integrating AI/ML tools to automate infrastructure operations and observability.
Leads enterprise observability and SRE strategy to ensure reliability and performance of Gap Inc.’s digital and in-store systems using automation, monitoring, and real-time insights.
Build and maintain automation for financial platforms on GCP/GKE, ensuring reliability and observability while troubleshooting production issues and driving continuous improvement.
The Senior SRE will manage and improve production and development infrastructure for a high-volume fintech company. The role focuses on implementing CI/CD pipelines, automating infrastructure workflows, and supporting Linux-based systems.
We couldn't check your fit for this role — add a CV to your profile to see it next time.