SRE Jobs in United States
There are 1,390 open SRE jobs in United States on freehire right now. 314 of them were posted recently. The skills employers ask for most often are cloud, automation and observability.
Salary
| Currency | Period | 25th | Median | 75th | Postings |
|---|---|---|---|---|---|
| USD | year | $160,875 | $187,000 | $208,754 | 323 |
From postings that disclose pay. Currencies are counted separately, never converted.
Most requested skills
- cloud 67%
- automation 63%
- observability 55%
- python 52%
- kubernetes 51%
- aws 45%
- ci-cd 44%
- ai 42%
How the work is done
- Remote 268 · 19%
- Hybrid 219 · 16%
- Onsite 73 · 5%
Visa sponsorship offered in 42% of the 317 postings that state a position on it.
Seniority
- Senior 470
- Staff 107
- Lead 97
- Principal 52
- C-level 25
- Junior 10
Who is hiring
- 1000+ employees 378
- 501-1000 employees 129
- 51-200 employees 33
- 11-50 employees 25
- 201-500 employees 5

Lead Site Reliability Engineer
Lead a team to architect, deploy, and optimize secure, scalable infrastructure for a mission-critical collaboration platform used by defense and government sectors.
Senior Site Reliability Engineer - Storage
Build and maintain Lambda’s high-performance storage infrastructure, ensuring reliability and scalability for AI workloads while automating incident response and fleet management.
Senior SRE Platform Software Engineer
Build and operate the NeoCloud SRE platform using GitOps/CI/CD, ensuring SLOs and preventing drift while maintaining on-call ownership and documentation.
Site Reliability Engineer
Design and maintain scalable cloud infrastructure for an auto-repair SaaS, using AWS/GCP, Kubernetes, and CI/CD pipelines to ensure reliability and performance.
Site Reliability Engineer
Builds automation and improves reliability for production systems using Python/R and ML libraries to reduce operational toil and enhance service availability.
SRE Solutions Architect – AI, HPC & GPU
Designs and maintains reliable, scalable AI/GPU and HPC infrastructure, focusing on observability, automation, and performance optimization for high-performance computing environments.
Site Reliability Engineer
Maintains and optimizes cloud infrastructure using Kubernetes/AKS and Azure DevOps, focusing on reliability, failover, and performance testing.
Sr. SRE (Storage Platforms)- W2 only
Senior SRE designs and maintains private-cloud storage and Kubernetes platforms using SDS, GitOps, and IaC to ensure scalable, resilient, and high-performance infrastructure.
Site Reliability Engineer
Maintain and automate healthcare IT infrastructure using cloud tools (AWS/Azure), IaC (Terraform/Ansible), and monitoring (Splunk/Zenoss) while debugging issues and supporting CI/CD pipelines.
Sr. Site Reliability Engineer (Compute Platform)
Senior Site Reliability Engineer designs and maintains Kubernetes clusters on bare-metal and hypervisor platforms in a private cloud, ensuring enterprise-grade compute and hypervisor environments are stable and standardized.
Site Reliability Engineer (Remote)
Build and operate ultra-reliable, low-latency cloud-native systems for global trading, using SRE practices, AI-driven automation, and observability tooling to keep 24/7 markets running smoothly.
Staff Site Reliability Engineer (Hybrid)
Lead the reliability and scalability of Cisco’s Splunk Agent Observability platform, designing Kubernetes-based infrastructure, SLOs, and automation to ensure resilient AI agent deployments in cloud and on-prem environments.
Senior Site Reliability Engineer (Remote)
Build and maintain the Kubernetes-based infrastructure and deployment systems that keep Splunk’s AI agent observability platform reliable and secure across cloud and air-gapped environments.
Senior Software / Site Reliability Lead Engineer
Lead the reliability engineering practice for AI services at a defense contractor, defining SLOs, observability stacks, and incident response while ensuring production readiness and eliminating operational toil.
Site Reliability Engineer
Build and maintain cloud-native telecom infrastructure using Kubernetes, Terraform, and CI/CD pipelines to ensure high availability and security for a cybersecurity product.
