Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Leads a team of SREs to design, implement, and maintain cloud/on-prem infrastructure observability, ensuring reliability, scalability, and automation across T. Rowe Price’s tech stack using AWS, DevOps tools, and chaos engineering.
The Principal Site Reliability Engineer will lead the development and implementation of SRE practices to ensure the observability, scalability, and reliability of cloud and on-prem infrastructure. The role involves hands-on engineering, automation, and strategic leadership to prevent service disruptions across a complex, distributed environment.
Position Overview At PNC, our people are our greatest differentiator and competitive advantage in the markets we serve. We are all united in delivering the best experience for our customers. We work together each day…
As a Principal member of the Site Reliability Engineering (SRE) team, you'll take ownership of highly available systems, influence service design, and work across teams to drive resiliency, automation, and…
Owns cloud platform reliability and scalability for a video sensor AI startup, troubleshooting Kubernetes/AWS issues, improving observability, and automating ops to reduce toil.
Software Guidance & Assistance, Inc., (SGA), is searching for a Database Site Reliability Engineer for a CONTRACT assignment with one of our premier Financial Services clients in Alpharetta, GA. We are seeking an…
Define and strengthen reliability, scalability, and operational excellence for a healthcare reimbursement platform using cloud infrastructure, observability, and automation.
Managing Consultant in SRE helping organisations transform their cloud operating models and reliability practices, advising senior stakeholders in an onsite capacity in Central London.
Staff SRE leads Plaid’s release engineering, designing reliability frameworks (SLOs, error budgets) and progressive delivery systems to enable safe, high-velocity deployments across fintech platforms.
Senior SRE builds and operates Fastly’s global edge network, focusing on BGP/IP routing, automation, and monitoring to ensure reliable internet-scale performance.
Principal Engineer at commercetools designs and implements resiliency processes for mission-critical commerce infrastructure, ensuring system reliability during high-traffic events like Black Friday using incident management, metrics, and cross-team collaboration.
Principal Engineer builds and operates AI-driven commerce features end-to-end, from agent orchestration and generative UI to production deployment and observability, using LLMs, MCP, and full-stack development.
Designs, maintains, and secures scalable cloud infrastructure for production systems, focusing on availability, performance, and automated deployments using IaC, Kubernetes, and monitoring tools like Prometheus/Grafana.
Design and operate a Kubernetes-based platform for Axon’s evidence storage systems, using GitOps, Terraform, and observability tools to ensure reliability and scalability.
Site Reliability Engineer focused on ensuring vLLM’s AI inference engine operates reliably, scalably, and with minimal downtime by designing resilient systems, improving observability, and driving incident response improvements.
Senior Site Reliability Engineer designs, deploys, and maintains cloud-based AI infrastructure on Kubernetes and CNCF tools, ensuring reliability and performance while mentoring teams and customers.
Who is Sonar? Sonar is driving the future of agent-centric software development. As the leader in AI code review and verification, we solve a critical problem: ensuring that software generated by AI-assisted developers…
Are you interested in working with the World’s leading AI-first Quality Engineering Company? Ready to advance your career, team up with global thought leaders across industries and make a difference every day? Join us…
Ensures reliability, monitoring, and operational health of Apple’s production data center services in Shanghai, focusing on security-critical hardware lifecycle services. Builds automation, resolves incidents, and partners with engineering teams to improve system resilience for global Apple products.
Designs and maintains GCP network architectures for a global conversational AI platform, ensuring high availability, security, and performance while collaborating with SRE and platform teams.
We couldn't check your fit for this role — add a CV to your profile to see it next time.