Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
As a Field/Site Reliability Engineering intern, you will maintain and improve the reliability of robotic systems by automating diagnostics, debugging hardware and software, and supporting field deployments. The role involves hands-on work with Python, C++, Linux, and various electromechanical systems.
Manage a team of engineers to ensure Memorystore’s in-memory database service remains reliable, scalable, and secure for global customers.
Миссия роли: усилить команду, которая отвечает за выстраивание концепции «мониторинга как сервис» в рамках департамента (в том числе встраивание в текущий релизный/производственный процессы департамента), сопровождение…
Senior Site Reliability Engineer responsible for maintaining and improving the reliability, performance, and scalability of distributed software systems using cloud platforms, automation, and observability tools.
Hybrid Cloud SRE responsible for cloud platform delivery, daily operations, alarm handling, and stability improvement in hybrid cloud scenarios, working with Linux, containers, K8s, and scripting languages.
The Site Reliability Engineer will manage and optimize Azure-based cloud infrastructure, CI/CD pipelines, and observability tools to ensure the reliability of AI-driven cancer diagnostic platforms. The role involves automating operational tasks, participating in an on-call rotation, and leading incident reviews.
Designs, implements, and maintains a reliable, scalable hybrid application platform with a focus on cloud databases and self-service tooling. Leads CICD automation, IaC, and DevOps strategies while collaborating with engineering, security, and business teams.
Maintain and optimize Backblaze’s distributed MySQL (Vitess) and Cassandra databases, automate operations, and ensure high availability and performance for cloud storage services.
Leads global iCloud SRE teams to design, automate, and maintain ultra-high-availability cloud services at Apple scale, ensuring seamless customer experiences worldwide.
The Site Reliability Engineer will manage, maintain, and improve the Tyk API Management platform, focusing on cloud infrastructure reliability, automation, and incident management. The role involves working with Kubernetes, AWS, and various monitoring tools within a distributed, remote-first team.
This 6-month internship involves supporting the Site Reliability Engineering team with Jira and Confluence cloud migrations, operational reporting, and documentation. The intern will gain hands-on experience in enterprise platform administration, data analytics, and process improvement within a global financial messaging environment.
The Principal Software Engineer will join the SRE & Architecture team to improve the resilience, performance, and cost-efficiency of the cloud-native SOPHiA DDM platform. The role involves hands-on full-stack development, architectural design, and the use of AI-assisted tools to optimize system operations.
This Manager role involves leading teams to design and deploy AI-driven security solutions and data infrastructure within a cybersecurity practice. The position focuses on strategic planning, mentoring staff, and utilizing Python and C++ to implement machine learning models for enterprise and cloud security.
Maintains and improves the reliability of cloud-native systems by monitoring with Dynatrace, managing Kubernetes workloads, and resolving incidents via ServiceNow.
Senior Staff SRE leading compute infrastructure initiatives across on-prem and cloud, designing and scaling core services like DNS, NTP, DHCP, and LDAP at global scale with a focus on automation, monitoring, and capacity planning.
Designs and maintains scalable cloud-based systems (AWS/AliCloud) for a fintech POS/SaaS provider, focusing on reliability, observability, and automation to support SME digitization across Asia.
Lead SRE focused on designing and implementing AI-driven solutions for observability, intelligent alerting, and automation in fintech applications, ensuring reliability and scalability of AI-powered infrastructure.
The Principal Site Reliability Engineer leads production support, incident management, and infrastructure automation for diabetes technology systems. The role focuses on implementing SRE practices like SLOs, IaC with Terraform, and CI/CD reliability while mentoring a distributed team.
The Senior Site Reliability Engineer will design, build, and maintain scalable, secure infrastructure while leading platform engineering initiatives and incident management. The role focuses on automating operational workflows, implementing observability, and driving DevSecOps practices across cloud and hybrid environments.
The Principal Software Engineer (SRE) will lead reliability, scalability, and resiliency initiatives for Onshape's large-scale distributed Java-based platform. This role involves driving architectural improvements, managing incident response, and mentoring engineering teams to ensure operational excellence.
We couldn't check your fit for this role — add a CV to your profile to see it next time.