freehire launches on Product Hunt on 26 August.

Follow →

Senior Site Reliability Engineer

Open 29d

Summary

Senior Site Reliability Engineer responsible for ensuring the availability, performance, and scalability of a SaaS platform through cloud-native architectures, observability, and automation.

OneRail seeks a Senior Site Reliability Engineer to help ensure the availability, performance, scalability, observability, and resilience of our SaaS platform. In this role, you’ll be responsible for designing and optimizing platform architecture, improving system reliability, driving automation initiatives, and enabling engineering teams to build and operate highly scalable services.

This is a hands-on role for someone with deep expertise in cloud-native platforms, distributed systems, observability, and performance optimization. You’ll work closely with Engineering, Product, and Operations teams to improve platform reliability, scalability, and operational efficiency while reducing manual overhead through automation. Over time, you’ll have opportunities to influence platform architecture, lead reliability initiatives, and establish Site Reliability Engineering best practices across the organization.
  • Bachelor’s degree in Computer Science, Engineering, or a related technical field.
  • 5+ years of experience in Site Reliability Engineering, Platform Engineering or a related role.
  • Strong experience with cloud-native architectures and distributed systems.
  • Hands-on experience implementing observability solutions, including metrics, logging, tracing, and application performance monitoring.
  • Experience designing scalable backend architecture using Node.js, TypeScript, .NET
  • Strong knowledge of database administration, performance tuning, and optimization, including Azure Cosmos DB, MySQL, or similar platforms.
  • Experience building automation frameworks, scripting solutions, and operational tooling.
  • Familiarity with CI/CD pipelines and continuous integration practices using GitHub Actions or similar platforms.
  • Experience with containerization technologies such as Docker.
  • Strong understanding of event-driven architectures, messaging systems, and real-time data processing platforms.
  • Experience with monitoring and observability tools such as Datadog, OpenTelemetry, ELK Stack, App Insights, Grafana, or Prometheus.
  • Excellent troubleshooting, analytical, and problem-solving skills.
  • Strong written and verbal communication skills with the ability to collaborate across technical and business teams.
  • Advanced proficiency in English & Polish, both written and spoken (B2+).

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available