freehire launches on Product Hunt on 26 August.

Follow →

Site Reliability Engineer

Open 36d

We are looking for an experienced Site Reliability Engineer (SRE) to help build and operate a reliable, scalable, and secure cloud platform. In this role, you will work closely with cross-functional engineering teams to automate infrastructure, improve system observability, and enable fast, safe, and efficient software delivery.

Your responsibilities:

  • Design, implement, and maintain scalable cloud infrastructure on AWS.

  • Develop and manage Infrastructure as Code using Terraform.

  • Automate infrastructure provisioning, configuration management, and operational workflows.

  • Build and optimize CI/CD pipelines to support reliable software delivery.

  • Monitor system reliability, availability, and performance through observability solutions, including monitoring, logging, tracing, and alerting.

  • Lead incident response, troubleshooting, root cause analysis, and post-incident reviews.

  • Collaborate with engineering teams to improve application reliability, performance, and operational readiness.

  • Contribute to security, disaster recovery, and business continuity initiatives.

  • Reduce operational overhead through automation and continuous process improvement.

  • Promote Site Reliability Engineering and DevOps best practices across engineering teams.

We are looking for you, if you have:

  • Experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure roles.

  • Strong hands-on experience with AWS.

  • Proven experience with Terraform and Infrastructure as Code (IaC).

  • Experience designing and maintaining CI/CD pipelines.

  • Proficiency in scripting or programming with Python, Bash, or similar languages.

  • Solid experience administering and troubleshooting Linux-based environments.

  • Good understanding of networking, cloud architecture, and security principles.

  • Experience implementing monitoring, logging, and observability solutions.

  • Strong analytical, troubleshooting, and problem-solving skills.

  • Excellent communication and collaboration skills.

Nice to Have

  • Software development experience with Java, Python, Go, or similar languages.

  • Experience with cloud-native or microservices-based applications.

  • Familiarity with Kubernetes or other container orchestration technologies.

  • Experience with modern observability platforms.

  • Knowledge of reliability engineering practices such as SLOs, SLIs, and Error Budgets.

  • Understanding of security, compliance, disaster recovery, and business continuity best practices.

  • AWS certifications.

We offer:

  • Participation in interesting and demanding projects.

  • Flexible working hours.

  • A great, non-corporate atmosphere.

  • Possibility to work remote or hybrid (2 days per week from the office).

  • Opportunities for development and promotion.

  • Attractive package of benefits.

We reserve the right to contact the selected candidates.

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available