Site Reliability Engineer

Summary

Maintains and improves the reliability, performance, and uptime of critical applications and infrastructure for a European tech consultancy, using automation, monitoring, and on-call support.

At Mindbox we connect top IT talents with technology projects for leading enterprises across Europe.

We are seeking a skilled Site Reliability Engineer (SRE) to join our global DevOps team. This role is focused on ensuring high availability, performance, and reliability of mission-critical applications and infrastructure while driving automation and operational excellence through SRE best practices.

You will collaborate with engineering and architecture teams to design resilient solutions, implement automation, enforce observability standards, and provide on-call support for global production systems operating 24x7.

Sounds like your kind of challenge?

What you get in return
  • Flexible cooperation model – choose the form that suits you best
    (B2B, employment contract, etc.)
  • Hybrid work setup – 6 days from the office per month
  • Collaborative team culture – work alongside experienced professionals eager to share knowledge
  • Continuous development – access to training platforms and growth opportunities
  • Comprehensive benefits – including Interpolska Health Care, Multisport card, Warta Insurance, and more
  • High quality equipment – laptop and essential software provided
  • 5+ years of experience in Site Reliability Engineering or Production Support roles.
  • Strong background in troubleshooting high-availability systems in complex enterprise environments.
  • Proficiency with automation and CI/CD tools such as Ansible, Jenkins.
  • Solid experience with monitoring and observability tools like Prometheus and Grafana.
  • Strong engineering capability in at least one programming language (e.g., Java, Python, Node.js) and SQL.
  • Comprehensive knowledge of the Software Development Life Cycle (SDLC) and Agile practices.
  • Excellent problem-solving, analytical mindset, and strong communication skills for effective collaboration with globally distributed teams.

Nice to have:

  • Hands-on experience managing large-scale Atlassian Jira and Confluence Data Center instances.
  • Thought leadership in disaster recovery strategies, capacity planning, and performance optimization.
  • Prior involvement in defining operational metrics and automation frameworks for enterprise platforms.

Joining this project you’ll become part of Mindbox – a tech-driven company where consulting, engineering, and talent meet to build meaningful digital solutions. We’ll back you up every step of the way, accelerate your development, and ensure your skills make a difference.