Senior Software Engineer, Site Reliability Engineering
Summary
Senior SRE on Google Cloud's Technical Infrastructure team in Warsaw, responsible for the full lifecycle of large-scale distributed systems—designing, deploying, monitoring, and automating to ensure reliability and performance at Google scale.
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance.
Poland: zł364000 - zł373000 (PLN) + 15% bonus target + equity + benefits
Learn more about benefits at Google.
- Engage in and improve the whole lifecycle of services, from inception and design, through to deployment, operation and refinement.
- Support services before they go live through activities such as system design consulting, developing software platforms and frameworks, capacity planning and launch reviews.
- Maintain services once they are live by measuring and monitoring availability, latency and overall system health.
- Scale systems sustainably through mechanisms like automation, and evolve systems by pushing for changes that improve reliability and velocity.
- Practice sustainable incident response and blameless postmortems.
Minimum qualifications:
- Bachelor’s degree in Computer Science, a related field, or equivalent practical experience.
- 5 years of experience with software development in one or more programming languages.
- 5 years of experience with data structures or algorithms.
- 3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems, and 2 years of experience leading projects and providing technical leadership.
Preferred qualifications:
- Master's degree in Computer Science or Engineering, or a related field.