Infrastructure Engineer with Programming Skills
Role Summary
We're looking for a strong infrastructure engineer to join the Performance Optimization Squad within the Platform mission. You'll work alongside senior engineers and a data scientist to execute on fleet-wide cost and performance optimization initiatives — building automation, implementing changes across production infrastructure, and helping the squad scale its impact across thousands of services.
This is a hands-on execution role. You'll take well-scoped optimization initiatives like rightsizing compute resources, JVM tuning, improving workload placement and drive them to completion across Spotify's fleet.
Responsibilities
- Drive initiatives to optimize infrastructure costs and performance.
- Implement automation to enhance efficiency and scalability.
- Collaborate with team members to identify and execute on optimization opportunities.
- Conduct performance tuning and reliability engineering tasks.
- Monitor and analyze cloud costs across services.
- Engage in incident response and troubleshooting as needed.
Key Requirements
- 3–5 years of experience in infrastructure, platform, or backend engineering roles.
- Solid experience with Kubernetes (ideally GKE).
- Proficient in at least two of: Java, Go, Python, with a preference for strong scripting and automation skills.
- Comfortable with Google Cloud Platform; compute, networking, IAM, and cost monitoring.
- Familiarity with IaC tools (Terraform, Helm) and CI/CD pipelines.
- Solid understanding of reliability engineering, performance tuning, and incident response.
Nice to Have
- Familiarity with cloud cost monitoring tools (GCP Billing, BigQuery cost exports).
- Background in reliability engineering or SRE; understanding of SLOs, error budgets, and safe rollout practices.
- Experience with JVM-based services at scale.
Other Details
Start Date: 2026-07-13 to End Date: 2027-01-12.
Workplace: Stockholm, Sweden.