Senior Platform Engineer (Reliability) - Unannounced Project

Open 37d posting dated 7 hours ago

Summary

Senior Platform Engineer (Reliability) to design and maintain scalable, reliable systems for a new multiplayer strategy game, using cloud-native tools like Kubernetes, AWS, and Terraform.

Scopely is looking for a Senior Platform Engineer (Reliability) to join a new truly unique multiplayer strategy game in Spain, Ireland, Portugal or the UK on a remote or hybrid basis. We can support with visa sponsorship and relocation assistance from any location.

At Scopely, we care deeply about what we do and want to inspire play every day — whether in our work environments alongside our talented colleagues or through our deep connections with our communities of players. We are a global team of game lovers who are developing, publishing and innovating the mobile games industry, connecting millions of people around the world daily.

We are in the early stages of development on an ambitious, unannounced Strategy/MMO title, creating a team of talented and passionate game makers to join us on this exciting journey!

What You'll Do

  • Messaging Systems Ownership: Operate, monitor, and continuously improve our messaging infrastructure — with a current focus on NATS cluster and NATS JetStream — ensuring it is reliable, observable, and well-understood by the teams that depend on it.
  • Signals & Observability: Design and own the observability layer for distributed backend systems, defining the signals (metrics, traces, logs) that make operational problems visible and actionable — and pushing for their adoption where they don't yet exist.
  • Cross-functional Diagnosis: Sit at the intersection of infrastructure and backend engineering: correlate infrastructure signals (IOPS, latency, resource saturation) with application behaviour (message throughput, consumer lag, retry storms) to diagnose root causes and guide the right teams toward the right fixes.
  • SLOs & Error Budgets: Define, implement, and maintain SLO related frameworks for backend services and messaging pipelines, making reliability measurable and helping teams make informed trade-offs between velocity and stability.
  • Reliability as Internal Product: Build and maintain reliability tooling, runbooks, and operational frameworks as internal products — enabling backend and infrastructure engineers to self-serve on operational concerns rather than creating dependency on SRE. Partner with backend, infrastructure, and product engineers to shape reliability standards, share operational context, and influence architecture decisions before they become production problems.
  • Incident Management: Lead or contribute to incident response across the messaging and backend layers, drive postmortems to systemic fixes, and embed preventative improvements into engineering workflows.
  • Code Literacy & Engineering Collaboration: Navigate the infrastructure and backend codebases confidently; identify poorly instrumented services, inadequate infrastructure architecture, missing error handling, or patterns that create operational risk, contributing and improving in close collaboration with other engineering teams.

What We're Looking For

  • Strong background in Site Reliability Engineering, production operations, or backend engineering with a significant operational focus
  • Hands-on experience operating NATS and/or NATS JetStream in production — or equivalent deep experience with distributed messaging systems such as Apache Kafka, AWS Kinesis, or similar. Regardless of the system, experience with Leader Election, RAFT consensus, log replication, are essential.
  • Ability to navigate and reason about application codebases (e.g. C#, Go, Python, or similar) while being capable of identifying instrumentation gaps, operational anti-patterns, and code-level root causes
  • Strong observability experience: designing and implementing metrics, logs, and traces strategies across distributed systems
  • Experience debugging complex distributed systems, particularly across the boundary between infrastructure and application layers
  • Solid understanding of cloud infrastructure (AWS preferred) and containerized workloads (ECS, Kubernetes/EKS, or equivalent)
  • Strong communication skills — able to translate between infrastructure and application engineers and articulate operational risk to non-technical stakeholders
  • Infrastructure as Code First - You believe in managing infrastructure, SLOs, monitors, dashboards, and others, primarily through code rather than manual configuration. You understand the benefits of IaC and advocate for them.

Bonus Points

  • Direct experience with NATS JetStream specifically — stream configuration, consumer groups, retention policies, delivery guarantees
  • Experience with Datadog or equivalent for observability at scale
  • Exposure to database-level operational concerns (query performance, connection pooling, replication lag), as well as exposure to data modeling.
  • Experience mentoring engineers or driving reliability culture across teams

About Scopely

Scopely is a leading video game and global interactive entertainment company, home to many of the world’s most beloved and enduring experiences, including two of the most successful mobile games of all-time “MONOPOLY GO!” and “Pokémon GO,” along with “Stumble Guys,” “Star Trek™ Fleet Command,” “MARVEL Strike Force,” “WWE Champions,” the Scrabble® franchise, “Yahtzee® With Buddies,” and many others. Across mobile, web, PC, and console, Scopely creates, develops, publishes, and live-operates one of the most diversified and award-winning portfolios in the games industry — bringing hundreds of millions of players together through a shared love of play.

Founded in 2011, Scopely is powered by its exceptional team — including thousands of world-class gamemakers around the globe, a distinctive tenet-driven culture, and its proprietary technology platform, Playgami. Together, these strengths have fueled Scopely’s position as the #1 mobile games company in the U.S. and #2 globally, generating more than $10 billion in lifetime revenue. Whether building global sensations like “MONOPOLY GO!” from the ground up, or expanding through strategic acquisitions, including the FoxNext, GSN, and Scopely Explore games businesses — Scopely consistently delivers experiences players love today and return to for years to come.

Recognized multiple times as one of the "100 Most Influential Companies in the World" by TIME magazine and one of Fast Company's "World's Most Innovative Companies" and “Best Workplaces for Innovators,” Scopely believes that video games can be a force for good — creating meaningful connections, vibrant communities, and making life better through play.

Scopely has global operations and partners across four continents in more than a dozen countries worldwide. For more information, visit: https://www.scopely.com/.

Notice to Candidates: Scopely will never request payment or financial information during the application or hiring process. Please apply only through our official website and verify that all Talent Partner communications come from an email address ending in @scopely.com.

Should you have any questions or encounter any fraudulent requests/emails/websites, please immediately contact recruiting@scopely.com. Our job applicant privacy policies are available here: California Privacy Notice and EEA/UK Privacy Notice.

Employment at Scopely is based solely on a person's merit and qualifications. Scopely does not discriminate against any employee or applicant because of race, creed, color, religion, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition (including breastfeeding), or any other basis protected by law. We also consider qualified applicants with arrest or conviction records, consistent with applicable federal, state and local law.