Principal Software Engineer
Summary
Principal Software Engineer at Riot Games designing internal platforms to integrate AI into daily workflows, focusing on reliability, observability, and developer productivity.
Principal Software Engineer at Riot Games.
About the role The AI Efficiency team at Riot Games is dedicated to developing the platforms and technical infrastructure that enable Rioters to seamlessly incorporate AI into their daily operations. As a Principal Software Engineer, your primary responsibility will be to design and enhance internal systems, automation processes, and safety measures that boost the scalability and reliability of our AI services. You will play a crucial role in guiding the company's approach to adopting AI-native engineering tools while ensuring that our internal developer platforms remain robust and efficient.
Key facts
Location: Singapore
Engagement: Full-time
Job ID: REQ-0010096
Team: AI Efficiency
What you’ll do
- Design and oversee the development of internal platform capabilities that facilitate the creation, deployment, and management of AI services.
- Develop self-service workflows and standardized processes that enhance developer productivity while maintaining security and governance standards.
- Improve the reliability of the platform by enhancing observability, monitoring systems, incident response protocols, and deployment safety measures.
- Define and monitor service health metrics, Service Level Indicators (SLIs), and Service Level Objectives (SLOs) to ensure a balance between development speed and system stability.
- Create automation tools to minimize operational burdens and enhance recovery times during incidents.
- Collaborate closely with engineering teams to integrate production readiness and maintainability throughout the system development lifecycle.
- Diagnose and resolve issues across various AI model-serving frameworks, runtime environments, and GPU hardware setups.
- Incorporate AI-assisted tools for code reviews, bug detection, and automated test generation into existing Continuous Integration/Continuous Deployment (CI/CD) workflows.
- Establish governance and reporting frameworks that ensure AI-driven automation is both trustworthy and effective.
Requirements
- A Bachelor's degree in Computer Science or a related field, or equivalent professional experience.
- A minimum of 5 years of experience in Platform Engineering, Infrastructure, Site Reliability Engineering, DevOps, or Developer Experience.
- Strong programming skills in languages such as Python, Go, or JavaScript/TypeScript.
- Proven experience in building or managing internal platforms, developer tools, or shared infrastructure.
- Familiarity with managing cloud-based production environments, including AWS, GCP, or Azure.
- Knowledge of observability practices, including metrics, logs, traces, and alerting mechanisms.
- Experience with container orchestration technologies like Kubernetes or ECS.
- Strong ability to influence technical direction and communicate effectively with stakeholders across different functions.
Nice to have
- Experience working with AI/ML platforms, inference services, or GPU-accelerated workloads.
- Background in defining Service Level Objectives (SLOs) and managing error budgets.
- Experience in platform product management, particularly in designing self-service solutions.
- Familiarity with Infrastructure as Code (IaC) tools such as Terraform or Pulumi.
- Knowledge of security practices, access control, and secrets management in production environments.
- Experience with browser automation frameworks like Playwright for UI or accessibility testing.
- A track record of mentoring engineers and establishing technical standards within teams.
Skills & tools
- Programming Languages: Python, Go, JavaScript, TypeScript
- Cloud Infrastructure: AWS, GCP, Azure, Kubernetes, ECS
- Observability Tools: Metrics, logs, traces, dashboards
- Automation Tools: CI/CD systems, Terraform, Pulumi, Playwright
Practical notes
- full relocation assistance
- extensive health insurance for employees and their families
- retirement plans with matching contributions
- open paid time off policy
- life insurance
- parental leave
- disability coverage
- a Play Fund for gaming activities
- a donation matching program to support charitable contributions
This position is an excellent opportunity for a seasoned software engineer looking to make a significant impact in a dynamic and innovative environment.