Staff Software Engineer, Platform
Summary
Staff-level platform engineer on Super.com's DevOps/Infrastructure team who owns the internal developer platform, self-serve tooling, and internal AI agent platform, acting as cloud admin for AWS/EKS and mentoring engineers. Hands-on work in Python/TypeScript with Kubernetes, Terraform, Crossplane, and a weekly on-call rotation.
About
We started to help maximize lives–both the lives of our customers and the lives of our employees– so that everyone can experience all that life has to offer. For our employees, our promise is that is more than just a job; it’s an opportunity to unlock one’s potential, where learning is celebrated and impact is realized.
We are more than a fast-paced, high-growth tech company; we care about our people and take career progression seriously. This is your career and our aim is to supercharge it through the people, the work, and the programs that fuel who we are.
About the role
We're looking for a Staff Platform Engineer to join our DevOps/Infrastructure team. This is a transformative, high-impact role spanning technical leadership, hands-on execution, platform development, and organizational enablement. You'll serve as a company-wide subject matter expert for platform engineering, with a primary mandate to define, drive, and operationalize the internal developer platform, tools, and shared AI infrastructure.
You'll own the platform engineering roadmap, lead major cross-functional initiatives, and architect the self-serve tooling. The systems you own will help teams build and deploy faster without sacrificing quality, reliability, or security. Your work will often cut across team boundaries, and you'll be expected to independently identify gaps, set direction, and drive work through to adoption. Your efforts will directly improve the core DORA metrics of , and you'll own how we measure and operationalize them.
In this hands-on role you'll operate as a cloud administrator within our AWS and Kubernetes environments, participate in our incident response rotation, and provide direct support to other engineers over Slack. You'll mentor engineers across the organization and raise the technical bar through guilds, Lunch & Learns, and knowledge-sharing sessions. This is a staff-level individual contributor role with significant autonomy, no people management responsibility, and close partnership with the Engineering Manager of DevOps and the Sr. Director of Infrastructure.
About the team
This DevOps & Infrastructure team is one of the oldest continually operating teams at , and has delivered significant milestones already. Across our teams we support upwards of 1000 production deployments per month, strong SLO performance, and continually improving DORA metrics. Our cloud infrastructure is modern and, thanks to this team's efforts, kept up-to-date. The basic early DevOps challenges such as trunk-based development, IAC adoption, canary releases, and Kubernetes microservice adoption have all been tackled and now we've set our sights on bringing in the newest AI technologies. We use high-quality tools like Datadog, EKS, IncidentIO, and Claude to support engineers and staff across the entire company. As the most senior IC on the team, you'll set the technical direction for how this platform evolves.
What you’ll be working on:
Define and drive the platform engineering roadmap and strategy, taking the initiative, owning projects that scale, and introducing new capabilities to both the Infrastructure org and Engineering as a whole.
Architect, build, and maintain our internal platforms and tools: self-serve infrastructure, paved paths, and AI-first automation. Identify opportunities and roadmap the future of our platform. Automation code to be authored primarily in Python and JS/TS. Specifically, you'll extend our Super CLI, improve our Cloud Development Environments, and expand on our IDP.
Drive the development, maintenance, and roadmap of our internal AI Agent Platform and the successful adoption of internal AI agents across the company.
Measurably improve Developer Experience by actioning findings from DevEx surveys and communicating with engineers directly.
Migrate traditional DevOps operational tasks to automated and agentic flows.
Maintain and extend our cloud infrastructure: Particularly Kubernetes (EKS), AWS, Postgres, Redis, Kafka, and Elasticsearch. Leverage tools like Terraform and Crossplane to keep our infrastructure declarative, and act as the escalation point for the problems other engineers can't solve.
Own the automated collection of platform and DORA metrics and operationalize the data to drive business outcomes. Identify and attribute trends in delivery and reliability data and socialize learnings.
Define infrastructure and platform best practices and ensure they're followed: drive adoption with POCs and a technical case, make build-vs-buy recommendations, and bake tribal knowledge into tooling and docs so the platform, not a single person, is the source of truth
Mentor engineers across the organization – from interns through senior engineers – and level up the broader team through Lunch & Learns, guilds, documentation, and general knowledge-sharing
Help maintain the availability, reliability, and security of our production infrastructure by participating in our on-call rotation, including monitoring, incident response, and resolving issues outside of standard business hours
Note: On-call is a core part of this role. Engineers take turns in a weekly on-call rotation and are compensated for each rotation per the company’s on-call compensation policy.
Our Technology:
Services are authored as K8s-based microservices powered by Python FastAPI backends and Typescript React front-end code.
Apps are built upon Postgres, Redis, Kafka, Elasticsearch, and Snowflake.
Our infrastructure runs on AWS and leverages hosted infrastructure technologies such as EKS, MSK, Elasticache, and RDS. We use Istio for service mesh, Helm, TF and Crossplane for IAC, and Kyverno for policy-as-code enforcement.
Trunk-based development and deploy-on merge through CI/CD is orchestrated with GitLab and ArgoCD, leveraging AI within the pipeline to deliver upwards of 1000 production deployments/month.
We invest heavily in reliability and observability using Datadog for monitoring and automated alerting, and use Amplitude for client-side analytics and experimentation.
Developer productivity and DevX are improved by this team through our CDE, Coder, DX360 surveys, and significant in-house productivity tooling.
As an AI-First engineering organization, we leverage modern AI development workflows using tools such as Cursor, Claude, and internally built AI agents to accelerate implementation and iteration across our codebases. You'll contribute directly to our AI infrastructure in this role.
What we’re looking for:
10+ years of experience across Software Engineering, DevOps, SRE, and/or Platform Engineering roles, at least 3 years at a senior+ level
AI-first thought & technical leadership: Professional experience safely applying AI to infrastructure and platform engineering, opinionated about safe AI usage around high-impact production workloads, and prepared to lead and drive AI infrastructure outcomes for the company
Deep platform engineering expertise: Thinks about developer productivity as a platform problem. Has owned internal developer platforms, self-service capabilities, or infrastructure automation from idea through to production and adoption.
Expert-level cloud and infrastructure skill: AWS, Kubernetes, IaC, and modern CI/CD tooling, plus administratively managing core service infrastructure – servers, databases, caches, message queues, and data networks
Strong full-stack coding, database, and architecture skills: Delivers solutions independently across the entire stack from the ground up
Demonstrated commitment to operational excellence including system reliability, on-call participation, incident response, and continuous improvement practices
Work with extreme initiative & autonomy: Identifying, delivering, measuring, and communicating staff-level initiatives in a high-speed scale-up environment with broad scope and rapid asynchronous remote communications
Staff-level communication and cross-functional influence
We’ve got you covered:
At , we believe in supporting our team so they can thrive—both at work and in life.
Remote-First Flexibility: Work from anywhere in the world and choose the hours that suit you best. We trust you to get great work done on your terms.
Time to Recharge: Enjoy unlimited PTO, company-wide recharge days, and annual team offsites.
Everyday Perks: Weekly UberEats credits and travel discounts on Super.com help you enjoy the little things.
Family-Friendly Benefits: We support growing families with generous parental leave and a flexible return-to-work plan.
Comprehensive Compensation: Competitive salary, equity options, and top-tier benefits starting on day one.
Investing in You: Access to wellness budgets, personal development funds, and team-level learning resources.
At , we are proud to leverage cutting-edge artificial intelligence (AI) technology to make our hiring process smarter, faster, and more inclusive. By integrating AI tools into our recruitment, we enhance our ability to identify top talent efficiently while promoting fairness and consistency for every applicant.
is an equal opportunity employer. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.
Accommodations are available on request for candidates taking part in all aspects of the selection process. If needed, please notify your Talent Acquisition Partner.
Skills
As published by ashby · 17 questions · 3 written answers
Basics
Resume, Name, Email
Short answers (7)
- Phone Number
- Location
- Current Company optional
- Link to Portfolio optional
- Other optional
- What City and Province/State are you located in within Canada/US?
Pick from a list (7)
- Do you currently hold citizenship status, a valid open work permit, open work visa, or permanent residency that allows you to work for any employer in Canada or the US for at least the next 12 months? I.e. PNP Express Entry, F-1 OPT, STEM OPT, EB, etc.
- Do you currently or will you in the future require sponsorship to work in the US or Canada?
- What are your compensation expectations for this role? (Please provide a range)
- How would you describe your gender identity? optional
- Which of these categories do you feel best describes you in terms of race/ethnicity? optional
- How would you describe your sexual orientation? optional
- Do you self-identify as a person with a disability? optional
Written answers (3)
- Describe the most impactful developer tooling, internal platform, or self service infrastructure that you have contributed to. How was its success measured and in what way did you contribute? optional
- What cloud and infrastructure technologies have you personally administered in production? Please include approximate scale. optional
- How are you currently using AI in your engineering workflow? optional
Y Combinator