Senior DevOps Engineer (Platform Reliability / SRE) - Ruby on Rails
Summary
Own production reliability for a Ruby on Rails e-commerce platform serving Australia’s largest dealerships, focusing on uptime, incident response, and infrastructure health on Heroku, AWS, and Cloudflare.
About Dealer Studio
Dealer Studio is not your typical automotive tech company. We're building and scaling eCommerce and digital solutions for some of Australia's largest dealership groups and manufacturers, and we're doing it differently. Faster. Smarter.
We're in a genuine growth phase and we're assembling a high-calibre team to take things to the next level. Everyone who joins us on this journey will be well looked after: we're building something big, and we reward the people who help us get there.
The Role
We're hiring a Senior DevOps Engineer to take direct, hands-on ownership of production reliability and platform health for our core Ruby on Rails application. This is a dedicated reliability mandate for someone who's already spent years running production systems, leading incidents, and turning one-off fires into permanent fixes.
You'll work closely with our engineering team, but the buck stops with you on uptime, incident response, and infrastructure health.
What You'll Own (First 6 Months)
Improving platform uptime and reducing incident response times (MTTR)
Establishing proactive server health practices across performance, patching, and capacity
Strengthening incident communication and post-incident follow-through
Driving root-cause analysis and implementing durable fixes - not workarounds
Responsibilities
Own day-to-day production reliability and infrastructure health across our environments
Detect, diagnose, and resolve platform and database performance issues quickly
Build and maintain high-signal alerting, logging, dashboards, and operational runbooks
Lead incident response when needed, including clear technical and stakeholder communication
Implement preventative maintenance rhythms (upgrades, patching, capacity planning, resilience checks)
Optimise and maintain CI/CD pipelines using GitHub Actions
Partner with engineering teams to improve operational standards and reliability outcomes
Support after-hours escalations for critical production issues (shared rotation)
Must-Have Experience
7+ years in software, systems, or infrastructure engineering
3+ years in a senior DevOps, SRE, or platform reliability role with clear, named production ownership (not "devops as a side duty")
Proven experience managing a Ruby on Rails application in production
Hands-on experience operating production apps on Heroku (or a comparable PaaS): dyno/process sizing, releases, add-ons, log drains, and day-to-day incident diagnosis
Strong hands-on experience with AWS and Cloudflare in production environments, specifically WAF / DNS & CDN / Load Balancing.
Deep PostgreSQL capability, including query performance diagnosis and practical optimisation
Proven experience building and improving observability using tools such as New Relic or Grafana
Strong incident management capability, including calm leadership and clear communication under pressure
Track record of driving root-cause investigations through to permanent fixes
Demonstrated ability to execute at pace with minimal supervision
Highly Desirable
Experience planning and executing migrations from Heroku or other PaaS platforms
Ability to uplift reliability practices and mentor teams through practical standards
Experience planning and executing a migration off Heroku (or another PaaS) onto AWS or equivalent infrastructure, with clear cutover and rollback thinking
What Success Looks Like Here
You take ownership without being asked
You improve production health proactively, not reactively
You communicate clearly during incidents and follow through afterwards
You solve the underlying issue and prevent repeat incidents
You move quickly while maintaining engineering quality
What We Offer
A competitive salary (based on experience) + super
Bonus incentives
Training and development opportunities
Hybrid flexibility (2 days/week in our Eight Mile Plains office) or remote across Australia
A genuine growth-stage company solving real problems for Australia's largest dealership groups
Hiring Process
Initial screening call
Technical interview focused on real operational scenarios
Practical reliability/incident exercise
Final conversation with leadership
Interested in joining the team? Apply now!
Please reach out to Darci if you would like to discuss this position or other opportunities with Dealer Studio - [email protected].