Senior DevOps Engineer (Platform Reliability / SRE) - Ruby on Rails (Queensland)

Senior DevOps Engineer (Platform Reliability / SRE) - Ruby on Rails (Queensland)

16 Aug
|
Total Dealer
|
Queensland

16 Aug

Total Dealer

Queensland

Senior DevOps Engineer (Platform Reliability / SRE) - Ruby on Rails

Dealer Studio is not your typical automotive tech company. We're building and scaling eCommerce and digital solutions for some of Australia's largest dealership groups and manufacturers, and we're doing it differently. Faster. Smarter.

We're in a genuine growth phase and we're assembling a high-calibre team to take things to the next level. Everyone who joins us on this journey will be well looked after: we're building something big, and we reward the people who help us get there.

About Dealer Studio

Dealer Studio is not your typical automotive tech company. We're building and scaling eCommerce and digital solutions for some of Australia's largest dealership groups and manufacturers, and we're doing it differently. Faster. Smarter.

We're in a genuine growth phase and we're assembling a high-calibre team to take things to the next level. Everyone who joins us on this journey will be well looked after: we're building something big, and we reward the people who help us get there.

The Role

We're hiring a Senior DevOps Engineer to take direct, hands-on ownership of production reliability and platform health for our core Ruby on Rails application. This is a dedicated reliability mandate for someone who's already spent years running production systems, leading incidents, and turning one-off fires into permanent fixes.

You'll work closely with our engineering team, but the buck stops with you on uptime, incident response, and infrastructure health.

What You'll Own (First 6 Months)

- Improving platform uptime and reducing incident response times (MTTR)

- Establishing proactive server health practices across performance, patching, and capacity





- Strengthening incident communication and post-incident follow-through

- Driving root-cause analysis and implementing durable fixes - not workarounds

- Responsibilities

- Own day-to-day production reliability and infrastructure health across our environments

- Detect, diagnose, and resolve platform and database performance issues quickly

- Build and maintain high-signal alerting, logging, dashboards, and operational runbooks

- Lead incident response when needed, including clear technical and stakeholder communication

- Implement preventative maintenance rhythms (upgrades, patching, capacity planning, resilience checks)

- Optimise and maintain CI/CD pipelines using GitHub Actions

- Partner with engineering teams to improve operational standards and reliability outcomes

- Support after-hours escalations for critical production issues (shared rotation)

Must-Have Experience

- 7+ years in software, systems, or infrastructure engineering

- 3+ years in a senior DevOps, SRE, or platform reliability role with clear, named production ownership (not "devops as a side duty")

- Proven experience managing a Ruby on Rails application in production

- Hands-on experience operating production apps on Heroku (or a comparable PaaS): dyno/process sizing, releases, add‑ons, log drains, and day‑to‑day incident diagnosis

- Strong hands‑on experience with AWS and Cloudflare in production environments,



specifically WAF / DNS & CDN / Load Balancing.

- Deep PostgreSQL capability, including query performance diagnosis and practical optimisation

- Proven experience building and improving observability using tools such as New Relic or Grafana

- Strong incident management capability, including calm leadership and transparent communication under pressure

- Track record of driving root‑cause investigations through to permanent fixes

- Demonstrated ability to execute at pace with minimal supervision

Highly Desirable

- Experience planning and executing migrations from Heroku or other PaaS platforms

- Ability to uplift reliability practices and mentor teams through practical standards

- Experience planning and executing a migration off Heroku (or another PaaS) onto AWS or equivalent infrastructure, with clear cutover and rollback thinking

What Success Looks Like Here

- You take ownership without being asked

- You improve production health proactively, not reactively

- You communicate clearly during incidents and follow through afterwards

- You solve the underlying issue and prevent repeat incidents

- You move quickly while maintaining engineering quality

What We Offer

- A competitive salary (based on experience) + super

- Bonus incentives

- Training and development opportunities

- Hybrid flexibility (2 days/week in our Eight Mile Plains office) or remote across Australia

- A genuine growth-stage company solving real problems for Australia's largest dealership groups

Hiring Process

- Initial screening call

- Technical interview focused on real operational scenarios

- Practical reliability/incident exercise

- Final conversation with leadership

#J-18808-Ljbffr

📌 Senior DevOps Engineer (Platform Reliability / SRE) - Ruby on Rails (Queensland)
🏢 Total Dealer
📍 Queensland

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior devops engineer (platform reliability / sre) - ruby on rails (queensland) / queensland

Subscribe to this job alert:

Get the latest job offers by email for: senior devops engineer (platform reliability / sre) - ruby on rails (queensland) / queensland