11 Sep
|
Attekus
|
Brisbane
Site Reliability Engineer
? Flexible Location | Australia ? Full time | Remote
Build reliability into everything. Keep technology performing when it matters most.
Attekus is helping organisations across Australia and New Zealand create better experiences for communities through technology.
Our enterprise solutions support local government, government and education organisations to manage bookings, venues, events, appointments and customer experiences more effectively.
As we continue to grow, we’re looking for a Site Reliability Engineer to help mature our SaaS software platform.
We are someone who loves solving complex infrastructure and reliability challenges, building robust engineering patterns and can participating in a fast-paced software engineering team.
You’ll take ownership of how we think about platform reliability, observability, incident management, cloud infrastructure and operational excellence across our Azure-hosted product suite, while working closely with our development teams to continually improve how we build and operate our technology.
What you’ll be doing
Platform operations
- Support Engineering leadership in defining and continually improving SLOs, SLIs and error budgets across Attekus products
- Participate in incident response, on-call rotation and post-incident reviews
- Contribute to the building of practical runbooks and processes that improve platform resilience and reduce recovery time
- Work with development teams to balance delivery velocity with reliability
Own and evolve our Azure environment
- Spearhead azure infrastructure utilisation across development, staging and production
- Optimise cloud performance, security, scalability and cost
- Drive infrastructure-as-code maturity using Bicep
- Ensure infrastructure is version-controlled and deployed through automated pipelines
Make observability meaningful
- Build and evolve monitoring across application performance, infrastructure health, logging and distributed tracing
- Develop proactive alerting that surfaces the signals that matter without creating unnecessary noise
- Help teams identify and resolve issues before they affect customers
Improve how we deploy
- Own and evolve CI/CD practices using GitHub Actions
- Build safe, repeatable deployment processes with appropriate controls, rollback capability and environment promotion
- Partner with development teams on capacity planning, performance testing and deployment risk
- Build SRE capability across the wider engineering organisation
- Share knowledge and lift practices across incident management, observability, infrastructure and operational excellence
Strengthen security and resilience
- Embed security-first thinking into infrastructure, access management and deployment practices
- Partner across the business to support the compliance expectations of our government and education customers
- Improve disaster recovery and business continuity processes
- Ensure backup, failover and recovery procedures are documented, tested and ready when needed
What we’re looking for
You’ll likely bring
- 8 years’ experience across software engineering, infrastructure, cloud engineering or site reliability
- Demonstrated experience working as a senior contributor in a a small engineering or operations team
- Deep hands-on Microsoft Azure experience across IaaS and PaaS
- Strong understanding of SRE principles including SLOs, SLIs, error budgets, toil reduction and incident management
- Strong infrastructure-as-code experience using Bicep, ARM, Terraform or similar
- CI/CD experience, ideally using GitHub Actions
- Experience implementing observability and monitoring solutions
- Strong communication skills and the ability to work effectively across technical teams
Bonus points if you have
- Strong .NET/C# experience and familiarity with TypeScript
- Experience with Azure Monitor, Application Insights, Grafana, Datadog or similar
- Experience supporting SaaS or enterprise software environments
- Experience working within government, education or similarly regulated environments
- Azure Solutions Architect, Azure DevOps Engineer Expert or related certifications
Why Attekus?
At Attekus, culture is not a poster on a wall. It shapes how we work and grow together.
We value people who are:
Confident - We back ourselves and what we deliver
Outgoing - We connect and bring energy to every interaction
Obsessive - We care deeply about outcomes and the details that matter
Our purpose is simple: helping communities thrive through better experiences.
Experience Starts Here.
If you’re excited by building reliable platforms, developing people and shaping how a growing engineering organisation operates, we’d love to hear from you.
Please note: Applicants must hold permanent working rights in Australia or New Zealand at the time of application.
No recruiters please.
📌 Site Reliability Engineer - Remote (Brisbane)
🏢 Attekus
📍 Brisbane