01 Sep
|
Tribus
|
Australia
Site Reliability Engineer – Data Infrastructure
Sydney | Quantitative Trading | Linux, Kafka, Python
We are working with a leading global quantitative trading firm that is growing its Data Engineering team in Sydney.
This is an SRE role focused on the infrastructure that powers large-scale data platforms. You will help operate a multi-petabyte workplace supporting millions of queries each day, working across Linux, distributed data systems, automation and production reliability.
This is not a traditional data engineering role focused on writing pipelines, DAGs or analytics workloads. The team owns and operates the underlying platforms themselves.
What you'll be doing
Operate and improve large-scale distributed data infrastructure
Work with technologies including Kafka, HDFS, Kubernetes and distributed query platforms
Own monitoring, alerting, incident response and production reliability
Automate infrastructure deployment, upgrades and operational processes using Python
Troubleshoot Linux, networking, storage and distributed systems issues
Improve capacity, resilience,
failure handling and deployment processes
Work directly with traders, researchers and developers to solve data infrastructure problems
Participate in an on-call rotation and engineer out recurring issues
What we're looking for
Hands-on Linux systems administration and troubleshooting experience
Experience owning production systems, including monitoring, incidents and on-call
Operator-side experience with at least one of Kafka, HDFS or Kubernetes
Python experience, ideally for infrastructure automation or operational tooling
An understanding of networking, storage, processes, memory and system performance
Experience with infrastructure automation, CI/CD or configuration management
Experience with Kafka or HDFS administration is particularly valuable, but you do not need to know the entire technology stack.
Engineers coming from SRE, infrastructure, platform engineering, systems engineering or production engineering backgrounds are encouraged to apply.
📌 Site Reliability Engineer Data Infrastructure Sydney (Australia)
🏢 Tribus
📍 Australia