- Own the service level of a critical, revenue‑generating e‑commerce platform and all supporting infrastructure and services. This role will focus on service reliability, highly‑scalable design and release management in a cloud‑native workplace.
- Define service level indicators and data‑driven objectives to uphold and improve uptime, latency and system health of a core TikTok production platform.
- Collaborate across teams with engineering and product to ensure that key requirements (such as capacity planning and launch reviews) are performed to enable transparent service delivery to customers.
- Implement automation geared towards infrastructure‑as‑code, scalability and service resiliency.
- Implement SRE practices around incident management and post‑mortems while being part of on‑call rotations.
Qualifications
Minimum Qualifications
- Good understanding of Unix/Linux operating system internals and networking.
- Experience writing code in Java, Go, Python or a similar language.
- Experience with algorithms, data structures, complexity analysis and software design.
- Experience developing tools and APIs to reduce manual interaction with systems and applications using a variety of coding and scripting standards.
- Systematic problem‑solving approach, coupled with effective communication skills and a sense of drive.
Preferred Qualifications
- Experience running production‑grade web services at scale in a cloud‑native environment.
- Experience implementing observability solutions such as monitoring, logging and tracing in complex service meshes.
- Expertise in designing, analyzing and troubleshooting large‑scale distributed systems.
#J-18808-Ljbffr
📌 Site Reliability Engineer, Global E-Commerce - USDS (City of Sydney)
🏢 TikTok USDS Joint Venture
📍 City of Sydney
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.