Team Intro
Responsibilities
Own the service level of a critical, revenue‐generating e‐commerce platform and all supporting infrastructure and services. This role will focus on service reliability, highly‐scalable design and release management in a cloud‐native workplace.
Define service level indicators and data‐driven objectives to uphold and improve uptime, latency and system health of a core TikTok production platform.
Collaborate across teams with engineering and product to ensure that key requirements (such as capacity planning and launch reviews) are performed to enable transparent service delivery to customers.
Implement automation geared towards infrastructure‐as‐code, scalability and service resiliency.
Implement SRE practices around incident management and post‐mortems while being part of on‐call rotations.
Qualifications
Minimum Qualifications
Good understanding of Unix/Linux operating system internals and networking.
Experience writing code in Java, Go, Python or a similar language.
Experience with algorithms, data structures, complexity analysis and software design.
Experience developing tools and APIs to reduce manual interaction with systems and applications using a variety of coding and scripting standards.
Systematic problem‐solving approach, coupled with effective communication skills and a sense of drive.
Preferred Qualifications
Experience running production‐grade web services at scale in a cloud‐native environment.
Experience implementing observability solutions such as monitoring, logging and tracing in complex service meshes.
Expertise in designing, analyzing and troubleshooting large‐scale distributed systems.
#J-*****-Ljbffr
📌 Site Reliability Engineer, Global E-Commerce - Usds (New South Wales)
🏢 TikTok USDS Joint Venture
📍 New South Wales
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.