Site Reliability Engineer - Infrastructure

TikTok

Sydney

On-site

AUD 150,000 - 210,000

Full time

38 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

TikTok is seeking a Site Reliability Engineer in Australia to own reliability and efficiency of our core infrastructure. You will tackle scale challenges using coding, algorithms, and large-scale system design, while promoting ownership and mentorship within the team.

The role emphasizes designing automated operations, implementing monitoring frameworks, and ensuring data protection and compliance across systems.

Qualifications

  • Solid basic knowledge of computer software.
  • Understanding of Linux OS, storage IO and related principles.
  • Familiarity with Python, Go, or Java.
  • Knowledge of design patterns and coding principles.

Responsibilities

  • Reliability: ensure core infrastructure reliability and efficiency; set up reliability standards and recovery SOPs.
  • Troubleshooting: locate issues, bottleneck analysis, maintain high availability architecture.
  • Efficiency: build automated operation solutions for large-scale systems; partner with development teams for iteration.
  • Efficiency: design and implement software platforms and monitoring frameworks for SOA governance.
  • Cost: monitor and budget systems to optimize company costs.
  • Compliance: design and implement data protection plans to meet standards.

Skills

Linux fundamentals
Programming (Python, Go, Java)
System reliability
Code, algorithms, and design
Monitoring and SOA governance

Education

Bachelor's / Master's in Computer Science or related

Tools

Kubernetes
Docker/Containers
Redis
MySQL
MongoDB
Kafka

Job description

Responsibilities

The team is responsible for infrastructure systems, including Storage/Computing/DB. We aim to be the leading SRE team across the industry. In the SRE team, you will have the opportunity to manage the complex challenges of scale, while using expertise in coding, algorithms, complexity analysis, and large-scale system design. We embrace a culture of diversity, intellectual curiosity, openness, and problem-solving. We also encourage ownership, self-governance and independence to work on various projects, and an environment that provides the support and mentorship needed to learn and grow as an engineer.

  • Reliability: Ensuring the reliability and efficiency of our core infrastructure, focusing on system capacity and stability; setting up reliability standards and recovery SOP.
  • Troubleshooting and locating technical issues, bottleneck analysis, managing system high availability architecture transformation and upgrading.
  • Efficiency: Building automated operation solutions for large-scale systems; partnering with system development teams for system iteration.
  • Efficiency: Designing and implementing software platforms and monitoring frameworks for efficient, automated, and intelligent service-oriented architecture (SOA) governance.
  • Cost: There are millions of CPUs. We should build delivery standards, and monitor and budget systems to optimize the cost of the company.
  • Compliance: Designing and setting up new IDC; designing and implementing a data protection plan to meet the standard requirement.
Qualifications

Minimum Qualification(s):

  • Solid basic knowledge of computer software
  • Understanding of Linux operating system, storage, network IO and related principles
  • Familiarity with one or more programming languages, such as Python, Go, and Java
  • Knowledge of design patterns and coding principles

Preferred Qualification(s):

  • Bachelor's / Master's Degree in Computer Science or related major
  • At least 3 years of relevant experience
  • Experience with storage systems and technologies such as KV, Table, Graph, Redis, MySQL, MongoDB, MQ, and Kafka
  • Experience with computing & big data systems and technologies such as Kubernetes, Docker/Containers, AIops, Spark, Flink, Function as a service, RPC Framework, and Service Mesh
About TikTok

TikTok is the leading destination for short-form mobile video. At TikTok, our mission is to inspire creativity and bring joy.

TikTok's global headquarters are in Los Angeles and Singapore, and we also have offices in New York City, London, Dublin, Paris, Berlin, Dubai, Jakarta, Seoul, and Tokyo.

Why Join Us

Inspiring creativity is at the core of TikTok's mission. Our innovative product is built to help people authentically express themselves, discover and connect – and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and bring joy - a mission we work towards every day.

We strive to do great things with great people. We lead with curiosity, humility, and a desire to make impact in a rapidly growing tech company. Every challenge is an opportunity to learn and innovate as one team. We're resilient and embrace challenges as they come. By constantly iterating and fostering an "Always Day 1" mindset, we achieve meaningful breakthroughs for ourselves, our company, and our users. When we create and grow together, the possibilities are limitless. Join us.

Diversity & Inclusion

TikTok is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At TikTok, our mission is to inspire creativity and bring joy. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.

Acknowledgment of Country

In the spirit of reconciliation, TikTok acknowledges the Traditional Custodians of country throughout Australia and their connections to land, sea and community. We pay our respect to elders past and present and extend that respect to all Aboriginal and Torres Strait Islander peoples today.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer - Infrastructure
Senior Site Reliability Engineer - Infrastructure

TikTok • Sydney

On-site
AUD 180,000 - 250,000
Site Reliability Engineer, Security Engineering
Site Reliability Engineer, Security Engineering

TikTok • Sydney

On-site
AUD 140,000 - 180,000
Senior Site Reliability Engineer, TikTok Product
Senior Site Reliability Engineer, TikTok Product

TikTok USDS Joint Venture • Sydney

On-site
AUD 150,000 - 190,000
Site Reliability Engineer, TikTok Product
Site Reliability Engineer, TikTok Product

TikTok USDS Joint Venture • Sydney

On-site
AUD 120,000 - 160,000
Site Reliability Engineer — Scale, Automation & Observability
Site Reliability Engineer — Scale, Automation & Observability

TikTok USDS Joint Venture • Sydney

On-site
AUD 120,000 - 160,000
Site Reliability Engineer-Cloud Infrastructure
Site Reliability Engineer-Cloud Infrastructure

TikTok • Sydney

On-site
AUD 100,000 - 130,000
Site Reliability Engineer, Tech Infra - USDS
Site Reliability Engineer, Tech Infra - USDS

TikTok USDS Joint Venture • Sydney

On-site
AUD 120,000 - 160,000
Network Engineer - Network Infrastructure
Network Engineer - Network Infrastructure

TikTok • Sydney

On-site
AUD 150,000 - 210,000
Site Reliability Engineer, Global E-Commerce - USDS
Site Reliability Engineer, Global E-Commerce - USDS

TikTok USDS Joint Venture • Sydney

On-site
AUD 180,000 - 250,000
Network Implementation Engineer - Physical Network Infrastructure
Network Implementation Engineer - Physical Network Infrastructure

TikTok • Council of the City of Sydney

On-site
AUD 160,000 - 210,000