Senior SRE - Compute: Scale, Automation & Reliability

TikTok USDS Joint Venture

Seattle (WA)

On-site

USD 178,000 - 342,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, vision insurance
401(k) with company match
Paid parental leave
Disability coverage

Job summary

TikTok USDS Joint Venture LLC is seeking a Site Reliability Engineer to help build and run large-scale, fault-tolerant systems. You will automate operations, partner with software teams, and drive reliability across services.

The role emphasizes on-call rotations, performance tuning, and scalable infrastructure to support massive web traffic. You will work with modern stacks (Go/Python) on Linux, collaborate across teams, and participate in defining reliability targets (SLOs/SLIs/SLAs).

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, or related field with 3+ years of experience.
  • Proven work experience as an SRE or systems engineer with automation focus.
  • Experience with Go/Python and large-scale distributed systems.
  • Strong Linux and open-source tech understanding.
  • Excellent problem-solving and cross-functional collaboration skills.

Responsibilities

  • Develop and maintain automation procedures to maximize system efficiency and minimize human intervention.
  • Work with software engineering teams to design, deploy and operate robust system components.
  • Ensure system scalability to handle web traffic and data growth.
  • Implement monitoring tools and set up metrics for system health and performance.
  • Participate in on-call rotations, incident management, and postmortems.
  • Conduct performance tests to identify bottlenecks.
  • Define and monitor SLOs, SLIs, and SLAs.
  • Foster blameless postmortems and sustainable user support.

Skills

Go
Python
Distributed systems
Linux
Networking
Problem solving
Communication

Education

Bachelor's degree in CS/IT or related field

Tools

Docker
Kubernetes
Prometheus
Grafana

Job description

TikTok USDS Joint Venture LLC is seeking a Site Reliability Engineer to help build and run large-scale, fault-tolerant systems. You will automate operations, partner with software teams, and drive reliability across services.

The role emphasizes on-call rotations, performance tuning, and scalable infrastructure to support massive web traffic. You will work with modern stacks (Go/Python) on Linux, collaborate across teams, and participate in defining reliability targets (SLOs/SLIs/SLAs).

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE: Scale & Automation for Global Infra
SRE: Scale & Automation for Global Infra

TikTok USDS Joint Venture • Seattle (WA)

On-site
USD 130,000 - 246,000
Senior SRE: Build Highly Available, Resilient Systems
Senior SRE: Build Highly Available, Resilient Systems

TikTok USDS Joint Venture • Seattle (WA)

On-site
USD 178,000 - 342,000
Senior Site Reliability Engineer – Scale & Automation
Senior Site Reliability Engineer – Scale & Automation

TikTok USDS Joint Venture • San Jose (CA)

On-site
USD 187,000 - 438,000
Medical insurance
Dental insurance
Vision insurance
+7
Senior Site Reliability Engineer Scalable Automated Systems
Senior Site Reliability Engineer Scalable Automated Systems

TikTok USDS Joint Venture • San Jose (CA)

On-site
USD 137,000 - 259,000
Senior SRE, AI Infrastructure & Automation
Senior SRE, AI Infrastructure & Automation

TikTok USDS Joint Venture • Seattle (WA)

On-site
USD 178,000 - 342,000
Senior SRE: Global Resilience & Automation
Senior SRE: Global Resilience & Automation

TikTok USDS Joint Venture • San Jose (CA)

On-site
USD 187,000 - 360,000
Medical, dental, and vision insurance
401(k) with company match
Paid parental leave
SRE for AI Infra & Global System Reliability
SRE for AI Infra & Global System Reliability

TikTok USDS Joint Venture • San Jose (CA)

On-site
USD 123,000 - 259,000
Site Reliability Engineer: Scale, Automation
Site Reliability Engineer: Scale, Automation

TikTok USDS Joint Venture • Seattle (WA)

On-site
USD 130,000 - 246,000
Video Platform SRE: Scale Reliability & Automation
Video Platform SRE: Scale Reliability & Automation

TikTok USDS Joint Venture • Seattle (WA)

On-site
USD 112,725 - 177,840
SRE Tech Lead Manager — Observability, Reliability & Scale
SRE Tech Lead Manager — Observability, Reliability & Scale

TikTok USDS Joint Venture • San Jose (CA)

On-site
USD 208,000 - 438,000