Tech Lead - Data Infrastructure Site Reliability

ByteDance

Seattle (WA)

On-site

USD 232,560 - 427,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical insurance
Dental and vision insurance
401(k) with company match
Parental leave
Disability coverage
Life insurance
Wellbeing benefits
Paid holidays
Paid sick days
Paid personal time

Job summary

ByteDance in Seattle seeks a Tech Lead for Data Infrastructure Site Reliability to design, build, and operate large-scale cloud infrastructure. You will lead reliability initiatives and collaborate with Advertising, ML, E-commerce, and Core Infra teams.

The role requires 5+ years in SRE/Dev, hands-on with databases, Kubernetes, or big data, and strong systems knowledge. It offers competitive base pay, comprehensive benefits, and opportunities to influence system design.

Qualifications

  • 5+ years of experience in Site Reliability Engineering, Software Development, or related fields.
  • Deep hands-on expertise in databases (SQL/NoSQL), Kubernetes, or Big Data processing and storage.
  • Strong knowledge of system architecture, distributed systems, and performance bottlenecks.
  • Excellent communication and collaboration skills across engineering, product, and data science teams.

Responsibilities

  • Design, develop, and operate large-scale cloud infrastructure and distributed systems.
  • Collaborate with cross-functional teams to drive reliability, performance, and scalability.
  • Lead automation initiatives to reduce toil and improve efficiency.
  • Troubleshoot complex production issues and perform root-cause analysis.
  • Promote best practices in observability, performance optimization, and cost efficiency.
  • Communicate complex technical concepts to both technical and non-technical stakeholders.

Skills

SRE/Dev experience
Distributed systems
Cloud infrastructure
Big Data
Databases
Communication

Tools

Kubernetes

Job description

Tech Lead - Data Infrastructure Site Reliability

Location: Seattle

Team: Technology

Employment Type: Regular

Job Code: A223980A

Responsibilities
  • Strong hands‑on skills in the design, development, and operation of large‑scale cloud infrastructure and distributed systems.
  • Collaborate with cross‑functional teams (Advertising, Machine Learning, E‑commerce, Core Infra) to drive system reliability, performance, and scalability.
  • Lead initiatives to automate operations, eliminate toil, and improve overall system efficiency.
  • Troubleshoot complex production issues, perform root‑cause analysis, and drive long‑term reliability improvements.
  • Promote best practices in system design, observability, performance optimization, and cost efficiency.
  • Communicate complex technical concepts effectively to both technical and non‑technical stakeholders.
Qualifications
  • Minimum Qualifications
    • 5+ years of experience in Site Reliability Engineering, Software Development, or related fields, focusing on designing, building, scaling, and operating cloud‑based systems.
    • Deep hands‑on expertise in at least one of the following: Databases (SQL/NoSQL); Kubernetes or container orchestration; Big Data processing and storage systems (streaming and batch).
    • Strong knowledge of system architecture, distributed systems, and performance bottlenecks.
    • Excellent communication and collaboration skills, with experience working across engineering, product, and data science teams.
  • Preferred Qualifications
    • Proven track record of driving automation, tooling, and process improvements that enhance reliability and efficiency.
    • Experience in cost optimization and performance tuning at scale, backed by data‑driven decision making.
    • Thought leadership in adopting new technologies, improving operational practices, and influencing system design.
Compensation

The base salary range for this position in Seattle is $232,560 – $427,500 annually.

Benefits

Employees have day‑one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short‑term and long‑term disability coverage, life insurance, wellbeing benefits, and additional perks. Employees also receive 10 paid holidays per year, 10 paid sick days per year, and 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure).

Fair Chance and Equal Opportunity

For Los Angeles County (unincorporated) candidates: Qualified applicants with arrest or conviction records will be considered for employment in accordance with all federal, state, and local laws including the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Our company believes that criminal history may have a direct, adverse and negative relationship on the following job duties, potentially resulting in the withdrawal of the conditional offer of employment:

  • Interacting and occasionally having unsupervised contact with internal/external clients and/or colleagues.
  • Appropriately handling and managing confidential information including proprietary and trade secret information and access to information technology systems; and
  • Exercising sound judgment.
Reasonable Accommodation

ByteDance is committed to providing reasonable accommodations in our recruitment processes for candidates with disabilities, pregnancy, sincerely held religious beliefs or other reasons protected by applicable laws. If you need assistance or a reasonable accommodation, please reach out to us at https://tinyurl.com/RA-request.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Tech Lead Cloud Site Reliability Engineer - DCS Cloud
Tech Lead Cloud Site Reliability Engineer - DCS Cloud

Socket.dev • Seattle (WA)

On-site
USD 232,560 - 427,500
Tech Lead Cloud Site Reliability Engineer - DCS Cloud
Tech Lead Cloud Site Reliability Engineer - DCS Cloud

ByteDance • San Jose (CA)

On-site
USD 244,800 - 450,000
Tech Lead - Machine Learning Platform Engineer
Tech Lead - Machine Learning Platform Engineer

ByteDance • San Jose (CA)

On-site
USD 244,800 - 450,000
Medical, dental, and vision insurance
401(k) with company match
Paid parental leave
+2
Tech Lead,Infrastructure Delivery Platform
Tech Lead,Infrastructure Delivery Platform

ByteDance • San Jose (CA)

On-site
USD 212,800 - 387,600
Tech Lead - Data Infrastructure Site Reliability
Tech Lead - Data Infrastructure Site Reliability

ByteDance • San Jose (CA)

On-site
USD 244,800 - 450,000
Health insurance
401(k) with company match
Paid parental leave
+3
Senior Software Engineer, Cloud Infrastructure
Senior Software Engineer, Cloud Infrastructure

ByteDance • San Jose (CA)

On-site
USD 212,800 - 387,600
Medical, dental and vision insurance
401(k) with company match
Parental leave
+6
Site Reliability Engineer - Data (Seattle) Seattle Regular
Site Reliability Engineer - Data (Seattle) Seattle Regular

ByteDance • Seattle (WA)

On-site
USD 177,000 - 342,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+1
Senior Software Development Engineer, SDN-Traffic Intelligence & Control
Senior Software Development Engineer, SDN-Traffic Intelligence & Control

ByteDance • Seattle (WA)

On-site
USD 202,160 - 368,220
Health insurance
401(k) with company match
Paid parental leave
+6
Tech Lead Software Engineer, Programming Language
Tech Lead Software Engineer, Programming Language

ByteDance • San Jose (CA)

On-site
USD 244,800 - 450,000
Medical Insurance
401k Matching
Parental Leave
+6
Senior Site Reliability Engineer - Data Infrastructure (San Jose)
Senior Site Reliability Engineer - Data Infrastructure (San Jose)

ByteDance • San Jose (CA)

On-site
USD 212,800 - 387,600