Site Reliability Engineer — Cloud-Scale & Automation
ByteDance
San Jose (CA)
On-site
USD 136,800 - 359,720
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
Short-term and long-term disability coverage
Life insurance
Paid holidays, sick days, and personal time
Job summary
A global technology company in San Jose is seeking an experienced Site Reliability Engineer to enhance service lifecycle management and develop cloud-managed infrastructure. You will leverage your software development skills to improve reliability and efficiency across distributed systems. Ideal candidates have a Bachelor's degree in Computer Science and strong programming experience in languages like C, C++, or Python, alongside familiarity with tools such as Kubernetes and MySQL. This role offers competitive compensation and comprehensive benefits.
Qualifications
At least 3 years of experience programming in one of the specified languages.
Familiar with Unix/Linux system internals, networking, and distributed systems.
Experience in designing and analyzing large-scale distributed systems is preferred.
Responsibilities
Participate in the complete service lifecycle from design to operation.
Design software platforms and monitoring frameworks.
Develop and manage components of cloud-managed data infrastructure.
Establish sustainable mechanisms for scaling systems.
Manage incident responses and conduct blameless postmortems.
Skills
C
C++
Java
Python
Go
Rust
Problem-solving skills
Communication skills
Education
Bachelor's degree in Computer Science or a related technical field
Tools
Kubernetes
Redis
MySQL
Flink
Nginx
Docker
Job description
A global technology company in San Jose is seeking an experienced Site Reliability Engineer to enhance service lifecycle management and develop cloud-managed infrastructure. You will leverage your software development skills to improve reliability and efficiency across distributed systems. Ideal candidates have a Bachelor's degree in Computer Science and strong programming experience in languages like C, C++, or Python, alongside familiarity with tools such as Kubernetes and MySQL. This role offers competitive compensation and comprehensive benefits.