Site Reliability Engineer, Hybrid Cloud Operation and Delivery - Data Infrastructure

BYTEDANCE PTE. LTD.

Singapore

On-site

SGD 120,000 - 180,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

ByteDance is seeking an experienced Site Reliability Engineer to strengthen our hybrid cloud infrastructure. You will deliver cloud platform services, deploy software, scale resources, and collaborate with R&D to ensure reliable, high‑performance systems.

You will operate cloud environments for internal and external customers, manage alarms, on‑call support, and change processes, while helping evolve high‑availability architecture and disaster recovery.

Qualifications

  • Bachelor's or Master's degree in computer science or a related field.
  • Deep knowledge of Linux, networking, and middleware principles.
  • Proficiency in Shell, Python, Go, or Java; ability to build automation tools.

Responsibilities

  • Deliver cloud platform solutions in hybrid cloud environments and collaborate with R&D to ensure timely project delivery.
  • Operate cloud platform environments for internal and external customers, handle alarms, on-call support, and changes.
  • Contribute to high-availability architecture, disaster recovery, and scalable system improvements.

Skills

Shell scripting
Python
Go
Java
Linux
Networking fundamentals
Distributed systems

Education

Bachelor's / Master's Degree in CS

Tools

Kubernetes (K8s)
Virtual Machines
Load Balancing
Middleware
AI models

Job description

About Us

Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Lemon8, CapCut and Pico as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, ByteDance has made it easier and more fun for people to connect with, consume, and create content.

Why Join ByteDance

Inspiring creativity is at the core of ByteDance's mission. Our innovative products are built to help people authentically express themselves, discover and connect – and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and enrich life - a mission we work towards every day.

Diversity & Inclusion

ByteDance is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At ByteDance, our mission is to inspire creativity and enrich life. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.

Responsibilities
Team Introduction

Our team is responsible for infrastructure systems of hybrid cloud, including products in IaaS/PaaS/SaaS/AI models. We strive to be a leading Site Reliability Engineering (SRE) team in the industry, driving reliability, scalability, and performance at scale.As part of the SRE team, you will tackle complex, large-scale challenges, leveraging your expertise in coding, algorithms, complexity analysis, and distributed system design.We foster a culture of diversity, intellectual curiosity, and open collaboration. Engineers are empowered with strong ownership, autonomy, and the opportunity to work across a wide range of impactful projects.

What you will be doing:
  1. 1. Responsible for delivery products in hybrid cloud scenarios, including cloud platform planning, software deployment, resource expansion, etc. Collaborate with R&D teams to complete project delivery.
  2. 2. Responsible for the operation of cloud platform environments for internal and external customers, including daily alarm handle, on-call support, change, as well as ensuring stability of cloud platform during important event periods.
  3. 3. Participate in stability construction of cloud products with R&D team, and continuously improve capabilities in high availability architecture, disaster recovery, alarm monitoring, etc, based on the experience we get from large-scale systems on site.
  4. 4. Continuously promote the improvement of hybrid cloud serviceability, participate in the standardized SOW of O&M and delivery for new product versions, and build the SRE serviceability acceptance standards to improve implement efficiency.
Qualifications
Minimum Qualifications:
  • - Bachelor's / Master's Degree in Computer Science or related major, with at least 5 years of relevant experience;
  • - Solid basic knowledge of computer software, understanding of Linux operating system, network ,middleware and other related principles.
  • - Familiar with one or more programming languages, such as Shell, Python, Go, or Java. Knowledge of building scripts or tools to handle different problems.
  • - Experience in operation and maintenance of one or more fields, including virtual machines, containers, K8s, load balancing, middleware, AI models, etc.
Preferred Qualifications
  • - Experience in operation and maintenance of IDC equipments such as switches and GPU servers
  • - Working experience in cloud platform related vendors
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Hybrid Cloud SRE: Reliability, Delivery & Scale
Hybrid Cloud SRE: Reliability, Delivery & Scale

BYTEDANCE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Hybrid Cloud SRE: Delivery & Reliability Engineer
Hybrid Cloud SRE: Delivery & Reliability Engineer

ByteDance • Singapore

Hybrid
SGD 80,000 - 120,000
Site Reliability Engineer, System - System Service Global
Site Reliability Engineer, System - System Service Global

ByteDance • Singapore

On-site
SGD 90,000 - 150,000
Site Reliability Engineer, System - System Service Global
Site Reliability Engineer, System - System Service Global

BYTEDANCE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Cloud Site Relibility Engineer - DCS
Cloud Site Relibility Engineer - DCS

ByteDance • Singapore

On-site
SGD 110,000 - 180,000
Site Reliability Engineer - Big Data Computer Platform
Site Reliability Engineer - Big Data Computer Platform

BYTEDANCE PTE. LTD. • Singapore

On-site
SGD 90,000 - 150,000
Tech Lead (SRE) - Cloud Infrastructure
Tech Lead (SRE) - Cloud Infrastructure

BYTEDANCE PTE. LTD. • Singapore

On-site
SGD 180,000 - 240,000
Site Reliability Engineer (Cloud) - Infrastructure Engineering Technology - DevOps Singapore Regular
Site Reliability Engineer (Cloud) - Infrastructure Engineering Technology - DevOps Singapore Regular

ByteDance • Singapore

On-site
SGD 60,000 - 90,000
Cloud Site Relibility Engineer - DCS Singapore Regular
Cloud Site Relibility Engineer - DCS Singapore Regular

ByteDance • Singapore

On-site
SGD 75,000 - 100,000
Backend Software Engineer (SRE) - Cloud Infrastructure Singapore Regular
Backend Software Engineer (SRE) - Cloud Infrastructure Singapore Regular

Bytedance • Singapore

On-site
SGD 110,000 - 180,000