Cloud Platform Engineer - Traffic Infrastructure

United States Digital Space LLC

Singapore

On-site

SGD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ByteDance is seeking a senior platform engineer to join the Global Traffic Infrastructure team. You will help build and scale the edge cloud and global traffic platform powering TikTok's services, focusing on edge resource orchestration, traffic scheduling, cloud governance, and platform automation.

The role emphasizes high scalability, cost efficiency, and global service stability for billions of users, with responsibilities spanning architecture, resource operations, and incident management.

Qualifications

  • Bachelor's degree or above in Computer Science, Software Engineering, Networking, Communications, or related fields, with minimum 3 years of hands-on R&D experience in cloud platform, cloud-native, or edge infrastructure domains.
  • Solid computer science fundamentals, with deep mastery of Linux system internals, process/memory/network models, and core networking principles including TCP/IP, virtual networking, and load balancing.
  • Proficiency in at least one of Go/C++/Python, with proven experience in large-scale backend or infrastructure development.
  • In-depth expertise in Kubernetes core architecture and component principles, with a track record of large-scale cluster optimization.
  • Familiarity with mainstream virtualization and container technologies, including working knowledge of KVM/QEMU, Virtio, container isolation, and image management underlying mechanisms.
  • Hands-on experience with cloud-native observability stacks; practical exposure to Prometheus, Grafana, logging pipelines, and distributed tracing systems is highly preferred.
  • Strong architectural thinking and incident root-cause analysis capabilities, with the ability to independently own complex module design, technical problem-solving, and cross-team collaboration.

Responsibilities

  • Lead core R&D and architecture iteration of the company's traffic infrastructure platform, covering underlying modules including business governance, orchestration & scheduling, and resource management, to support unified management and elastic delivery of global cloud resources.
  • Build a full-lifecycle resource operations management platform spanning resource planning, provisioning, sales, runtime, and billing. Continuously optimize resource utilization and delivery efficiency, drive cost optimization, and ensure business resource competitiveness and customer cost-effectiveness.
  • Refine core business pipelines, polish domain models and engineering workflows, and design secure, robust systems aligned with business requirements and architectural characteristics.
  • Own on-call incident troubleshooting, performance bottleneck resolution, and capacity planning to guarantee year-round high availability of the cloud platform and stable iteration of core business workloads.

Skills

Go
C++
Python
Kubernetes
Linux fundamentals
Networking (TCP/IP)
Cloud-native

Education

Bachelor's degree or higher in CS/Engineering/related

Tools

KVM/QEMU
Virtio
Prometheus
Grafana

Job description

Location:

Singapore

Team:

Infrastructure

Employment Type:

Regular

Job Code:

A185720

Responsibilities

About the TeamThe Traffic Infrastructure team leverages unified platform capabilities to manage global edge infrastructure (China & Non-China), both self-built and third-party, providing standardized, compliant, scalable, and cost-effective traffic infrastructure capabilities for edge services. Our vision is to build a global edge traffic infrastructure platform and become the long-term cornerstone of the company's global edge business in terms of scale, performance, and cost. The Platform Engineering team is a comprehensive business building team. We focus on the operational efficiency of core workflows within the GTI business system, continuously evaluating the system's operational performance from perspectives including business modeling and collaboration processes. By leveraging modern engineering practices, we drive the GTI business system to operate automatically, efficiently, and stably, thereby enhancing customer experience.

Our core focus areas include:

  • 1. Customer Requirements & Product Delivery : Collaborate with product teams to optimize user experience and fulfill customer functional and resource requirements; deliver product capabilities that meet customer expectations, optimize product inventory turnover efficiency, and drive efficient resource delivery.
  • 2. Resource Operations & Cost Optimization : Collaborate with the operations system to improve resource pool management efficiency, advance system automation, platformization, and datafication; monitor resource input-output efficiency to support continuous improvement of operational capabilities.

Job SummaryJoin the company's Global Traffic Infrastructure (GTI) team to build and scale the edge cloud and global traffic platform that powers TikTok's worldwide service delivery. You will focus on edge resource orchestration, traffic scheduling, cloud resource governance, and platform automation, ensuring high scalability, cost efficiency, and global service stability for billions of users.

Responsibilities
  • 1. Lead core R&D and architecture iteration of the company's traffic infrastructure platform, covering underlying modules including business governance, orchestration & scheduling, and resource management, to support unified management and elastic delivery of global cloud resources.
  • 2. Build a full-lifecycle resource operations management platform spanning resource planning, provisioning, sales, runtime, and billing. Continuously optimize resource utilization and delivery efficiency, drive cost optimization, and ensure business resource competitiveness and customer cost-effectiveness.
  • 3. Refine core business pipelines, polish domain models and engineering workflows, and design secure, robust systems aligned with business requirements and architectural characteristics.
  • 4. Own on-call incident troubleshooting, performance bottleneck resolution, and capacity planning to guarantee year-round high availability of the cloud platform and stable iteration of core business workloads.
Qualifications
Minimum Qualifications
  • 1. Bachelor's degree or above in Computer Science, Software Engineering, Networking, Communications, or related fields, with minimum 3 years of hands-on R&D experience in cloud platform, cloud-native, or edge infrastructure domains.
  • 2. Solid computer science fundamentals, with deep mastery of Linux system internals, process/memory/network models, and core networking principles including TCP/IP, virtual networking, and load balancing.
  • 3. Proficiency in at least one of Go/C++/Python, with proven experience in large-scale backend or infrastructure development.
  • 4. In-depth expertise in Kubernetes core architecture and component principles, with a track record of large-scale cluster optimization.
  • 5. Familiarity with mainstream virtualization and container technologies, including working knowledge of KVM/QEMU, Virtio, container isolation, and image management underlying mechanisms.
  • 6. Hands-on experience with cloud-native observability stacks; practical exposure to Prometheus, Grafana, logging pipelines, and distributed tracing systems is highly preferred.
  • 7. Strong architectural thinking and incident root-cause analysis capabilities, with the ability to independently own complex module design, technical problem-solving, and cross-team collaboration.
Preferred Qualifications
  • 1. Prior R&D experience in edge computing, edge cloud, or distributed computing scheduling platforms.
  • 2. Sound understanding of cloud computing technology stack, including eBPF, Cilium, DPDK, and high-performance network forwarding fundamentals.
  • 3. Relevant work experience in multi-cloud/hybrid cloud architecture, OpenStack, or cloud infrastructure cost and billing strategy design.
  • 4. Domain knowledge and architectural understanding of audio/video, CDN, or streaming media systems.
Job Information
About Us

Founded in 2012, the company's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Lemon8, CapCut and Pico as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, the company has made it easier and more fun for people to connect with, consume, and create content.

Why Join the company

Inspiring creativity is at the core of the company's mission. Our innovative products are built to help people authentically express themselves, discover and connect - and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and enrich life - a mission we work towards every day.

As ByteDancers, we strive to do great things with great people. We lead with curiosity, humility, and a desire to make impact in a rapidly growing tech company. By constantly iterating and fostering an "Always Day 1" mindset, we achieve meaningful breakthroughs for ourselves, our Company, and our users. When we create and grow together, the possibilities are limitless. Join us.

Diversity & Inclusion

the company is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At the company, our mission is to inspire creativity and enrich life. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Traffic Infrastructure
Site Reliability Engineer - Traffic Infrastructure

United States Digital Space LLC • Singapore

On-site
SGD 90,000 - 140,000
Site Reliability Engineer - Traffic Infrastructure Technology - Infrastructure Singapore Regular
Site Reliability Engineer - Traffic Infrastructure Technology - Infrastructure Singapore Regular

Bytedance • Singapore

On-site
SGD 120,000 - 180,000
Site Reliability Engineer Graduate (Video and Edge, CDN Platform) - 2027 Start
Site Reliability Engineer Graduate (Video and Edge, CDN Platform) - 2027 Start

United States Digital Space LLC • Singapore

On-site
SGD 120,000 - 240,000
Software Engineer Intern (Traffic Architecture) - 2027 Start
Software Engineer Intern (Traffic Architecture) - 2027 Start

ByteDance • Singapore

On-site
SGD 13,000 - 27,000
Internship program
Backend Software Engineer (Cloud Platform) - Cloud Infrastructure Singapore Regular
Backend Software Engineer (Cloud Platform) - Cloud Infrastructure Singapore Regular

ByteDance • Singapore

On-site
SGD 80,000 - 120,000
Site Reliability Engineer - Traffic Infrastructure
Site Reliability Engineer - Traffic Infrastructure

ByteDance • Singapore

On-site
SGD 90,000 - 150,000
Software Engineer - Service Platform
Software Engineer - Service Platform

ByteDance • Singapore

On-site
SGD 120,000 - 180,000
Cloud Site Relibility Engineer - DCS
Cloud Site Relibility Engineer - DCS

ByteDance • Singapore

On-site
SGD 90,000 - 150,000
Edge POP Engineer, DCS
Edge POP Engineer, DCS

ByteDance • Singapore

On-site
SGD 90,000 - 130,000
Security Engineer (Infrastructure Security SDLC) Graduate (Security BP) - 2027 Start
Security Engineer (Infrastructure Security SDLC) Graduate (Security BP) - 2027 Start

United States Digital Space LLC • Singapore

On-site
SGD 48,000 - 75,000