Backend Inference Framework Engineer Graduate (AML Inference) - 2027 Start

BYTEDANCE PTE. LTD.

Singapore

On-site

SGD 120,000 - 180,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ByteDance is seeking a skilled engineer to architect and implement a scalable model inference system for large-parameter AI models. You will address challenges across deployment scenarios, optimize inference framework core modules, and drive improvements in performance, latency, and resource utilization.

The role requires handling high-concurrency distributed services, staying current with inference tech, and collaborating across teams to deliver system-wide improvements.

Qualifications

  • Bachelor's or Master's in computing or related field.
  • Solid C/C++ programming skills and knowledge of data structures and algorithms.
  • Familiar with multi-threaded concurrency and basic performance tuning in multi-threaded scenarios.
  • Experience in R&D of high-concurrency distributed services and awareness of latency/resource optimization.
  • Good learning, execution, and cross-team collaboration skills.

Responsibilities

  • Design and implement architecture for high-performance model inference services.
  • Optimize core modules of the inference framework (scheduling, monitoring, canary release).
  • Track and adopt new inference technologies and promote standardization of the team’s tech system.

Skills

C/C++ programming
Linux
Concurrency
Performance tuning
Distributed services

Education

Bachelor's or Master's in computing

Tools

Redis
RocksDB
BRPC
GRPC

Job description

Team Introduction

Data AML is ByteDance's Machine Learning mid-platform, providing training and inference systems for recommendation/advertising for businesses such as Douyin, Jinri Toutiao, and Xigua Video. It provides powerful Machine Learning computing power for internal business units within the company and conducts research on some general and innovative algorithms for issues in these businesses.

Responsibilities
  • Responsible for the overall architecture design and implementation of model inference services, building a high-performance, highly available, and scalable enterprise-level inference system for large-parameter, high-complexity AI models, overcoming various architectural challenges in the implementation of complex model inference, and supporting the efficient launch of models across all business scenarios.
  • Responsible for the R&D and optimization of the core modules of the inference framework, covering core capabilities such as inference engine scheduling, monitoring and alerting, canary release, etc., continuously iterating on the framework performance, and resolving performance bottlenecks, resource bottlenecks, and stability issues in high-concurrency and large-model inference scenarios.
  • Keep track of the latest inference technologies in the industry, conduct technology selection and innovation in combination with business scenarios, accumulate distributed high-concurrency service architecture solutions, and promote the upgrade and standardization of the team's technical system.
Qualifications
Minimum Qualifications:
  • Individuals who are completing or have recently completed a Bachelor's/ Master's degree in computing or a related discipline.
  • Familiar with basic Linux commands, with solid C/C++ programming skills and knowledge of data structures and algorithms
  • Familiar with the basic principles of multi-threaded concurrency, proficient in basic usages such as thread usage, synchronization locks, and thread pools, able to identify common concurrency issues, and possess the ability to perform basic performance tuning in multi-threaded scenarios
  • Have experience in R&D projects of high-concurrency distributed services, and be familiar with service latency and resource optimization
  • Possess good learning and execution abilities, be willing to proactively understand the model inference service architecture, and have problem analysis and abstraction capabilities
  • Possess good cross-team collaboration skills, communication and presentation skills, and document writing skills, have strong sense of responsibility and stress tolerance, and be able to drive the resolution of complex technical issues and the implementation of projects
Preferred Qualifications:
  • Have practical project experience and understanding of source code in high-concurrency services/frameworks such as Redis, RocksDB, BRPC, GRPC, etc.
  • Understand the operating mechanism of GPUs, and have relevant project experience and optimization capabilities in GPU service resource management and control
Job Information
About Us

Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Lemon8, CapCut and Pico as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, ByteDance has made it easier and more fun for people to connect with, consume, and create content.

Why Join ByteDance

Inspiring creativity is at the core of ByteDance's mission. Our innovative products are built to help people authentically express themselves, discover and connect - and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and enrich life - a mission we work towards every day.

As ByteDancers, we strive to do great things with great people. We lead with curiosity, humility, and a desire to make impact in a rapidly growing tech company. By constantly iterating and fostering an "Always Day 1" mindset, we achieve meaningful breakthroughs for ourselves, our Company, and our users. When we create and grow together, the possibilities are limitless. Join us.

Diversity & Inclusion

ByteDance is committed to creating an inclusive space where employees are val

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Backend Inference Framework Engineer Graduate (AML Inference) - 2027 Start
Backend Inference Framework Engineer Graduate (AML Inference) - 2027 Start

ByteDance • Singapore

On-site
SGD 120,000 - 180,000
Backend Inference Framework Engineer Graduate (AML Inference) - 2027 Start Technology - Backend Bachelor/Master Graduate - 2027 Start Singapore Regular
Backend Inference Framework Engineer Graduate (AML Inference) - 2027 Start Technology - Backend Bachelor/Master Graduate - 2027 Start Singapore Regular

Bytedance • Singapore

On-site
SGD 180,000 - 280,000
Machine Learning Backend Engineer Graduate (AML MLdev) - 2027 Start
Machine Learning Backend Engineer Graduate (AML MLdev) - 2027 Start

BYTEDANCE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start

BYTEDANCE PTE. LTD. • Singapore

On-site
SGD 100,000 - 160,000
Inference Framework Engineer for High-Scale AI
Inference Framework Engineer for High-Scale AI

ByteDance • Singapore

On-site
SGD 120,000 - 180,000
Machine Learning System Engineer- Data AML- Soaring Star Talent Program
Machine Learning System Engineer- Data AML- Soaring Star Talent Program

Pangleglobal • Singapore

On-site
SGD 80,000 - 120,000
Graduate ML System Engineer - Large-Scale Systems
Graduate ML System Engineer - Large-Scale Systems

ByteDance • Singapore

On-site
SGD 60,000 - 80,000
Production Engineer Intern (AML Serving) - 2027 Start
Production Engineer Intern (AML Serving) - 2027 Start

ByteDance • Singapore

On-site
SGD 20,000 - 29,000
Backend Applied Machine Learning Engineer Graduate (AML Efficiency Tool) - 2027 Start Technology - Backend Bachelor/Master Graduate - 2027 Start Singapore Regular
Backend Applied Machine Learning Engineer Graduate (AML Efficiency Tool) - 2027 Start Technology - Backend Bachelor/Master Graduate - 2027 Start Singapore Regular

Bytedance • Singapore

On-site
SGD 60,000 - 120,000
Site Reliability Engineer - AI Application
Site Reliability Engineer - AI Application

ByteDance • Singapore

On-site
SGD 120,000 - 170,000