Student Researcher - Compiler (Seed Infra) - 2026 Start

ByteDance

San Jose (CA)

On-site

USD 35,000 - 60,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

ByteDance Seed team in San Jose invites you to join as an intern on the Seed Infrastructures team, tackling AI compiler optimizations and MLIR-based passes for large-scale models.

You will work with researchers to optimize GPU/TPU/NPU performance and contribute to distributed training pipelines, for 12 weeks in 2026.

Qualifications

  • Pursuing a Bachelor's or Master's in CS, EE, or related field.
  • Experience with open source LLM inference frameworks (e.g., vLLM, SGLang) and at least one DL framework.
  • Understanding of computing systems and how they impact ML workloads.
  • Familiarity with compilers or model optimization pipelines (e.g., PyTorch Dynamo).
  • Available for a 12-week internship in 2026.

Responsibilities

  • Contribute to AI compiler optimizations for training and inference workloads
  • Develop and extend MLIR-based compiler passes for graph lowering, optimization, and code generation
  • Optimize model execution on GPUs and NPUs for performance and memory efficiency
  • Support model deployment pipelines, including compilation and runtime integration
  • Assist with distributed training and inference acceleration, including parallel execution and scheduling
  • Benchmark and profile large-scale models across hardware backends
  • Collaborate with researchers and engineers to translate requirements into compiler/runtime improvements

Skills

LLM inference
PyTorch
Megatron
DeepSpeed
JAX
MLIR
Compiler optimization
Distributed ML systems
GPU/TPU/NPU programming

Education

Bachelor’s or Master’s in CS/EE

Tools

vLLM
SGLang
PyTorch Dynamo
CUDA
Triton
NCCL

Job description

Join us as we work together to inspire creativity and enrich life around the globe.

Location:

San Jose

Team:

Technology

Employment Type:

Intern

Job Code:

A92858

Responsibilities

About the TeamThe Seed Infrastructures team oversees the distributed training, reinforcement learning framework, high-performance inference, and heterogeneous hardware compilation technologies for AI foundation models.

  • Responsibilities- Contribute to AI compiler optimizations for training and inference workloads
  • Develop and extend MLIR-based compiler passes for graph lowering, optimization, and code generation
  • Optimize model execution on GPU and NPU accelerators, focusing on performance, memory efficiency, and scalability
  • Support model deployment pipelines, including compilation, packaging, and runtime integration
  • Assist with distributed training and inference acceleration, such as parallel execution, communication optimization, and runtime scheduling
  • Benchmark, profile, and analyze performance of large-scale models across different hardware backends
  • Collaborate with researchers and engineers to translate model and system requirements into compiler and runtime improvements
Qualifications
  • Minimum Qualifications- Currently pursuing a Bachelor’s or Master’s degree in Computer Science, Electrical Engineering, or related technical fields
  • Experience using or developing open source frameworks for LLM inference such as vLLM or SGLang. Proficient in at least one deep learning framework (e.g., PyTorch, Megatron, DeepSpeed, JAX), with experience in model inference workflows
  • Understanding of modern computing systems, including hardware, storage, and networking, and how they impact ML workloads
  • Familiarity with compilers or model optimization pipelines (e.g., PyTorch Dynamo), or related model execution workflows
  • Able to commit to working for 12 weeks in 2026Preferred Qualifications- Experience with distributed or large-scale ML systems, including training or inference pipelines and related optimizations (e.g., FSDP, DeepSpeed, Megatron, GSPMD)
  • Experience with GPU/TPU/NPU programming and performance optimization, or high-performance computing and communication (e.g., CUDA, Triton, NCCL, RDMA)
  • Understanding of AI compiler and model optimization stacks (e.g., torch.fx, PyTorch Dynamo, XLA, MLIR)
Job Information
About Doubao (Seed)

Established in 2023, the ByteDance Seed team is dedicated to pioneering new paths toward artificial general intelligence. We aspire to advance the frontier of intelligence to drive progress for both technology and society.

With a long-term vision for the AI sector, the Seed team's research spans MLLM, GenMedia, AI for Science, and Robotics. We maintain a global presence with laboratories and career opportunities across China, Singapore, and the United States. To date, we have launched industry-leading general foundation models and cutting-edge multimodal capabilities. Our technology powers over 50 application scenarios — including Doubao, Jimeng, TRAE, Dola and Dreamnia — and serves enterprise customers through Volcano Engine and BytePlus. Third-party data shows that the Doubao App ranks first in user volume in the Chinese market, while Doubao foundation models lead the industry in average daily token consumption.

Why Join ByteDance

Inspiring creativity is at the core of ByteDance's mission. Our innovative products are built to help people authentically express themselves, discover and connect – and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and enrich life - a mission we work towards every day.

As ByteDancers, we strive to do great things with great people. We lead with curiosity, humility, and a desire to make impact in a rapidly growing tech company. By constantly iterating and fostering an “Always Day 1” mindset, we achieve meaningful breakthroughs for ourselves, our Company, and our users. When we create and grow together, the possibilities are limitless. Join us.

Diversity & Inclusion

ByteDance is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At ByteDance, our mission is to inspire creativity and enrich life. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.

Reasonable Accommodation

ByteDance is committed to providing reasonable accommodations in our recruitment processes for candidates with disabilities, pregnancy, sincerely held religious beliefs or other reasons protected by applicable laws. If you need assistance or a reasonable accommodation, please reach out to us at https://tinyurl.com/RA-request

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Student Researcher - Compiler (Seed Infra) - 2026 Start (PhD)
Student Researcher - Compiler (Seed Infra) - 2026 Start (PhD)

ByteDance • San Jose (CA)

On-site
USD 100,000 - 167,000
Research Engineer Graduate (AI Training Systems & RL Infrastructure - Seed Infra) - 2026 Start (PhD)
Research Engineer Graduate (AI Training Systems & RL Infrastructure - Seed Infra) - 2026 Start (PhD)

ByteDance • San Jose (CA)

On-site
USD 254,000 - 480,000
Medical/Dental/Vision insurance
401(k) with company match
Parental leave
+5
Student Researcher (AI Foundation Models Infrastructure – Seed Infra) – 2026 Start
Student Researcher (AI Foundation Models Infrastructure – Seed Infra) – 2026 Start

ByteDance • San Jose (CA)

On-site
USD 63,000 - 89,000
Student Researcher (AI Foundation Model Infrastructure - Seed) - 2027 Start Seed Foundation Mod[...]
Student Researcher (AI Foundation Model Infrastructure - Seed) - 2027 Start Seed Foundation Mod[...]

Bytedance • San Jose (CA), Northern (KY)

On-site
USD 25,000 - 39,000
Research Engineer Graduate (Agent Systems & AI Coding Environment – Seed Infra) – 2026 Start (PhD)
Research Engineer Graduate (Agent Systems & AI Coding Environment – Seed Infra) – 2026 Start (PhD)

ByteDance • Seattle (WA)

On-site
USD 242,000 - 456,000
Medical, dental and vision insurance
401(k) with company match
Paid parental leave
+2
Research Scientist Graduates (Seed AI Foundation Model Infrastructure) - 2027 Start
Research Scientist Graduates (Seed AI Foundation Model Infrastructure) - 2027 Start

ByteDance • San Jose (CA)

On-site
USD 218,000 - 388,000
Medical, dental, vision insurance
401(k) with company match
Paid parental leave
+6
Research Engineer – Reinforcement Learning (RL) Systems & Infrastructure (Seed Infra) San Jose [...]
Research Engineer – Reinforcement Learning (RL) Systems & Infrastructure (Seed Infra) San Jose [...]

Bytedance • San Jose (CA)

On-site
USD 244,800 - 450,000
Student Researcher (AI Foundation Model Infrastructure - Seed) - 2027 Start
Student Researcher (AI Foundation Model Infrastructure - Seed) - 2027 Start

ByteDance • San Jose (CA)

On-site
USD 63,000 - 89,000
Health insurance
Housing allowance
Paid holidays
+1
Student Researcher (AI Foundation Models Infrastructure - Seed Infra) - 2026 Start (PhD)
Student Researcher (AI Foundation Models Infrastructure - Seed Infra) - 2026 Start (PhD)

ByteDance • San Jose (CA)

On-site
USD 97,000 - 138,000
Health insurance
Housing allowance
Paid holidays
Research Scientist (LLM) - 2026 Start (PhD)
Research Scientist (LLM) - 2026 Start (PhD)

ByteDance • San Jose (CA)

On-site
USD 254,000 - 480,000
Health insurance
401(k) with company match
Parental leave
+1