AI Systems & Cloud Infrastructure Intern

ByteDance

San Jose (CA)

On-site

USD 50,000 - 83,000

Full time

12 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Health insurance
Life insurance
Wellbeing benefits
Housing allowance
Paid holidays
Paid sick time

Job summary

ByteDance is seeking PhD students for a Summer 2027 internship to contribute to our DPU team, building scalable cloud-native GPU and AI accelerator infrastructure, and advancing LLM inference platforms. You will work with vLLM, SGLang, TensorRT-LLM, and related technologies across cloud and AI systems.

You will collaborate with global teams, gain hands-on experience, and participate in professional development events designed for researchers and engineers pursuing cutting-edge cloud and AI

Qualifications

  • Currently pursuing a PhD in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.
  • Able to commit to working for 12 weeks during Summer 2027.
  • Strong understanding of large model inference, distributed and parallel systems, and/or high-performance networking systems.
  • Hands‑on experience building cloud or ML infrastructure in areas such as resource management, scheduling, request routing, monitoring, or orchestration.
  • Solid knowledge of container and orchestration technologies (Docker, Kubernetes).
  • Proficiency in at least one major programming language (Go, Rust, Python, or C++).

Responsibilities

  • Design and build large-scale, container-based cluster management and orchestration systems with extreme performance, scalability, and resilience.
  • Architect next-generation cloud-native GPU and AI accelerator infrastructure to deliver cost-efficient and secure ML platforms.
  • Collaborate across teams to deliver world-class inference solutions using vLLM, SGLang, TensorRT-LLM, and other LLM engines.
  • Stay current with open source and AI/ML infrastructure advances; integrate best practices into production systems.
  • Write high-quality, production-ready code that is maintainable, testable, and scalable.

Skills

Large model inference
Distributed systems
High-performance networking
Cloud infrastructure
Container orchestration
Programming: Go/Python/C++

Education

PhD in CS/CE/EE or related

Tools

Docker
Kubernetes

Job description

ByteDance is seeking PhD students for a Summer 2027 internship to contribute to our DPU team, building scalable cloud-native GPU and AI accelerator infrastructure, and advancing LLM inference platforms. You will work with vLLM, SGLang, TensorRT-LLM, and related technologies across cloud and AI systems.

You will collaborate with global teams, gain hands-on experience, and participate in professional development events designed for researchers and engineers pursuing cutting-edge cloud and AI

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Intern, Cloud & AI Systems
Research Intern, Cloud & AI Systems

ByteDance • Seattle (WA)

On-site
USD 55,000 - 83,000
AI Infrastructure Intern — Cloud Compute Platform
AI Infrastructure Intern — Cloud Compute Platform

ByteDance • San Jose (CA)

On-site
USD 51,000 - 73,000
Health insurance from day one
Housing allowance
Paid holidays
PhD Research Intern — AI Infra & Cloud Compute
PhD Research Intern — AI Infra & Cloud Compute

Bytedance • San Jose (CA), Northern (KY)

Hybrid
USD 34,000 - 48,000
PhD Research Intern — Cloud AI Systems & HPC
PhD Research Intern — Cloud AI Systems & HPC

ByteDance • San Jose (CA)

On-site
USD 68,000 - 97,000
Health insurance
Life insurance
Wellbeing benefits
+2
AI Compute Research Intern: Cloud & GPUs (PhD, Summer 2027)
AI Compute Research Intern: Cloud & GPUs (PhD, Summer 2027)

ByteDance • Seattle (WA)

On-site
USD 65,000 - 92,000
Health insurance
Housing allowance
Paid holidays
PhD Cloud & AI Systems Research Intern
PhD Cloud & AI Systems Research Intern

ByteDance • San Jose (CA)

On-site
USD 68,000 - 97,000
Health insurance
Wellbeing benefits
Paid holidays
+2
AI Infra Engineer Intern: Build Scalable AI Compute
AI Infra Engineer Intern: Build Scalable AI Compute

ByteDance • San Jose (CA)

On-site
USD 123,984,000 - 137,760,000
Health insurance
Housing allowance
Paid holidays
+1
PhD AI Infra Research Intern — Compute Platform
PhD AI Infra Research Intern — Compute Platform

ByteDance • San Jose (CA)

On-site
USD 68,000 - 97,000
Health insurance from day one
Wellbeing benefits
10 paid holidays per year
+1
AI Compute & DPU Research Scientist
AI Compute & DPU Research Scientist

ByteDance • San Jose (CA)

On-site
USD 212,000 - 388,000
Medical, dental and vision insurance
401(k) with company match
Paid parental leave
+6
AI Infrastructure & ML Systems Intern
AI Infrastructure & ML Systems Intern

ByteDance • San Jose (CA)

On-site
USD 97,000 - 138,000
Health insurance
Housing allowance
Paid holidays