AI Compute Research Intern: Cloud & GPUs (PhD, Summer 2027)

ByteDance

Seattle (WA)

On-site

USD 65,000 - 92,000

Full time

10 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Health insurance
Housing allowance
Paid holidays

Job summary

ByteDance in Seattle is seeking a Research Intern (AI Compute) for a 12-week Summer 2027 program in our Technology team. You will design and build large-scale, container-based cluster management and orchestration systems with extreme performance and scalability, and architect GPU/AI accelerator infrastructure for cost-efficient ML platforms.

The ideal candidate is a PhD student in CS/CE with hands-on experience in Kubernetes, GPU programming, and distributed systems, and can contribute to

Qualifications

  • Currently pursuing a PhD in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.
  • Able to commit to working for 12 weeks during Summer 2027.
  • Strong understanding of large model inference, distributed and parallel systems, and/or high-performance networking systems.
  • Hands-on experience building cloud or ML infrastructure in areas such as resource management, scheduling, request routing, monitoring, or orchestration.
  • Solid knowledge of container and orchestration technologies (Docker, Kubernetes).
  • Proficiency in at least one major programming language (Go, Rust, Python, or C++).
  • Experience contributing to or operating large-scale cluster management systems (e.g., Kubernetes, Ray).
  • Experience with workload scheduling, GPU orchestration, scaling, and isolation in production environments.
  • Hands-on experience with GPU programming (CUDA) or inference engines (vLLM, SGLang, TensorRT-LLM).
  • Familiarity with public cloud providers (AWS, Azure, GCP) and their ML platforms (SageMaker, Azure ML, Vertex AI).
  • Strong knowledge of ML systems (Ray, DeepSpeed, PyTorch) and distributed training/inference platforms.
  • Excellent communication skills and ability to collaborate across global, cross-functional teams.
  • Passion for system efficiency, performance optimization, and open-source innovation.

Responsibilities

  • Design and build large-scale, container-based cluster management and orchestration systems with extreme performance, scalability, and resilience.
  • Architect next-generation cloud-native GPU and AI accelerator infrastructure to deliver cost-efficient and secure ML platforms.
  • Collaborate across teams to deliver world-class inference solutions using vLLM, SGLang, TensorRT-LLM, and other LLM engines.
  • Stay current with the latest advances in open source (Kubernetes, Ray, etc.), AI/ML and LLM infrastructure; integrate best practices into production systems.
  • Write high-quality, production-ready code that is maintainable, testable, and scalable.

Skills

Docker
Kubernetes
Python
C++
Go

Education

PhD in Computer Science/Engineering

Tools

vLLM
SGLang
TensorRT-LLM
Ray

Job description

ByteDance in Seattle is seeking a Research Intern (AI Compute) for a 12-week Summer 2027 program in our Technology team. You will design and build large-scale, container-based cluster management and orchestration systems with extreme performance and scalability, and architect GPU/AI accelerator infrastructure for cost-efficient ML platforms.

The ideal candidate is a PhD student in CS/CE with hands-on experience in Kubernetes, GPU programming, and distributed systems, and can contribute to

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Intern, Cloud & AI Systems
Research Intern, Cloud & AI Systems

ByteDance • Seattle (WA)

On-site
USD 55,000 - 83,000
PhD Research Intern — AI Infra & Cloud Compute
PhD Research Intern — AI Infra & Cloud Compute

Bytedance • San Jose (CA), Northern (KY)

Hybrid
USD 34,000 - 48,000
Graduate Research Engineer, AI Infra & Compute
Graduate Research Engineer, AI Infra & Compute

ByteDance • Seattle (WA)

On-site
USD 130,000 - 180,000
AI Compute Scheduler Intern (PhD)
AI Compute Scheduler Intern (PhD)

Bytedance • San Jose (CA), Northern (KY)

Hybrid
USD 48,000 - 62,000
PhD AI Infra Research Intern — Compute Platform
PhD AI Infra Research Intern — Compute Platform

ByteDance • San Jose (CA)

On-site
USD 68,000 - 97,000
Health insurance from day one
Wellbeing benefits
10 paid holidays per year
+1
Cloud Acceleration Research Intern (DPU & AI Infra) - 2027 Start (PhD)
Cloud Acceleration Research Intern (DPU & AI Infra) - 2027 Start (PhD)

ByteDance • Seattle (WA)

On-site
USD 201,000 - 268,000
AI Infra Intern: Cloud Compute & Prototyping
AI Infra Intern: Cloud Compute & Prototyping

ByteDance • Seattle (WA)

On-site
USD 20,000 - 31,000
AI Infra Intern: Build Smarter Compute at Scale
AI Infra Intern: Build Smarter Compute at Scale

Bytedance • Seattle (WA), Northern (KY)

Hybrid
USD 34,000 - 55,000
Graduate Research Scientist, DPU & AI/ML Infrastructure
Graduate Research Scientist, DPU & AI/ML Infrastructure

ByteDance • Seattle (WA)

On-site
USD 180,000 - 240,000
PhD Research Intern: AI Systems & Scale Infrastructure
PhD Research Intern: AI Systems & Scale Infrastructure

ByteDance • Seattle (WA)

On-site
Health insurance
Life insurance
Wellbeing benefits
+2