Graduate AI Model Optimization Engineer

ByteDance

San Jose (CA)

On-site

USD 128,000 - 256,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ByteDance is seeking an AI Model Optimization Engineer to design and implement techniques to speed up AI models, improve efficiency, and simplify deployment at scale. You will collaborate with research and engineering to push performance in production.

Responsibilities include developing optimization algorithms, benchmarking, and optimizing GPU pipelines across distributed systems. A strong background in Python/C++, DL frameworks, and ML compilers is expected.

Qualifications

  • Bachelor's or master's in computer engineering or related field.
  • Strong coding skills in Python and C++.
  • Experience with deep learning frameworks and distributed training systems.
  • Familiarity with GPU programming (CUDA, Triton, or similar) and ML compilers like TVM/XLA/TensorRT is a plus.

Responsibilities

  • Develop and implement model optimization algorithms (quantization, pruning, distillation, efficient architectures).
  • Build and maintain performance benchmarking frameworks for large-scale training and inference.
  • Optimize training and inference pipelines on GPUs and across distributed systems.
  • Collaborate with ML researchers to productionize optimized models.
  • Stay current with research in model efficiency, compilers, and systems.

Skills

Python
C++
GPU programming (CUDA)
Deep learning frameworks
Distributed training systems
ML compilers familiarity

Education

Bachelor's or Master's in Computer Engineering or related

Tools

TVM
XLA
TensorRT
CUDA

Job description

ByteDance is seeking an AI Model Optimization Engineer to design and implement techniques to speed up AI models, improve efficiency, and simplify deployment at scale. You will collaborate with research and engineering to push performance in production.

Responsibilities include developing optimization algorithms, benchmarking, and optimizing GPU pipelines across distributed systems. A strong background in Python/C++, DL frameworks, and ML compilers is expected.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Graduate GPU AI Platform Engineer — System Optimization
Graduate GPU AI Platform Engineer — System Optimization

ByteDance • San Jose (CA)

On-site
USD 162,000 - 317,000
Graduate Research Scientist, AI Systems & Infrastructure
Graduate Research Scientist, AI Systems & Infrastructure

ByteDance • San Jose (CA)

On-site
USD 218,000 - 388,000
Medical, dental, vision insurance
401(k) with company match
Paid parental leave
+6
Graduate ML Engineer: Scalable AI Infrastructure
Graduate ML Engineer: Scalable AI Infrastructure

ByteDance • San Jose (CA)

On-site
USD 162,000 - 317,000
Medical, dental, and vision insurance
401(k) with company match
Paid parental leave
+6
Large-Model Inference Runtime Engineer
Large-Model Inference Runtime Engineer

Bytedance • San Jose (CA)

On-site
USD 128,000 - 256,000
Graduate Research Engineer, AI Training Systems
Graduate Research Engineer, AI Training Systems

ByteDance • Seattle (WA)

On-site
USD 242,000 - 456,000
Graduate Research Scientist - AI for Infra & Optimization
Graduate Research Scientist - AI for Infra & Optimization

Bytedance • San Jose (CA)

On-site
USD 150,000 - 210,000
Graduate Backend Inference Engine Engineer
Graduate Backend Inference Engine Engineer

ByteDance • San Jose (CA)

On-site
USD 128,000 - 256,000
Medical insurance
Dental insurance
Vision insurance
+8
Research Scientist, AI for Infrastructure & Optimization
Research Scientist, AI for Infrastructure & Optimization

ByteDance • San Jose (CA)

On-site
USD 218,000 - 388,000
Medical benefits
Dental benefits
Vision benefits
+6
Remote Model Optimization Engineer - AI Systems & GPU
Remote Model Optimization Engineer - AI Systems & GPU

Triwill Group • United States

Remote
USD 150,000 - 175,000
Graduate Software Engineer: AI Infrastructure & Scalable Systems
Graduate Software Engineer: AI Infrastructure & Scalable Systems

ByteDance • San Jose (CA)

On-site
USD 150,000 - 200,000