Efficient GenAI Research Engineer: Distillation

ByteDance

San Jose (CA)

On-site

USD 254,000 - 588,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

ByteDance is seeking a Research Engineer/Scientist to design and implement efficient models for large-scale generative AI, with emphasis on distillation, compression, and deployment. You will work on model acceleration and hardware-aware inference within the Vision-Applied Research team in San Jose.

The role requires experience in quantization, diffusion/ autoregressive methods, and a strong background in PyTorch or JAX to optimize model performance and scalability for real-world applications.

Qualifications

  • Bachelor's degree in CS or related field or equivalent experience.
  • Expertise in efficient models with understanding of bottlenecks and acceleration methods.
  • Experience training generative AI or LLM models using PyTorch or JAX.

Responsibilities

  • Develop efficient algorithms and architectures for large-scale generative and multimodal models, using distillation and quantization to improve efficiency.
  • Advance scalable generative modeling approaches, including diffusion and autoregressive models, with focus on acceleration and efficiency.

Skills

Efficient models
Model distillation
Quantization
Large-scale generation
Hardware acceleration
PyTorch
JAX

Education

Bachelor's degree in Computer Science or related field
Ph.D. in GenAI, MLSys or equivalent

Tools

PyTorch
JAX

Job description

ByteDance is seeking a Research Engineer/Scientist to design and implement efficient models for large-scale generative AI, with emphasis on distillation, compression, and deployment. You will work on model acceleration and hardware-aware inference within the Vision-Applied Research team in San Jose.

The role requires experience in quantization, diffusion/ autoregressive methods, and a strong background in PyTorch or JAX to optimize model performance and scalability for real-world applications.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

GenAI Research Engineer: Distillation & Efficient Models
GenAI Research Engineer: Distillation & Efficient Models

ByteDance • Seattle (WA)

On-site
USD 242,000 - 456,000
Medical, dental, and vision insurance
401(k) with company match
Paid parental leave
+6
Generative AI Research Engineer — Efficient Models
Generative AI Research Engineer — Efficient Models

ByteDance • San Jose (CA)

On-site
USD 162,000 - 388,000
Senior GenAI Research Engineer—Efficient Large Models
Senior GenAI Research Engineer—Efficient Large Models

TikTok • Seattle (WA)

On-site
USD 242,000 - 456,000
Generative AI Research Engineer
Generative AI Research Engineer

TikTok • San Jose (CA)

On-site
USD 150,000 - 210,000
Senior GenAI Researcher: Efficient Models & Distillation
Senior GenAI Researcher: Efficient Models & Distillation

ByteDance • San Jose (CA)

On-site
USD 208,800 - 616,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+2
Research Engineer/Scientist(all levels), Efficient Models
Research Engineer/Scientist(all levels), Efficient Models

TikTok • San Jose (CA)

On-site
USD 150,000 - 210,000
Research Engineer – Distributed AI Infrastructure
Research Engineer – Distributed AI Infrastructure

ByteDance • San Jose (CA)

On-site
USD 254,000 - 480,000
Medical/Dental/Vision insurance
401(k) with company match
Parental leave
+5
Ai Distilation Expert
Ai Distilation Expert

RB Labs • San Francisco (CA)

On-site
USD 130,000 - 180,000
Graduate AI Model Optimization Engineer
Graduate AI Model Optimization Engineer

ByteDance • San Jose (CA)

On-site
USD 128,000 - 256,000
Real-Time AI Distillation Engineer
Real-Time AI Distillation Engineer

RB Labs • San Francisco (CA)

On-site
USD 130,000 - 180,000