Research Engineer/Scientist(all levels), Efficient Models

TikTok

San Jose (CA)

On-site

USD 150,000 - 210,000

Full time

8 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

TikTok’s Vision-Applied Research team seeks a Research Engineer/Scientist to design and implement efficient large-scale generative AI models, with emphasis on distillation and compression. You will prototype methods and infrastructure to transfer capabilities from foundation models into smaller, scalable systems.

You will work on distillation frameworks, model acceleration, and hardware-efficient inference, collaborating with researchers to push state-of-the-art in image/video generation and

Qualifications

  • B.S. in Computer Science or related field or equivalent experience.
  • Expertise in efficient models with understanding of bottlenecks and acceleration methods.
  • Proficiency in training generative AI/LLM models using PyTorch and JAX.
  • Strong communication and collaboration skills in fast-paced environments.

Responsibilities

  • Develop efficient algorithms and architectures for large-scale generative and multimodal models, using distillation, quantization, and other methods to improve efficiency.
  • Advance scalable generative modeling approaches, including diffusion and autoregressive models, with focus on acceleration.

Skills

Efficient models
Communication
Collaboration

Education

B.S. in Computer Science or related fields

Tools

PyTorch
JAX

Job description

Responsibilities

About the Team

The Vision-Applied Research team focuses on applied research in Generative AI and CV/Multimodal Understanding, and delivering intelligent solutions to TikTok products, enabling users to make and share creative content in a much easier way. The team has research groups dedicated to generative models for content creation, image generation, video synthesis, intelligent image/video editing, and virtual humans.

The team is looking for a Research Engineer / Scientist who can take initiatives in designing and implementing efficient models for large-scale generative AI, with a particular emphasis on large model distillation and compression. The candidate will work on developing methods and infrastructure for transferring capabilities from foundation models into smaller, more efficient models, enabling scalable training, optimization, and deployment. Responsibilities may include, but are not limited to, distillation frameworks, model acceleration, hardware-efficient inference, and their applications.

Responsibilities
  • Develop efficient algorithms and architectures for large-scale generative and multimodal models, using techniques such as step distillation, cfg distillation, quantization, and other methods to improve model efficiency (e.g., image generation, video generation, VLM).
  • Advance scalable generative modeling approaches, including diffusion and autoregressive models, with a focus on acceleration and efficiency.
Qualifications
Minimum Qualifications:
  • B.S. in Computer Science or related fields, or equivalent experience
  • Expertise in efficient models with deep understanding of computational bottlenecks and acceleration methods.
  • Proficiency in training generative AI or LLM models using widely adopted frameworks and tools such as PyTorch and JAX.
  • Strong communication and collaboration skills in fast-paced environments.
Preferred Qualifications
  • Ph.D. in GenAI, MLSys or equivalent experience
  • Extensive research experiences in broad GenAI, MLSys, LLM areas.
  • Proven experiences in at least one of the following areas: image/video generation and editing; model compression (e.g., quantization, step/cfg distillation); efficient architectures (e.g., MoE, window attention); efficient model design; or reinforcement learning training methods (e.g., RLHF, DPO, GR
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior GenAI Research Engineer—Efficient Large Models
Senior GenAI Research Engineer—Efficient Large Models

TikTok • Seattle (WA)

On-site
USD 242,000 - 456,000
Generative AI Research Engineer
Generative AI Research Engineer

TikTok • San Jose (CA)

On-site
USD 150,000 - 210,000
Research Engineer
Research Engineer

Harnham • California (MO)

On-site
USD 120,000 - 180,000
GenAI Research Engineer: Distillation & Efficient Models
GenAI Research Engineer: Distillation & Efficient Models

ByteDance • Seattle (WA)

On-site
USD 242,000 - 456,000
Medical, dental, and vision insurance
401(k) with company match
Paid parental leave
+6
Generative AI Research Engineer — Efficient Models
Generative AI Research Engineer — Efficient Models

ByteDance • San Jose (CA)

On-site
USD 162,000 - 388,000
Research Engineer/Scientist, Efficient Models
Research Engineer/Scientist, Efficient Models

ByteDance • San Jose (CA)

On-site
USD 254,000 - 588,000
Senior GenAI Researcher: Efficient Models & Distillation
Senior GenAI Researcher: Efficient Models & Distillation

ByteDance • San Jose (CA)

On-site
USD 208,800 - 616,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+2
Research Engineer/Scientist(all levels), Efficient Models
Research Engineer/Scientist(all levels), Efficient Models

ByteDance • San Jose (CA)

On-site
USD 162,000 - 388,000
Sr. Research Engineer/Scientist(all levels), Efficient Models
Sr. Research Engineer/Scientist(all levels), Efficient Models

ByteDance • San Jose (CA)

On-site
USD 208,800 - 616,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+2
Efficient GenAI Research Engineer: Distillation
Efficient GenAI Research Engineer: Distillation

ByteDance • San Jose (CA)

On-site
USD 254,000 - 588,000