LLM/VLM Inference Optimization Engineer

ByteDance

Seattle (WA)

On-site

USD 232,560 - 427,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
Generous paid holidays and sick days

Job summary

ByteDance is seeking a Research Engineer - LLM/VLM Inference Optimization in Seattle. The role involves designing and optimizing high-performance inference systems for large-scale LLMs and VLMs, requiring expertise in C/C++ and Python, and familiarity with GPU optimization techniques.

Ideal candidates will have a Bachelor's degree in Computer Science or related fields and experience in production-level inference systems. Benefits include competitive salaries, comprehensive health insurance, and generous paid time off.

Qualifications

  • Bachelor's degree or above in Computer Science, Electrical Engineering, or related.
  • Strong proficiency in C/C++ and Python.
  • Experience deploying LLM/VLM inference at production scale.
  • Familiarity with GPU architecture.
  • Experience with performance optimization.

Responsibilities

  • Design and optimize high-performance inference systems for large-scale LLMs and VLMs.
  • Build model inference engines using advanced optimization techniques.
  • Collaborate with research teams to identify performance issues.

Skills

C/C++ programming
Python
Algorithms and data structures
Machine learning frameworks (e.g., PyTorch, TensorFlow)
GPU architecture optimization

Education

Bachelor's degree in Computer Science or related field

Tools

CUDA
TensorRT
Triton

Job description

ByteDance is seeking a Research Engineer - LLM/VLM Inference Optimization in Seattle. The role involves designing and optimizing high-performance inference systems for large-scale LLMs and VLMs, requiring expertise in C/C++ and Python, and familiarity with GPU optimization techniques.

Ideal candidates will have a Bachelor's degree in Computer Science or related fields and experience in production-level inference systems. Benefits include competitive salaries, comprehensive health insurance, and generous paid time off.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM/VLM Inference Optimization Research Engineer
LLM/VLM Inference Optimization Research Engineer

Bytedance • San Jose (CA)

On-site
USD 244,000 - 450,000
Medical, dental, and vision insurance
401(k) with company match
Paid parental leave
+1
LLM Training & Inference Scientist (GPU-Optimized)
LLM Training & Inference Scientist (GPU-Optimized)

ByteDance • San Jose (CA)

On-site
USD 212,000 - 450,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+3
LLM Training Infrastructure Research Engineer
LLM Training Infrastructure Research Engineer

ByteDance • Seattle (WA)

On-site
USD 232,000 - 428,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+2
Senior LLM Storage Systems Engineer & Researcher
Senior LLM Storage Systems Engineer & Researcher

ByteDance • Seattle (WA)

On-site
USD 202,160 - 368,220
Medical, dental, vision insurance
401(k) matching
Paid parental leave
+1
LLM Inference Systems Engineer
LLM Inference Systems Engineer

ByteDance • San Jose (CA)

On-site
USD 128,000 - 256,000
Medical insurance
Dental insurance
Vision insurance
+4
Research Engineer - LLM/VLM Inference Optimization (Seed Infra) Seattle Regular
Research Engineer - LLM/VLM Inference Optimization (Seed Infra) Seattle Regular

ByteDance • Seattle (WA)

On-site
USD 232,000 - 428,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+1
LLM Storage Systems Research Engineer
LLM Storage Systems Research Engineer

ByteDance • San Jose (CA)

On-site
USD 156,000 - 388,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+2
LLM Inference Optimization Engineer — Frontier Performance
LLM Inference Optimization Engineer — Frontier Performance

NLP PEOPLE • Sonoma (CA)

On-site
USD 120,000 - 160,000
LLM Research Scientist: Advancing Large-Scale AI
LLM Research Scientist: Advancing Large-Scale AI

ByteDance • San Jose (CA)

On-site
USD 254,000 - 480,000
Health insurance
401(k) with company match
Parental leave
+1
AI Infrastructure & LLM Systems Research Scientist
AI Infrastructure & LLM Systems Research Scientist

ByteDance • San Jose (CA)

On-site
USD 212,000 - 388,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+1