Post-Training LLM Inference Platform Engineer

Advanced Micro Devices

San Jose (CA)

Hybrid

USD 100,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Comprehensive benefits
Collaborative work environment
Opportunities for career advancement

Job summary

Advanced Micro Devices is looking for a systems-minded engineer in San Jose, CA, focusing on ML infrastructure and performance optimization for large-scale model inference. Ideal candidates should have a strong background in systems engineering and experience with GPU workloads. This hybrid position offers collaboration opportunities with research teams to enhance model performance. A bachelor's or master's degree in relevant fields is required. AMD promotes a diverse and inclusive culture.

Qualifications

  • Experience with GPU-accelerated workloads.
  • Familiarity with LLM inference workflows.
  • Solid understanding of KV cache behavior.

Responsibilities

  • Research modern LLM inference frameworks.
  • Design and implement KV cache lifecycle.
  • Analyze and optimize performance bottlenecks.

Skills

Systems engineering
Distributed systems
ML infrastructure
Python
C++
Analytical skills
Performance optimization

Education

Bachelor's or master's degree in computer science, computer engineering, electrical engineering

Tools

Profiling tools
Open-source codebases

Job description

Advanced Micro Devices is looking for a systems-minded engineer in San Jose, CA, focusing on ML infrastructure and performance optimization for large-scale model inference. Ideal candidates should have a strong background in systems engineering and experience with GPU workloads. This hybrid position offers collaboration opportunities with research teams to enhance model performance. A bachelor's or master's degree in relevant fields is required. AMD promotes a diverse and inclusive culture.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead GPU ML Training Performance Engineer
Lead GPU ML Training Performance Engineer

Advanced Micro Devices, Inc. • San Jose (CA)

On-site
USD 120,000 - 160,000
Principal ML Engineer: Large-Scale Training Performance
Principal ML Engineer: Large-Scale Training Performance

Advanced Micro Devices • San Jose (CA)

Hybrid
USD 120,000 - 180,000
Comprehensive benefits package
Innovative work culture
ML Systems Engineer for RL & Inference Infrastructure
ML Systems Engineer for RL & Inference Infrastructure

Advanced Micro Devices • Santa Clara (CA)

Hybrid
USD 160,000 - 210,000
AMD benefits
AMD GPU Performance Engineer — Inference Backends & Kernels
AMD GPU Performance Engineer — Inference Backends & Kernels

Inferact • San Francisco (CA)

Hybrid
USD 200,000 - 400,000
Health benefits
Dental benefits
Vision benefits
+1
LLM Inference Systems Performance Engineer
LLM Inference Systems Performance Engineer

3M HEALTHCARE • Austin (TX)

On-site
USD 150,000 - 210,000
Medical, dental, and vision coverage
Income protection benefits
Paid family leave
+1
RL & Inference ML Systems Engineer for Engineering AI
RL & Inference ML Systems Engineer for Engineering AI

AMD • Santa Clara (CA)

On-site
USD 180,000 - 250,000
AMD benefits
Lead RL Infra Engineer - Scalable GPU Training Platforms
Lead RL Infra Engineer - Scalable GPU Training Platforms

Advanced Micro Devices • Santa Clara (CA)

On-site
USD 120,000 - 170,000
Open Source ML Systems Engineer: GPU Kernels & Inference
Open Source ML Systems Engineer: GPU Kernels & Inference

Advanced Micro Devices • San Jose (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Principal ML Engineer: Large-Scale Training & Performance
Principal ML Engineer: Large-Scale Training & Performance

AMD • San Jose (CA)

Hybrid
USD 130,000 - 160,000
Senior Inference Systems Engineer (AMD/TPU)
Senior Inference Systems Engineer (AMD/TPU)

Inferact • San Francisco (CA)

On-site
USD 200,000 - 400,000
Health, dental, and vision benefits
401(k) company match
Equity compensation