Senior GPU Performance Engineer for AI Training

CareerArc

San Jose (CA)

Hybrid

USD 150,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary
Comprehensive benefits

Job summary

A leading semiconductor company is seeking a Fellow GPU Performance Optimization Engineer to maximize performance of AI workloads on AMD GPU platforms. The role requires expertise in GPU performance analysis and distributed systems. Candidates should have experience optimizing large-scale training and strong knowledge of GPU architecture. This position is hybrid in San Jose, California, and is not eligible for visa sponsorship.

Qualifications

  • Deep expertise in GPU performance optimization and distributed training.
  • Proven experience optimizing workloads across thousands of GPUs.
  • Strong understanding of ML frameworks for performance tuning.

Responsibilities

  • Lead performance optimization of large-scale AI training workloads.
  • Identify and eliminate system bottlenecks for compute and memory.
  • Optimize distributed training strategies for efficiency.

Skills

GPU architecture knowledge
Performance optimization
Distributed training expertise
Communication patterns understanding

Education

Ph.D. in Computer Science or related field

Tools

ROCm tools
NVIDIA Nsight
PyTorch
TensorFlow

Job description

A leading semiconductor company is seeking a Fellow GPU Performance Optimization Engineer to maximize performance of AI workloads on AMD GPU platforms. The role requires expertise in GPU performance analysis and distributed systems. Candidates should have experience optimizing large-scale training and strong knowledge of GPU architecture. This position is hybrid in San Jose, California, and is not eligible for visa sponsorship.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Fellow GPU Performance Optimizer for AI Training
Fellow GPU Performance Optimizer for AI Training

Advanced Micro Devices • San Jose (CA)

On-site
USD 140,000 - 180,000
Health insurance
Retirement plan
Paid time off
Lead GPU Performance Engineer for AI Training and Finetuning
Lead GPU Performance Engineer for AI Training and Finetuning

AMD • San Jose (CA)

Hybrid
USD 120,000 - 160,000
Senior GPU ML Training Performance Engineer
Senior GPU ML Training Performance Engineer

Advanced Micro Devices • San Jose (CA)

On-site
USD 130,000 - 170,000
Comprehensive health benefits
Inclusive workplace culture
Opportunities for career advancement
Lead GPU AI Inference Performance Engineer
Lead GPU AI Inference Performance Engineer

AMD • San Jose (CA)

On-site
USD 150,000 - 200,000
Senior GPU Training Performance Engineer
Senior GPU Training Performance Engineer

Advanced Micro Devices • San Jose (CA)

On-site
USD 120,000 - 160,000
Senior AI Performance Architect — GPU & Network
Senior AI Performance Architect — GPU & Network

Advanced Micro Devices • San Jose (CA)

Hybrid
USD 130,000 - 160,000
Comprehensive benefits package
Fellow GPU Performance Optimization Engineer
Fellow GPU Performance Optimization Engineer

Advanced Micro Devices • San Jose (CA)

On-site
USD 140,000 - 180,000
Health insurance
Retirement plan
Paid time off
Staff Engineer: GPU Kernels & AI Performance
Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
GenAI Inference Optimization Lead — GPU Performance
GenAI Inference Optimization Lead — GPU Performance

Advanced Micro Devices • San Jose (CA)

Hybrid
USD 150,000 - 200,000
Lead GPU ML Training Performance Engineer
Lead GPU ML Training Performance Engineer

Advanced Micro Devices, Inc. • San Jose (CA)

On-site
USD 120,000 - 160,000