Senior GPU Performance Engineer for AI Training

CareerArc

San Jose (CA)

Hybrid

USD 150,000 - 200,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Competitive salary
Comprehensive benefits

Job summary

A leading semiconductor company is seeking a Fellow GPU Performance Optimization Engineer to maximize performance of AI workloads on AMD GPU platforms. The role requires expertise in GPU performance analysis and distributed systems. Candidates should have experience optimizing large-scale training and strong knowledge of GPU architecture. This position is hybrid in San Jose, California, and is not eligible for visa sponsorship.

Qualifications

  • Deep expertise in GPU performance optimization and distributed training.
  • Proven experience optimizing workloads across thousands of GPUs.
  • Strong understanding of ML frameworks for performance tuning.

Responsibilities

  • Lead performance optimization of large-scale AI training workloads.
  • Identify and eliminate system bottlenecks for compute and memory.
  • Optimize distributed training strategies for efficiency.

Skills

GPU architecture knowledge
Performance optimization
Distributed training expertise
Communication patterns understanding

Education

Ph.D. in Computer Science or related field

Tools

ROCm tools
NVIDIA Nsight
PyTorch
TensorFlow

Job description

A leading semiconductor company is seeking a Fellow GPU Performance Optimization Engineer to maximize performance of AI workloads on AMD GPU platforms. The role requires expertise in GPU performance analysis and distributed systems. Candidates should have experience optimizing large-scale training and strong knowledge of GPU architecture. This position is hybrid in San Jose, California, and is not eligible for visa sponsorship.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead GPU Performance Engineer for AI Training and Finetuning
Lead GPU Performance Engineer for AI Training and Finetuning

AMD • San Jose (CA)

Hybrid
USD 120,000 - 160,000
Senior AI Performance Architect — GPU & Network
Senior AI Performance Architect — GPU & Network

Advanced Micro Devices • San Jose (CA)

Hybrid
USD 130,000 - 160,000
Comprehensive benefits package
Staff Engineer: GPU Kernels & AI Performance
Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
GenAI Inference Optimization Lead — GPU Performance
GenAI Inference Optimization Lead — GPU Performance

Advanced Micro Devices • San Jose (CA)

Hybrid
USD 150,000 - 200,000
Lead GPU ML Training Performance Engineer
Lead GPU ML Training Performance Engineer

Advanced Micro Devices, Inc. • San Jose (CA)

On-site
USD 120,000 - 160,000
Fellow GPU Performance Optimization Engineer
Fellow GPU Performance Optimization Engineer

AMD • San Jose (CA)

On-site
USD 150,000 - 200,000
Competitive salary
Comprehensive benefits
AI Performance Software Engineer – GPU & DL Optimizations
AI Performance Software Engineer – GPU & DL Optimizations

AMD • Santa Clara (CA)

On-site
USD 120,000 - 160,000
Comprehensive benefits package
Inclusive culture
Career advancement opportunities
Senior GPU/AI Systems Engineer - Performance & ML
Senior GPU/AI Systems Engineer - Performance & ML

AMD • Santa Clara (CA)

On-site
USD 170,000 - 250,000
AMD Benefits
Principal AI Performance Engineer, GPU Systems Lead
Principal AI Performance Engineer, GPU Systems Lead

AMD • San Jose (CA)

On-site
USD 150,000 - 200,000
Competitive salary
Health benefits
Career advancement opportunities
Senior GPU Runtime & System Software Architect
Senior GPU Runtime & System Software Architect

AMD • San Jose (CA)

Hybrid
USD 230,000 - 420,000
AMD benefits