Senior GPU Kernel Performance Engineer

Designworks Talent

Bellevue (WA)

Hybrid

USD 140,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical insurance
Dental insurance
Vision insurance
401(k) with company match
Paid holidays

Job summary

Designworks Talent in Bellevue, WA is seeking a GPU Performance / Kernel Engineer to optimize the data plane powering large-scale AI workloads. You will tune kernels, reduce latency, and maximize throughput across training and inference environments, collaborating across AI infrastructure, ML, and platform teams.

The role focuses on identifying performance bottlenecks, profiling GPU kernels, and driving efficiency in a fast-moving engineering environment.

Qualifications

  • Experience with GPU kernel development and performance optimization (CUDA/ROCm).
  • Strong understanding of GPU architecture and memory hierarchy.
  • Experience profiling and debugging performance issues in AI/distributed environments.

Responsibilities

  • Profile, analyze, and optimize GPU kernels to improve latency and throughput.
  • Identify and eliminate data-plane bottlenecks impacting GPU performance across large-scale workloads.
  • Tune performance-critical workloads across training and inference environments.
  • Collaborate with AI infrastructure, ML, and platform engineering teams to optimize system behavior.
  • Develop benchmarking methodologies and performance measurement practices across GPU infrastructure.
  • Evaluate emerging GPU technologies and optimization techniques for evolving hardware.
  • Contribute to engineering practices that improve GPU efficiency, scalability, and reliability across the fleet.

Skills

CUDA
Kernel optimization
GPU architecture
Profiling tools
Distributed AI workloads
C/C++ systems programming
Nsight/ROCm tools

Tools

Nsight Systems
Nsight Compute
ROCm profiling tools

Job description

Designworks Talent in Bellevue, WA is seeking a GPU Performance / Kernel Engineer to optimize the data plane powering large-scale AI workloads. You will tune kernels, reduce latency, and maximize throughput across training and inference environments, collaborating across AI infrastructure, ML, and platform teams.

The role focuses on identifying performance bottlenecks, profiling GPU kernels, and driving efficiency in a fast-moving engineering environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff GPU Performance / Kernel Engineer
Staff GPU Performance / Kernel Engineer

Designworks Talent • Bellevue (WA)

Hybrid
USD 140,000 - 210,000
Medical insurance
Dental insurance
Vision insurance
+2
GPU/SOC Kernel Engineer - Performance, Power, Equity
GPU/SOC Kernel Engineer - Performance, Power, Equity

NVIDIA AI • Redmond (WA)

On-site
USD 180,000 - 260,000
Equity
GPU Kernel Engineer — High-Performance ML at Scale
GPU Kernel Engineer — High-Performance ML at Scale

The Consensus • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation with equity
100% medical, dental, and vision insurance coverage
Flexible PTO policy including a Winter Break
+2
Performance Engineer: GPU Kernel & Inference Optimize
Performance Engineer: GPU Kernel & Inference Optimize

WORLD LABS • San Francisco (CA)

On-site
USD 200,000 - 300,000
Senior GPU Kernel Engineer for High-Performance AI Inference
Senior GPU Kernel Engineer for High-Performance AI Inference

CoreWeave • Sunnyvale (CA)

On-site
USD 182,000 - 242,000
Medical, dental, and vision insurance
401(k) with employer match
ESPP
+2
Senior GPU Software Engineer — AI & Kernel Performance
Senior GPU Software Engineer — AI & Kernel Performance

AMD • United States

On-site
USD 150,000 - 210,000
GPU Kernel Engineer for High-Performance AI Inference
GPU Kernel Engineer for High-Performance AI Inference

Baseten • San Francisco (CA)

On-site
USD 180,000 - 360,000
Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
+3
Senior GPU Performance Engineer - Low-Level AI Kernels
Senior GPU Performance Engineer - Low-Level AI Kernels

Intel Corporation • Hillsboro (OR)

Hybrid
USD 195,000 - 276,000
Stock bonuses
Health benefits
Retirement plan
GPU Kernel Engineer: Build Fast AI Inference at Scale
GPU Kernel Engineer: Build Fast AI Inference at Scale

Baseten • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
100% medical coverage
Generous PTO policy
+2
GPU Kernel Engineer for AI Inference & Performance
GPU Kernel Engineer for AI Inference & Performance

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Flexible working hours
Daily lunch and dinner
Health check-up support
+3