Senior GPU Kernel & Performance Engineer

Designworks Talent LLC

Bellevue (WA)

Hybrid

USD 180,000 - 240,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Medical insurance
Dental insurance
Vision insurance
401(k) match
Paid holidays

Job summary

Designworks Talent LLC is seeking a Senior GPU Performance / Kernel Engineer to optimize the data plane powering large-scale AI workloads. Hybrid role in Bellevue, WA, with a focus on kernel tuning, latency reduction, and maximizing throughput across training and inference environments.

You will analyze workload behavior, identify bottlenecks, and develop tooling to operate AI systems efficiently at scale, collaborating with AI infra, ML, and platform teams.

Qualifications

  • Strong experience with GPU kernel development and performance optimization using CUDA/ROCm.
  • Proven track record of improving GPU utilization, reducing latency, or increasing throughput for production AI workloads.
  • Deep understanding of GPU architecture, memory hierarchy, and the data path from app to hardware.
  • Experience profiling and debugging performance issues in complex AI or distributed environments.
  • Ability to own technically complex problems and drive solutions in a fast-moving team.
  • Strong systems programming and performance engineering mindset.

Responsibilities

  • Profile, analyze, and optimize GPU kernels to improve latency, throughput, and utilization.
  • Identify and eliminate data-plane bottlenecks affecting GPU performance at scale.
  • Tune workloads across training and inference environments.
  • Collaborate with AI infra, ML, and platform teams to optimize workload behavior.
  • Develop benchmarking methodologies and performance measurement practices.
  • Evaluate emerging GPU tech and tools as hardware evolves.
  • Contribute to engineering practices improving GPU efficiency, scalability, and reliability.

Skills

CUDA
ROCm
GPU kernel development
Performance optimization
Profiling
Systems programming
Distributed AI workloads

Tools

Nsight Systems
Nsight Compute
ROCm profiling tools
GPU profiling tools

Job description

Designworks Talent LLC is seeking a Senior GPU Performance / Kernel Engineer to optimize the data plane powering large-scale AI workloads. Hybrid role in Bellevue, WA, with a focus on kernel tuning, latency reduction, and maximizing throughput across training and inference environments.

You will analyze workload behavior, identify bottlenecks, and develop tooling to operate AI systems efficiently at scale, collaborating with AI infra, ML, and platform teams.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior GPU Performance / Kernel Engineer
Senior GPU Performance / Kernel Engineer

Designworks Talent LLC • Bellevue (WA)

Hybrid
USD 180,000 - 240,000
Medical insurance
Dental insurance
Vision insurance
+2
Kernel Engineer: GPU Performance & Inference
Kernel Engineer: GPU Performance & Inference

Acceler8 Talent • San Francisco (CA)

On-site
USD 180,000 - 240,000
GPU/SOC Kernel Engineer - Performance, Power, Equity
GPU/SOC Kernel Engineer - Performance, Power, Equity

NVIDIA AI • Redmond (WA)

On-site
USD 180,000 - 260,000
Equity
Senior Performance Engineer: AI Workload Optimization
Senior Performance Engineer: AI Workload Optimization

NVIDIA • Redmond (WA)

On-site
USD 224,000 - 431,250
Equity
Benefits
CUDA Kernel Engineer — Optimize GPU Performance at Scale
CUDA Kernel Engineer — Optimize GPU Performance at Scale

Pragmatike • California (MO)

On-site
USD 180,000 - 240,000
Health, Dental, and Vision
Sign-on bonus
401k
Principal GPU Server Architect for AI Infra
Principal GPU Server Architect for AI Infra

Designworks Talent LLC • Bellevue (WA)

Hybrid
USD 180,000 - 230,000
Medical, dental, and vision insurance
401(k) plan with company match
Paid holidays
GPU Kernel Engineer — High-Performance ML at Scale
GPU Kernel Engineer — High-Performance ML at Scale

The Consensus • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation with equity
100% medical, dental, and vision insurance coverage
Flexible PTO policy including a Winter Break
+2
Senior AI Training Performance Engineer (GPU & Scale)
Senior AI Training Performance Engineer (GPU & Scale)

figure.ai • San Jose (CA), Northern (KY)

Hybrid
USD 200,000 - 400,000
GPU Kernel Engineer for High-Performance AI Inference
GPU Kernel Engineer for High-Performance AI Inference

Baseten • San Francisco (CA)

On-site
USD 180,000 - 360,000
Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
+3
Performance Engineer: GPU Kernel & Inference Optimize
Performance Engineer: GPU Kernel & Inference Optimize

WORLD LABS • San Francisco (CA)

On-site
USD 200,000 - 300,000