Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

A cutting-edge technology company in San Francisco is seeking a Member of Technical Staff focused on kernels and GPU performance. This role involves optimizing GPU and accelerator kernels for AI workloads by analyzing performance across various hardware. Ideal candidates have strong software engineering foundations and experience with performance-critical systems. Familiarity with tools like CUDA and performance profiling is preferred. This position offers a dynamic environment focused on real-world performance optimizations.

Qualifications

  • Strong software engineering fundamentals.
  • Experience working on performance-critical systems close to hardware.
  • Comfort reasoning about low-level execution behavior and performance tradeoffs.

Responsibilities

  • Design, implement, and optimize GPU and accelerator kernels for AI workloads.
  • Analyze and tune performance across the GPU execution stack.
  • Work with compilers and runtimes to ensure kernel performance.
  • Bring up and optimize execution on new accelerators.
  • Profile, benchmark, and debug performance issues.
  • Ensure performance optimizations are production-ready.

Skills

Software engineering fundamentals
Performance-critical systems
Low-level execution behavior

Tools

CUDA
Triton
CUTLASS
Profiling tools

Job description

A cutting-edge technology company in San Francisco is seeking a Member of Technical Staff focused on kernels and GPU performance. This role involves optimizing GPU and accelerator kernels for AI workloads by analyzing performance across various hardware. Ideal candidates have strong software engineering foundations and experience with performance-critical systems. Familiarity with tools like CUDA and performance profiling is preferred. This position offers a dynamic environment focused on real-world performance optimizations.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Kernel & Compiler Performance Engineer (GPU/AI)
Senior Kernel & Compiler Performance Engineer (GPU/AI)

RadixArk • Palo Alto (CA)

On-site
USD 210,000 - 290,000
Competitive compensation
Comprehensive benefits
Flexible work arrangements
Staff GPU Kernel & Performance Engineer
Staff GPU Kernel & Performance Engineer

Gimlet Labs, Inc. • San Francisco (CA)

On-site
USD 150,000 - 350,000
Staff GenAI Kernel & Performance Engineer
Staff GenAI Kernel & Performance Engineer

Databricks • San Francisco (CA)

On-site
USD 190,900 - 232,800
Annual performance bonus
Equity options
Comprehensive benefits package
Kernel Engineer: GPU Performance & Inference
Kernel Engineer: GPU Performance & Inference

Acceler8 Talent • San Francisco (CA)

On-site
USD 180,000 - 240,000
GPU Kernel Engineer for AI Inference & Performance
GPU Kernel Engineer for AI Inference & Performance

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Flexible working hours
Daily lunch and dinner
Health check-up support
+3
GPU Kernel Engineer for High-Performance AI Inference
GPU Kernel Engineer for High-Performance AI Inference

Baseten • San Francisco (CA)

On-site
USD 180,000 - 360,000
Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
+3
Senior GPU Performance Engineer for AI Training
Senior GPU Performance Engineer for AI Training

CareerArc • San Jose (CA)

Hybrid
USD 150,000 - 200,000
Competitive salary
Comprehensive benefits
GPU Kernel Engineer: Build Fast AI Inference at Scale
GPU Kernel Engineer: Build Fast AI Inference at Scale

Baseten • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
100% medical coverage
Generous PTO policy
+2
Kernel Performance Engineer - AI Tooling & Systems
Kernel Performance Engineer - AI Tooling & Systems

OpenAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
AI Performance & Kernel Engineer for Frontier-Scale ML
AI Performance & Kernel Engineer for Frontier-Scale ML

Zyphra • San Francisco (CA)

On-site
USD 120,000 - 160,000
Comprehensive medical, dental, vision, and FSA plans
Competitive compensation and 401(k) plan
Relocation and immigration support
+3