GPU Kernel Engineer for AI Inference & Performance

FriendliAI

San Francisco (CA)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Flexible working hours
Daily lunch and dinner
Health check-up support
Competitive compensation
Startup equity
Health insurance

Job summary

FriendliAI is seeking a GPU Kernel Engineer in San Francisco to design and optimize GPU kernels for AI inference. This role requires expertise in CUDA, C++, and performance-critical systems. You will work on cutting-edge GPU technology and contribute to a highly collaborative, supportive work environment with competitive compensation and startup equity. Join us at the forefront of AI infrastructure innovation.

Qualifications

  • 3+ years of experience in GPU programming.
  • Strong proficiency in CUDA for NVIDIA GPUs or ROCm/HIP for AMD GPUs.
  • Deep understanding of GPU architecture and tuning.

Responsibilities

  • Design, implement, and optimize high-performance GPU kernels for AI inference.
  • Develop and maintain GPU code in CUDA and C++.
  • Benchmark and ensure performance parity between NVIDIA and AMD hardware.

Skills

GPU programming
Performance-critical systems
CUDA proficiency
C++ proficiency
Performance tuning

Education

Bachelor’s or Master’s in Computer Science, Computer Engineering, Electrical Engineering

Tools

CUDA
C++
ROCm/HIP

Job description

FriendliAI is seeking a GPU Kernel Engineer in San Francisco to design and optimize GPU kernels for AI inference. This role requires expertise in CUDA, C++, and performance-critical systems. You will work on cutting-edge GPU technology and contribute to a highly collaborative, supportive work environment with competitive compensation and startup equity. Join us at the forefront of AI infrastructure innovation.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Kernel Engineer for High-Performance AI Inference
GPU Kernel Engineer for High-Performance AI Inference

Baseten • San Francisco (CA)

On-site
USD 180,000 - 360,000
Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
+3
GPU Kernel Engineer: Build Fast AI Inference at Scale
GPU Kernel Engineer: Build Fast AI Inference at Scale

Baseten • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
100% medical coverage
Generous PTO policy
+2
Senior GPU Inference Engine Engineer
Senior GPU Inference Engine Engineer

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Flexible working hours
Daily lunch and dinner provided; unlimited snacks and beverages
Health check-up support and top-tier equipment/hardware support
+2
Senior AI Kernel Engineer — GPU Inference & Kernel Optimization
Senior AI Kernel Engineer — GPU Inference & Kernel Optimization

Modular • United States

Hybrid
USD 198,000 - 286,000
Staff GenAI Kernel & Performance Engineer
Staff GenAI Kernel & Performance Engineer

Databricks • San Francisco (CA)

On-site
USD 190,000 - 233,000
GPU Kernel Engineer for High-Performance AI Inference
GPU Kernel Engineer for High-Performance AI Inference

Baseten • New York (NY)

On-site
USD 180,000 - 360,000
Staff Engineer: GPU Kernels & AI Performance
Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
GPU Performance Engineer: Scale AI Inference
GPU Performance Engineer: Scale AI Inference

Anthropic • San Francisco (CA)

On-site
USD 315,000 - 560,000
Competitive salary
Equity opportunities
Flexible working hours
+1
CUDA GPU Kernel Architect for High-Throughput AI
CUDA GPU Kernel Architect for High-Throughput AI

TypeSafe AI • San Francisco (CA)

On-site
USD 180,000 - 280,000
Base salary $180k–$280k plus equity
100% covered health insurance
Daily lunch and dinner
+2
Senior Kernel & Compiler Performance Engineer (GPU/AI)
Senior Kernel & Compiler Performance Engineer (GPU/AI)

RadixArk • Palo Alto (CA)

On-site
Competitive compensation
Comprehensive benefits
Flexible work arrangements