GPU Kernel Engineer: Build Fast AI Inference at Scale
Baseten
San Francisco (CA)
On-site
USD 120,000 - 160,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
Competitive compensation
100% medical coverage
Generous PTO policy
Paid parental leave
Company-facilitated 401(k)
Job summary
A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal candidates have 1–5 years of CUDA development experience and a strong understanding of GPU architecture. This position offers competitive compensation, including equity, and comprehensive benefits including medical coverage and generous PTO.
Qualifications
1–5 years of experience in CUDA development.
Strong understanding of GPU architecture and programming paradigms.
Proficient in C++ and GPU performance profiling tools.
Responsibilities
Design and implement high-performance GPU kernels for ML operations.
Write and optimize code using CUDA and architecture-specific techniques.
Identify and resolve performance bottlenecks using profiling tools.
Skills
CUDA development
GPU architecture understanding
C++ proficiency
Performance profiling tools
Tools
CUDA C++ API
Nsight Systems
Nsight Compute
Job description
A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal candidates have 1–5 years of CUDA development experience and a strong understanding of GPU architecture. This position offers competitive compensation, including equity, and comprehensive benefits including medical coverage and generous PTO.