A cutting-edge technology company in San Francisco is seeking a Member of Technical Staff focused on kernels and GPU performance. This role involves optimizing GPU and accelerator kernels for AI workloads by analyzing performance across various hardware. Ideal candidates have strong software engineering foundations and experience with performance-critical systems. Familiarity with tools like CUDA and performance profiling is preferred. This position offers a dynamic environment focused on real-world performance optimizations.
Qualifications
Strong software engineering fundamentals.
Experience working on performance-critical systems close to hardware.
Comfort reasoning about low-level execution behavior and performance tradeoffs.
Responsibilities
Design, implement, and optimize GPU and accelerator kernels for AI workloads.
Analyze and tune performance across the GPU execution stack.
Work with compilers and runtimes to ensure kernel performance.
Bring up and optimize execution on new accelerators.
Profile, benchmark, and debug performance issues.
Ensure performance optimizations are production-ready.
Skills
Software engineering fundamentals
Performance-critical systems
Low-level execution behavior
Tools
CUDA
Triton
CUTLASS
Profiling tools
Job description
A cutting-edge technology company in San Francisco is seeking a Member of Technical Staff focused on kernels and GPU performance. This role involves optimizing GPU and accelerator kernels for AI workloads by analyzing performance across various hardware. Ideal candidates have strong software engineering foundations and experience with performance-critical systems. Familiarity with tools like CUDA and performance profiling is preferred. This position offers a dynamic environment focused on real-world performance optimizations.