Get more replies from employers
Send a job-specific resume in minutes.
uRun, located in San Francisco, is seeking a founding ML Performance Engineer to drive AI infrastructure performance. In this role, you will write custom CUDA kernels and optimize model inference for real-time applications, significantly impacting performance across the stack.
The ideal candidate will possess deep knowledge of CUDA, experience with AI workloads, and a strong capacity for optimization. The position offers a competitive salary, equity, and top-tier tools for an exceptional contributor.
uRun, located in San Francisco, is seeking a founding ML Performance Engineer to drive AI infrastructure performance. In this role, you will write custom CUDA kernels and optimize model inference for real-time applications, significantly impacting performance across the stack.
The ideal candidate will possess deep knowledge of CUDA, experience with AI workloads, and a strong capacity for optimization. The position offers a competitive salary, equity, and top-tier tools for an exceptional contributor.