Stand out for this role — generate a tailored resume and cover letter in about a minute.
Voltai is hiring in California to build and optimize high-performance ML pipelines and inference systems for LLMs in enterprise environments. You will design CUDA kernels, implement low-precision techniques, and Architect distributed training across multiple GPUs and nodes.
The role emphasizes production-ready research translate, collaboration with researchers and infra teams, and tuning for real-world hardware deployments in a fast-paced startup setting.
Voltai is hiring in California to build and optimize high-performance ML pipelines and inference systems for LLMs in enterprise environments. You will design CUDA kernels, implement low-precision techniques, and Architect distributed training across multiple GPUs and nodes.
The role emphasizes production-ready research translate, collaboration with researchers and infra teams, and tuning for real-world hardware deployments in a fast-paced startup setting.