Get more replies from employers
Send a job-specific resume in minutes.
Inferact in San Francisco is seeking a Head of Engineering to build and lead the team developing systems that power vLLM and AI inference. You’ll ensure technical credibility at the inference layer, optimizing runtimes, memory and performance across GPUs and accelerators.
You’ll partner with founders to scale a senior-heavy organization, recruit rare ML systems talent, translate ambitious work into execution plans, and deliver high-performance inference across models, hardware, and deployment
Inferact in San Francisco is seeking a Head of Engineering to build and lead the team developing systems that power vLLM and AI inference. You’ll ensure technical credibility at the inference layer, optimizing runtimes, memory and performance across GPUs and accelerators.
You’ll partner with founders to scale a senior-heavy organization, recruit rare ML systems talent, translate ambitious work into execution plans, and deliver high-performance inference across models, hardware, and deployment