Stand out for this role — generate a tailored resume and cover letter in about a minute.
Prime Intellect in San Francisco is seeking an experienced ML systems engineer to scale LLM serving and RL components across cloud GPU fleets.
You will build multi-tenant serving platforms, optimize inference and integrate frameworks like vLLM, SGLang, and TensorRT-LLM, while implementing robust CI/CD and observability.
This hybrid role offers remote or San Francisco office options, visa sponsorship, relocation support, and meaningful equity to shape frontier AI infrastructure.
Prime Intellect in San Francisco is seeking an experienced ML systems engineer to scale LLM serving and RL components across cloud GPU fleets.
You will build multi-tenant serving platforms, optimize inference and integrate frameworks like vLLM, SGLang, and TensorRT-LLM, while implementing robust CI/CD and observability.
This hybrid role offers remote or San Francisco office options, visa sponsorship, relocation support, and meaningful equity to shape frontier AI infrastructure.