An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Prime Intellect is building an open frontier AI platform with a focus on scalable LLM serving and RL integration. This hybrid role spans cloud LLM serving, inference optimization, and RL systems.
You will advance our ability to evaluate and serve models trained with our RL Lab at scale, building multi-tenant serving, scheduling, and robust failover across regions. Requires 3+ years on ML/LLM infra, experience with vLLM or TensorRT-LLM, and strong Python/PyTorch skills.
Prime Intellect is building an open frontier AI platform with a focus on scalable LLM serving and RL integration. This hybrid role spans cloud LLM serving, inference optimization, and RL systems.
You will advance our ability to evaluate and serve models trained with our RL Lab at scale, building multi-tenant serving, scheduling, and robust failover across regions. Requires 3+ years on ML/LLM infra, experience with vLLM or TensorRT-LLM, and strong Python/PyTorch skills.