Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
NVIDIA is seeking a senior product leader to own the inference performance roadmap across AI models and serving stacks. You will orchestrate strategy for memory/state management, request scheduling, and token generation while coordinating across TensorRT-LLM, vLLM, SGLang, and NVIDIA Dynamo.
You will build scalable platforms rather than one-offs, drive cross-functional execution, and ensure credible benchmarking and production readiness for diverse deployments.
NVIDIA is seeking a senior product leader to own the inference performance roadmap across AI models and serving stacks. You will orchestrate strategy for memory/state management, request scheduling, and token generation while coordinating across TensorRT-LLM, vLLM, SGLang, and NVIDIA Dynamo.
You will build scalable platforms rather than one-offs, drive cross-functional execution, and ensure credible benchmarking and production readiness for diverse deployments.