Get more replies from employers
Send a job-specific resume in minutes.
NVIDIA is seeking highly skilled software engineers to build AI inference systems that scale across multi-GPU, multi-node, and multi-cloud deployments. You will architect high-performance inference stacks, optimize kernels and compilers, and collaborate across inference, compiler, scheduling, and performance teams to push the frontier of accelerated AI computing.
Responsibilities include contributing features to vLLM, benchmarking, and driving MLPerf Inference submissions, with emphasis on
NVIDIA is seeking highly skilled software engineers to build AI inference systems that scale across multi-GPU, multi-node, and multi-cloud deployments. You will architect high-performance inference stacks, optimize kernels and compilers, and collaborate across inference, compiler, scheduling, and performance teams to push the frontier of accelerated AI computing.
Responsibilities include contributing features to vLLM, benchmarking, and driving MLPerf Inference submissions, with emphasis on