An application made for this job — a tailored resume and cover letter that speak straight to the posting.
NVIDIA is seeking a Senior Deep Learning Software Engineer to scale automated inference and deployment for generative AI models in Santa Clara, CA. You will train and deploy state-of-the-art LLMs and diffusion models, using the Torch 2.0 stack to extract graph representations and optimize inference with advanced parallelism and kernels.
Work across teams to push performance, profiling GPU kernels and refining deployment strategies for TRT and related tools.
NVIDIA is seeking a Senior Deep Learning Software Engineer to scale automated inference and deployment for generative AI models in Santa Clara, CA. You will train and deploy state-of-the-art LLMs and diffusion models, using the Torch 2.0 stack to extract graph representations and optimize inference with advanced parallelism and kernels.
Work across teams to push performance, profiling GPU kernels and refining deployment strategies for TRT and related tools.