Senior Developer Technology Engineer - Windows AI Platform

NVIDIA Gruppe

Santa Clara (CA)

On-site

USD 184,000 - 287,500

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Equity options
Comprehensive benefits package

Job summary

NVIDIA Gruppe is looking for an experienced GPU Deployment Engineer to tackle end-to-end AI deployment challenges on the NVIDIA RTX AI platform. The role involves analyzing GPU-accelerated applications, improving user experiences, and collaborating with teams to influence next-gen GPU features.

The ideal candidate will have over 8 years of experience and expertise in C/C++, Python, and CUDA. Strong problem-solving skills and a passion for AI technology are essential. An inclusive, supportive environment awaits you at NVIDIA.

Qualifications

  • 8+ years of professional experience in local GPU deployment, profiling, and optimization.
  • Experience with open-source LLM and GenAI software.
  • Strong background in AI technology advancements.

Responsibilities

  • Work on end-to-end AI GPU deployment challenges.
  • Analyze and improve GPU-accelerated applications.
  • Conduct hands-on training and develop sample code.

Skills

C/C++
Python
GPU profiling
Optimization techniques
Interpersonal skills
Problem-solving

Education

Bachelor's or Master's degree in Computer Science, Engineering, or related field

Tools

CUDA
NVIDIA Nsight
Windows OS
Vulkan
TensorRT

Job description

At NVIDIA, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As a NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

Responsibilities
  • Work closely with internal engineering and product teams and external app developers on solving local end-to-end AI GPU deployment challenges on the NVIDIA RTX AI platform.
  • Apply powerful profiling and debugging tools for analyzing most demanding GPU-accelerated end-to-end AI applications to detect insufficient GPU utilization resulting in suboptimal runtime performance.
  • Conduct hands‑on trainings, develop sample code and host presentations to give good guidance on efficient end-to-end AI deployment targeting optimal runtime performance on NVIDIA ARM‑based SoCs.
  • Improve Windows LLM & GenAI user experience on NVIDIA RTX by working on feature and performance enhancements of OSS software, including but not limited to projects like GGML, Llama.cpp, Ollama, ONNX Runtime.
  • Collaborate with GPU driver and architecture teams as well as NVIDIA research to influence next generation GPU features by providing real‑world workflows and giving feedback on partner and customer needs.
  • Providing technical leadership and mentorship to junior engineers, encouraging an inclusive and high‑performing team environment.
Qualifications
  • A proven track record of 8+ years of professional experience in local GPU deployment, profiling and optimization.
  • Bachelor's or Master's degree or equivalent experience in Computer Science, Engineering, or a related field.
  • Strong proficiency in C/C++, Python, software design, programming techniques.
  • Familiarity with and development experience on the Windows operating system.
  • Experience working with open‑source LLM and GenAI software.
  • Experience with CUDA and NVIDIA's Nsight GPU profiling and debugging suite.
  • Some travel is required for conferences and for on‑site visits with external partners.
  • Strong problem‑solving skills and the ability to work both independently and collaboratively in a fast‑paced environment.
  • Excellent interpersonal and communication skills and a passion for keeping track with the latest advancements in AI technology.
Additional Qualifications
  • Experience with GPU‑accelerated AI inference driven by NVIDIA APIs, specifically cuDNN, CUTLASS, TensorRT.
  • Confirmed expert knowledge in Vulkan and / or DX12.
  • Detailed knowledge of the latest generation GPU architectures.
  • Experience with AI deployment on NPUs and ARM architectures.
Benefits

The base salary range for Level 4 is $184,000 – $287,500 and for Level 5 is $224,000 – $356,500. You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until May 25, 2026.

EEO Statement

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Developer Technology Engineer - Windows AI Platform
Senior Developer Technology Engineer - Windows AI Platform

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Comprehensive benefits
Diversity and inclusion
Senior Developer Technology Engineer - Edge Agentic AI
Senior Developer Technology Engineer - Edge Agentic AI

Segment (Twilio) • Santa Clara (CA)

On-site
USD 152,000 - 287,500
Equity
Benefits package
Senior Developer Technology Engineer - Edge Agentic AI
Senior Developer Technology Engineer - Edge Agentic AI

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Competitive salary
Senior Developer Technology Engineer - Agentic AI
Senior Developer Technology Engineer - Agentic AI

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior Software Engineer - GPU Local AI Platforms
Senior Software Engineer - GPU Local AI Platforms

NVIDIA • Westford (MA)

On-site
USD 224,000 - 431,250
Senior Software Engineer - GPU Local AI Platforms
Senior Software Engineer - GPU Local AI Platforms

NVIDIA • Austin (TX)

On-site
USD 224,000 - 431,250
Equity
Benefits
Senior Software Engineer - GPU Local AI Platforms
Senior Software Engineer - GPU Local AI Platforms

NVIDIA • Durham (NC)

On-site
USD 224,000 - 356,500
Equity
Benefits
Senior Systems Software Engineer - GPU Performance at Scale
Senior Systems Software Engineer - GPU Performance at Scale

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity
Benefits
Senior Software Engineer - GPU Local AI Platforms
Senior Software Engineer - GPU Local AI Platforms

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 224,000 - 431,000
Senior Developer Technology Engineer - AI
Senior Developer Technology Engineer - AI

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000