Senior Customer Success Engineer - DGX Cloud

NVIDIA Corporation

Santa Clara (CA)

Remote

USD 200,000 - 322,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

NVIDIA Corporation is seeking a Senior Customer Success Engineer for the DGX Cloud organization. You will design and implement distributed cloud infrastructure, contribute code, and build tooling to automate workflows and manage GPU capacity.

You will collaborate with internal research, product, finance, and operations teams to align infrastructure roadmaps with business needs and drive efficiency. The role emphasizes deep technical execution, with opportunities to write code, build tooling, and

Qualifications

  • BS or MS in Computer Science, Engineering or related field, or equivalent experience.
  • 12+ years of experience designing and building distributed systems and cloud infrastructure.
  • Production code written in Golang, Java, C, C++, Python, or Rust.
  • Experience with Kubernetes and/or distributed task scheduling.
  • Strong background in Infrastructure, Networking, Storage, and DevOps tooling.
  • Experience deploying AI/ML workloads at scale.
  • Strong communication and relationship-building skills across teams.

Responsibilities

  • Design and implement distributed cloud infrastructure at scale across compute, storage, networking, and GPU capacity management across IaaS, PaaS, and SaaS.
  • Contribute production code when needed and codify patterns into reusable tools and playbooks.
  • Build and maintain agentic tooling to automate workflows and resource management.
  • Analyze DGX Cloud workloads to forecast demand and drive capacity efficiency with cross-functional teams.
  • Present roadmaps, decisions, and demos to stakeholders and leadership to align infrastructure strategy.

Skills

Distributed systems
Cloud infrastructure
Golang
Java
C++
Python
Rust
Kubernetes
DevOps scripting
Communication

Education

BS/MS CS or related

Tools

Kubernetes
Terraform
Docker
Git

Job description

The DGX Cloud organization bridges customer success and cloud infrastructure engineering, partnering directly with NVIDIA's internal research and product teams to accelerate AI workload development. As a Customer Success Engineer, you'll embed deeply with internal customers — gaining a thorough understanding of their applications and translating that knowledge into architectural guidance, best practices, and hands‑on solutions. This role sits at the unique intersection of solutions architecture and platform strategy: you'll write code, build tooling, and help shape NVIDIA's GPU capacity management from the inside. Working across Engineering, Product, Finance, and Operations, you'll connect infrastructure roadmaps to business needs in a way that directly influences how NVIDIA's most advanced AI teams move faster. If you thrive where deep technical work meets high‑stakes collaboration, this role was built for you.

What you’ll be doing:
  • Design and implement distributed cloud infrastructure at scale — spanning compute, storage, networking, and GPU capacity management across IaaS, PaaS, and SaaS models zendesk Partner with internal research and product teams to understand workloads from both a technology and business perspective, providing architectural guidance that drives their success
  • Contribute code directly when needed to move projects forward, and codify working patterns into tools, playbooks, and building blocks that others can reuse
  • Build and maintain agentic tooling to automate operational workflows and infrastructure resource management
  • Analyze the DGX Cloud ecosystem to understand current customer demand and future capacity needs, driving infrastructure efficiency initiatives in partnership with Engineering, Finance, and Product
  • Present technical roadmaps, architecture decisions, and demos to internal stakeholders and NVIDIA leadership, driving cross‑functional consensus on infrastructure strategy
What we need to see:
  • BS or MS in Computer Science, Engineering, or a related field, or equivalent experience.
  • 12+ years of experience designing and building distributed systems and cloud infrastructure, with demonstrated experience in GPU capacity management for high‑performance computing
  • Demonstrated ability to write production code in Golang, Java, C, C++, Python, or Rust
  • Experience with Kubernetes and/or distributed task scheduling
  • Strong background in Infrastructure, Networking, Storage, and DevOps scripting/tooling
  • Experience deploying AI/ML workloads at scale
  • Strong communication and relationship‑building skills, with a demonstrated ability to drive cross‑functional consensus and align stakeholders across departments

NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High‑Performance Computing, and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction inventions from artificial intelligence to autonomous cars. NVIDIA is looking for phenomenal people like you to help us accelerate the next wave of artificial intelligence. NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward‑thinking and dedicated people in the world working for us.

#LI-Remote Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 200,000 USD - 322,000 USD. You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August14,2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer.

As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry.

Learn more about NVIDIA.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Customer Success Engineer - DGX Cloud
Senior Customer Success Engineer - DGX Cloud

NVIDIA • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Senior Software Engineer, Capacity Management - DGX Cloud
Senior Software Engineer, Capacity Management - DGX Cloud

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 168,000 - 270,000
Senior Full-Stack Lead Engineer
Senior Full-Stack Lead Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits
Principal Software Engineer - DGX Cloud
Principal Software Engineer - DGX Cloud

NVIDIA Gruppe • Seattle (WA)

On-site
USD 272,000 - 431,000
Equity
Benefits package
Senior Software Engineer, Capacity Management - DGX Cloud
Senior Software Engineer, Capacity Management - DGX Cloud

NVIDIA • Santa Clara (CA)

On-site
USD 140,000 - 270,000
Software Platform Support Engineer - GPU Cloud
Software Platform Support Engineer - GPU Cloud

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 108,000 - 173,000
Principal Software Engineer - DGX Cloud
Principal Software Engineer - DGX Cloud

NVIDIA • Washington

On-site
USD 272,000 - 431,000
Principal Software Engineer, Distributed Systems Engineer - DGX Cloud
Principal Software Engineer, Distributed Systems Engineer - DGX Cloud

NVIDIA • Durham (NC)

On-site
USD 272,000 - 431,000
Equity
Benefits
Senior Software Engineer, Distributed Systems Engineer - DGX Cloud
Senior Software Engineer, Distributed Systems Engineer - DGX Cloud

NVIDIA • United States

Remote
USD 184,000 - 288,000
Senior DGX Cloud AI Infrastructure Software Engineer
Senior DGX Cloud AI Infrastructure Software Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits