Senior Solutions Architect - AI Infrastructure

NVIDIA

Santa Clara (CA)

On-site

USD 184,000 - 287,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA is looking for an experienced professional in GPU cluster design to work with Cloud Partners on innovative architectures. Responsibilities include guiding design and addressing performance issues while providing documentation and customer support.

The ideal candidate will have 8+ years of experience in GPU or HPC clusters, and a degree in a related field. This position offers a competitive salary range of 184,000 to 356,500 USD based on experience and level, along with equity and benefits.

Qualifications

  • 8+ years of experience in cluster design, validation, and issue resolution, specifically on GPU and HPC clusters.
  • Proven expertise in designing large-scale distributed systems or AI clusters.
  • Ability to translate sophisticated engineering concepts into customer-ready documentation.

Responsibilities

  • Partner with NVIDIA Cloud Partners for GPU cluster design and process information.
  • Guide cluster design while weighing design principles against situational limitations.
  • Provide hands-on support for debugging issues related to cluster design.

Skills

Cluster design
GPU and HPC expertise
Documentation skills
Customer communication
Problem-solving

Education

BS, MS, or PhD in Computer Science, Electrical Engineering, or related field

Job description

NVIDIA is building the world’s most groundbreaking and innovative accelerated computing platforms for AI and HPC. Because of our work, scientists, researchers, and engineers can push the boundaries of what’s possible. We pioneered a supercharged form of computing that powers everything from breakthrough AI research to the world’s fastest supercomputers.

What You’ll Be Doing
  • Partner with NVIDIA Cloud Partners in GPU cluster design and networking and convey architecture and optimal process information for building next‑generation architectures.
  • Guide NVIDIA Cloud Partners in cluster design, weighing design principles but also complex, situational limitations to make the most performant and supportable GPU clusters possible.
  • Work closely with NVIDIA Cloud Partners to ensure successful first deployments with new products, including new network architectures and topologies.
  • Feedback customer/field perspectives on cluster design and workflows back to engineering teams designing internal clusters.
  • Perform hands‑on work to assist NVIDIA Cloud Partners debugging issues relating to cluster design, configuration, and performance employing internal engineering expertise and known bugs.
  • Support NPI customer deployments with new GPU/Networking architectures.
What We Need To See
  • BS, MS, or PhD in Computer Science, Electrical Engineering, Computer Engineering, Physics, or related field (or equivalent experience).
  • 8+ years of experience in cluster design, validation, and issue resolution, specifically on GPU and HPC clusters.
  • Proven expertise in designing large-scale distributed systems, AI clusters, or HPC infrastructure.
  • Ability to translate sophisticated engineering concepts into customer‑ready documentation, diagrams, and reference material.
  • Expertise in driving customer/partner issues to a close with product and engineering teams.
  • Ability to handle multi-functional communications across customer, product team, support team, engineering team, etc.
Ways To Stand Out From The Crowd
  • Experience leading large‑scale AI Factory or HPC cluster bring‑ups or builds.
  • Hands‑on experience with NVIDIA products including, but not limited to, GPUs, NVLink, NVIDIA Networking, etc.; specifically debugging issues that occur during deployment on NVLink, etc.
  • Knowledge of NCCL, MPI, IMEX, NMX, and collectives in distributed training as it pertains to cluster designs.
  • External customer facing skill‑set and background.
  • Effective time management and capability to balance multiple tasks and customers while thinking creatively to debug and solve problems.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD – 287,500 USD for Level 4, and 224,000 USD – 356,500 USD for Level 5. You will also be eligible for equity and benefits.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Solutions Architect - AI Infrastructure
Senior Solutions Architect - AI Infrastructure

NVIDIA • New York (NY)

On-site
USD 184,000 - 357,000
Senior Solutions Architect - AI Infrastructure
Senior Solutions Architect - AI Infrastructure

NVIDIA • California (MO)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Solutions Architect, NVIDIA Cloud Partners - Telco
Senior Solutions Architect, NVIDIA Cloud Partners - Telco

NVIDIA • Seattle (WA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Inclusive work environment
Senior Solutions Architect, NVIDIA Cloud Partners - Telco
Senior Solutions Architect, NVIDIA Cloud Partners - Telco

NVIDIA • Town of Texas (WI)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Solutions Architect, NVIDIA Cloud Partners - Telco
Senior Solutions Architect, NVIDIA Cloud Partners - Telco

NVIDIA • New Jersey

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Solutions Architect, NVIDIA Cloud Partners - Telco
Senior Solutions Architect, NVIDIA Cloud Partners - Telco

NVIDIA • Colorado

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Solutions Architect, AI Hyperscalers
Senior Solutions Architect, AI Hyperscalers

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Equity and benefits
Senior Solutions Architect, AI Infrastructure, Senior Solutions Architect, AI Infrastructure
Senior Solutions Architect, AI Infrastructure, Senior Solutions Architect, AI Infrastructure

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Equities and benefits
Remote work options
Senior Solutions Architect, NVIDIA Cloud Partners
Senior Solutions Architect, NVIDIA Cloud Partners

NVIDIA • Town of Texas (WI)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior Solutions Architect, Generative AI
Senior Solutions Architect, Generative AI

NVIDIA • California (MO)

Hybrid
USD 184,000 - 357,000