Senior Solutions Architect - AI Infrastructure

NVIDIA

New York (NY)

On-site

USD 184,000 - 356,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA is looking for a senior cluster design engineer based in New York, NY. The role involves collaborating with NVIDIA Cloud Partners to design next-generation GPU clusters, providing support for initial deployments, and translating engineering concepts into customer-facing documentation.

Candidates should have a strong background in AI clusters and HPC infrastructure, with at least 8 years of experience. The position offers a competitive salary range and an inclusive work environment.

Qualifications

  • 8+ years of experience in cluster design, validation, and issue resolution.
  • Proven expertise in designing large-scale distributed systems or HPC infrastructure.
  • Strong ability to translate complex engineering concepts into documentation.

Responsibilities

  • Partner with NVIDIA Cloud Partners in designing GPU clusters.
  • Guide design principles and optimize process information.
  • Help debug issues in cluster design and performance.

Skills

Cluster design
Networking
AI clusters
HPC infrastructure
Customer communication

Education

BS, MS, or PhD in Computer Science, Electrical Engineering, Computer Engineering, Physics

Tools

NVIDIA products
NVIDIA Networking

Job description

NVIDIA is building the world’s most groundbreaking and innovative accelerated computing platforms for AI and HPC. Because of our work, scientists, researchers, and engineers can push the boundaries of what’s possible. We pioneered a supercharged form of computing that powers everything from breakthrough AI research to the world’s fastest supercomputers.

What You’ll Be Doing
  • Partner with NVIDIA Cloud Partners in GPU cluster design and networking and convey architecture and optimal process information for building next‑generation architectures.
  • Guide NVIDIA Cloud Partners in cluster design, weighing design principles but also complex, situational limitations to make the most performant and supportable GPU clusters possible.
  • Work closely with NVIDIA Cloud Partners to ensure successful first deployments with new products, including new network architectures and topologies.
  • Feedback customer/field perspectives on cluster design and workflows back to engineering teams designing internal clusters.
  • Perform hands‑on work to assist NVIDIA Cloud Partners debugging issues relating to cluster design, configuration, and performance employing internal engineering expertise and known bugs.
  • Support NPI customer deployments with new GPU/Networking architectures.
What We Need To See
  • BS, MS, or PhD in Computer Science, Electrical Engineering, Computer Engineering, Physics, or related field (or equivalent experience).
  • 8+ years of experience in cluster design, validation, and issue resolution, specifically on GPU and HPC clusters.
  • Proven expertise in designing large-scale distributed systems, AI clusters, or HPC infrastructure.
  • Ability to translate sophisticated engineering concepts into customer‑ready documentation, diagrams, and reference material.
  • Expertise in driving customer/partner issues to a close with product and engineering teams.
  • Ability to handle multi-functional communications across customer, product team, support team, engineering team, etc.
Ways To Stand Out From The Crowd
  • Experience leading large‑scale AI Factory or HPC cluster bring‑ups or builds.
  • Hands‑on experience with NVIDIA products including, but not limited to, GPUs, NVLink, NVIDIA Networking, etc.; specifically debugging issues that occur during deployment on NVLink, etc.
  • Knowledge of NCCL, MPI, IMEX, NMX, and collectives in distributed training as it pertains to cluster designs.
  • External customer facing skill‑set and background.
  • Effective time management and capability to balance multiple tasks and customers while thinking creatively to debug and solve problems.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD – 287,500 USD for Level 4, and 224,000 USD – 356,500 USD for Level 5. You will also be eligible for equity and benefits.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Solutions Architect - AI Infrastructure
Senior Solutions Architect - AI Infrastructure

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Senior Solutions Architect - AI Infrastructure
Senior Solutions Architect - AI Infrastructure

NVIDIA • California (MO)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Solutions Architect, Cluster Design and Architecture - Networking
Senior Solutions Architect, Cluster Design and Architecture - Networking

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Equity
Benefits
Senior Solutions Architect, NVIDIA Cloud Partners - Telco
Senior Solutions Architect, NVIDIA Cloud Partners - Telco

NVIDIA • Seattle (WA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Inclusive work environment
Senior Solutions Architect, NVIDIA Cloud Partners - Telco
Senior Solutions Architect, NVIDIA Cloud Partners - Telco

NVIDIA • New Jersey

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Solutions Architect, NVIDIA Cloud Partners - Telco
Senior Solutions Architect, NVIDIA Cloud Partners - Telco

NVIDIA • Colorado

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Solutions Architect, AI Hyperscalers
Senior Solutions Architect, AI Hyperscalers

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Equity and benefits
Senior Solutions Architect, NVIDIA Cloud Partners - Telco
Senior Solutions Architect, NVIDIA Cloud Partners - Telco

NVIDIA • Town of Texas (WI)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Solutions Architect, AI Hyperscalers
Senior Solutions Architect, AI Hyperscalers

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity options
Comprehensive benefits
Senior Solutions Architect, AI Infrastructure, Senior Solutions Architect, AI Infrastructure
Senior Solutions Architect, AI Infrastructure, Senior Solutions Architect, AI Infrastructure

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Equities and benefits
Remote work options