Solutions Architect, Infrastructure

NVIDIA Corporation

Washington (Washington County)

On-site

USD 152,000 - 288,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

NVIDIA Corporation seeks an Infrastructure Solutions Architect to lead bring‑up of Data Center GPU and networking platforms for hyperscalers and large enterprises.

You will bridge product strategy, cloud engineering, and customer deployments, driving end‑to‑end execution and cross‑functional collaboration for scale‑up and platform readiness across worldwide users.

Qualifications

  • BS/MS/PhD in Electrical/Computer Engineering, Computer Science, Physics, or similar, or equivalent experience.
  • 4+ years experience in Solutions Architecture, Infrastructure Engineering, or similar technical roles.
  • Hands-on experience with bring-up and validation of large-scale NVIDIA GPU platforms, including multi-GPU and multi-node architectures.
  • Understanding of high-performance networking technologies (RDMA, congestion control, high-bandwidth interconnects) and their role in distributed AI workloads.
  • Proficiency with NVIDIA system software stacks: CUDA, NCCL, NVSwitch/NVLink, driver behavior, and performance tuning.
  • Proficiency with Linux system tools for identifying issues and evaluating system performance (e.g., dmesg, journalctl, lspci, numactl, ethtool, iostat, perf, nvidia-smi, top/htop, ipmitool, container‑level tooling).
  • Understanding of server hardware architecture, including PCIe topologies, firmware, NUMA, BIOS/UEFI, power/thermal envelopes, and memory/subsystem behavior.
  • Understanding of BMC/IPMI/Redfish for remote management and out-of-band debugging during bring-up.
  • Strong Linux fundamentals across drivers, kernel subsystems, cgroups, containers, and node-level performance analysis.
  • Ability to identify performance bottlenecks at cluster, node, accelerator, network, or application layer.

Responsibilities

  • Lead end-to-end execution for Hyperscaler customers to bring NVIDIA GPU platforms to market at scale.
  • Drive partnerships with Product teams to understand roadmap, co-define metrics, and align directions.
  • Influence across Product, Engineering, Sales, Operations, and CSP customers to unblock scale‑up.
  • Analyze deployment and performance data to identify health trends, bottlenecks, and risks.
  • Solve challenging technical problems involving GPUs, networking, drivers, containers, firmware, and distributed systems.
  • Deliver streamlined executive‑level status updates on progress, risks, and decisions.
  • Collaborate with Product and Engineering to improve platform design and workflows.

Skills

Solutions Architecture
Infrastructure Engineering
NVIDIA GPUs
RDMA networking
Linux systems
Driver tuning
nvidia-smi
Performance analysis
Remote management
Container tooling

Education

Electrical/Computer Engineering, Computer Science, Physics

Tools

CUDA
NCCL
NVSwitch/NVLink
lspci
numactl
ethtool
perf
ipmitool

Job description

Do you thrive on taking a strategic product from launch to go‑to‑market at scale across the world’s largest customers? NVIDIA is looking for an Infrastructure Solutions Architect to lead deployment and bring‑up of our next‑generation Data Center GPUs and networking platforms. As part of the NVIDIA Solutions Architecture team, we navigate uncharted technical and organizational spaces — serving as the bridge between early platform readiness, cloud engineering teams, product strategy, and large‑scale customer deployments. We are looking for Solution Architects to combine hands‑on infrastructure expertise with multi‑functional leadership to accelerate adoption of NVIDIA technologies across worldwide cloud hosting providers and large enterprise environments.

What You’ll Be Doing:

Lead end‑to‑end execution for Hyperscaler customers to rapidly bring NVIDIA Data Center GPU and networking platforms to market at scale. Drive strategic partnership and alignment with Product teams to understand roadmap intent, co‑define critical metrics, and ensure unified direction across technical, sales, and leadership organizations. Influence without authority across Product, Engineering, Sales, Operations, and CSP customers, driving clarity, alignment, and unblock paths for scale‑up. Analyze deployment and performance data, identifying product health trends, system bottlenecks, and operational risks. Solve challenging technical problems involving GPUs, networking, drivers, containers, firmware, and distributed system interactions. Deliver streamlined executive‑level communication on status, risks, progress, and required decisions. Collaborate with Product and Engineering, enabling future improvements in platform design, validation, and operational workflows.

What We Need to See:

BS/MS/PhD in Electrical/Computer Engineering, Computer Science, Physics, or similar, or equivalent experience. 4+ years experience in Solutions Architecture, Infrastructure Engineering, or similar technical roles. Hands‑on experience with bring‑up and validation of large‑scale NVIDIA GPU platforms, including multi‑GPU and multi‑node architectures. Understanding of high‑performance networking technologies (e.g., RDMA, congestion control, high‑bandwidth interconnects) and their role in distributed AI workloads. Familiarity with NVIDIA system software stacks: CUDA, NCCL, NVSwitch/NVLink, driver behavior, and performance tuning. Proficiency with Linux systems tools for identifying issues and evaluating system performance, such as: dmesg, journalctl, lspci, numactl, ethtool, iostat, perf, nvidia-smi, top/htop, ipmitool, container‑level tooling, and related utilities. Understanding of server hardware architecture, including PCIe topologies, system firmware, NUMA, BIOS/UEFI configuration, power/thermal envelopes, and memory/subsystem behavior. Understanding of BMC/IPMI/Redfish for remote management, hardware health monitoring, and out‑of‑band debugging during early‑stage bring‑up. Strong Linux fundamentals across drivers, kernel subsystems, cgroups, containers, and node‑level performance analysis. Ability to identify performance bottlenecks at the cluster, node, accelerator, network, or application layer.

Ways to Stand Out from the Crowd:

Outstanding interpersonal skills and the ability to build clarity and direction across diverse, fast paced technical teams. Knowledgeof Computeand networking infrastructure (e.g., Instance types, networking primitives, high‑performance communication paths etc)atHyperscalersorCloudServiceProviders.Demonstrated leadership resolving multi‑team infrastructure challenges across engineering, product and customer groups. A consistentrecordoftakingGPUorinfrastructureproductsfrom pilottohigh‑volume deploymentinlargedatacenter environments.Familiaritywithmoderndeep learningLLM architectures, and distributed training/inference challenges at scale.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level3, and 184,000 USD - 287,500 USD for Level4. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until August30, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law. NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Solutions Architect, Infrastructure
Solutions Architect, Infrastructure

NVIDIA Gruppe • Redmond (WA)

On-site
USD 152,000 - 242,000
Equity and benefits
Diverse work environment
Solutions Architect, Infrastructure
Solutions Architect, Infrastructure

NVIDIA AI • Seattle (WA)

On-site
USD 152,000 - 288,000
Equity compensation
Benefits package
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

NVIDIA • Austin (TX)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

NVIDIA • Town of Texas (WI)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

NVIDIA • California (MO)

On-site
USD 152,000 - 288,000
Equity
Benefits
Solutions Architect - NVIDIA Cloud Partners
Solutions Architect - NVIDIA Cloud Partners

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 184,000 - 357,000
Infrastructure Solutions Architect - OEM Deployment
Infrastructure Solutions Architect - OEM Deployment

NVIDIA • California (MO)

On-site
USD 124,000 - 242,000
Equity
Benefits
Travel up to 50%
Infrastructure Solutions Architect - OEM Deployment
Infrastructure Solutions Architect - OEM Deployment

NVIDIA • Massachusetts

On-site
USD 124,000 - 242,000
Equity
Benefits
Infrastructure Solutions Architect - OEM Deployment
Infrastructure Solutions Architect - OEM Deployment

NVIDIA • Santa Clara (CA)

On-site
USD 124,000 - 242,000
Equity
Benefits
Infrastructure Solutions Architect - OEM Deployment
Infrastructure Solutions Architect - OEM Deployment

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 124,000 - 242,000
Equity
Benefits