Distinguished Engineer, End-to-End Scaling Performance Architecture

NVIDIA Corporation

Santa Clara (CA)

On-site

USD 320,000 - 489,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

NVIDIA Corporation seeks a Distinguished Engineer to guide long-term performance strategy from single processors to multi-die, multi-GPU and multi-node platforms. You will drive architectural direction across DRAM, NVLink, and C2C interconnects, translating workloads into architectural requirements and prioritizing investments.

You will mentor teams, shape end-to-end performance targets, and align architecture with software, systems, and applications across multiple product generations.

Qualifications

  • 18+ years of relevant industry or academic experience.
  • Experience setting architecture direction for complex, high-performance systems.
  • Deep understanding of system performance, including DRAM behavior, high-bandwidth fabrics like NVLink, and C2C interconnects.
  • Ability to connect workload behavior to architecture choices and measurable outcomes.

Responsibilities

  • Define multi-generation strategy for application scaling across DRAM, NVLink, C2C and software stack.
  • Translate AI/HPC workloads into architectural requirements and performance targets.
  • Collaborate with teams to align architecture decisions with product roadmaps and technology investments.
  • Mentor system performance architects and communicate trade-offs to executives.

Skills

Architectural thinking
Performance modeling
Cross-functional leadership

Education

MSEE/MSCE/PhD or equivalent

Job description

We are looking for a Distinguished Engineer to join NVIDIA's architecture organization and help define how future accelerated computing systems scale from a single processor to multi-die, multi-GPU, and multi-node platforms! In this role, we will rely on you to set long-term performance strategy across applications, systems, and architecture, including DRAM, NVLink, and chip-to-chip (C2C) interconnects. We focus this role on architectural direction and application outcomes. We need someone who can identify where data movement, communication, memory behavior, topology, and compute limit scaling, then turn those insights into priorities that guide multiple product generations. Our domain teams own detailed implementation and delivery. We will count on you to align their decisions around a shared end-to-end strategy so local improvements create meaningful system-level gains.

What You Will Be Doing

Ask you to define the multi-generation strategy for application scaling across DRAM, NVLink, C2C, compute, and the supporting software stack. Translate the behavior of important AI, HPC, and accelerated computing applications into architectural requirements, performance targets, and investment priorities. Rely on you to build a clear view of how bottlenecks shift as workloads scale across dies, GPUs, nodes, model sizes, data sets, and communication patterns. Evaluate system-level trade-offs across bandwidth, latency, capacity, topology, coherence, power, area, cost, programmability, and resiliency. Use your leadership to establish common workload scenarios, scaling metrics, models, and decision frameworks so architecture teams can compare proposals against application outcomes. Count on you to identify architectural discontinuities and emerging technology opportunities early enough to shape product and technology decisions. We will partner with you to align DRAM, NVLink, C2C, GPU, CPU, system, and software architects around shared performance limits and high-value opportunities. You will work with application, framework, compiler, runtime, modeling, and post-silicon teams to connect measured behavior with future architecture choices. Value your clear recommendations to senior technical and business leaders, including assumptions, sensitivities, risks, and expected impact. We will also ask you to mentor system performance architects, strengthen technical communities across teams, generate sustained intellectual property, and help influence the direction of large-scale accelerated computing.

What We Need To See

MSEE, MSCE, PhD, or equivalent experience in Electrical Engineering, Computer Engineering, Computer Science, or a related field. 18+ years of relevant industry or academic experience, including experience setting architecture direction for complex, high-performance systems. Deep understanding of system performance and scaling, including interactions among DRAM behavior, high-bandwidth fabrics such as NVLink, and C2C communication. Strong application-level intuition, including the ability to connect workload algorithms, parallelism, communication, locality, and data movement to architecture choices and measurable outcomes. We value experience with workload characterization, analytical or simulation-based performance modeling, bottleneck analysis, and architecture trade-off evaluation. We look for a record of identifying cross-domain opportunities that may not be visible when teams optimize individual components separately. We need demonstrated ability to create and advance a multi-generation technical strategy through influence across silicon, systems, software, and application teams. We value clear communication and sound judgment in ambiguous technical areas, with the ability to explain complex system trade-offs to specialists and executive leaders. We also look for experience mentoring senior engineers into broader architecture leadership roles and building strong technical communities.

Ways To Stand Out from the crowd

If you have shaped product or technology roadmaps around application-level scaling needs, helped architecture teams build a quantitative understanding of bottlenecks across memory, interconnect, compute, and software, or identified high-impact trade-offs early enough to guide product, architecture, or technology investment. We will also value examples where your work delivered measurable end-to-end improvements in performance, efficiency, or scaling for priority applications.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 320,000 USD - 488,750 USD. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until August 1, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Distinguished Engineer, Power Architecture
Distinguished Engineer, Power Architecture

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 320,000 - 489,000
Equity
Benefits
Senior GPU Architect - Performance and Yield Optimization
Senior GPU Architect - Performance and Yield Optimization

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior GPU System Architect
Senior GPU System Architect

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Equity compensation
Benefits package
Hybrid work model
Distinguished Engineer - Memory Architecture
Distinguished Engineer - Memory Architecture

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 320,000 - 489,000
Senior Data Center System Architect
Senior Data Center System Architect

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Principal System Power Management and Performance Architect
Principal System Power Management and Performance Architect

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

On-site
USD 232,000 - 368,000
Senior GPU Memory System Architect
Senior GPU Memory System Architect

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Datacenter Product Architect
Datacenter Product Architect

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 168,000 - 322,000
Principal GPU Memory Architect
Principal GPU Memory Architect

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Equity
Benefits
Senior Accelerated Computing Architect
Senior Accelerated Computing Architect

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits