Principal Deep Learning Communication Architect

NVIDIA

Austin (TX)

On-site

USD 272,000 - 431,250

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA in Austin, Texas, is seeking a highly experienced professional to drive architecture leadership and develop next-generation communication libraries for their platforms. Candidates must possess a Ph.D. or M.S. in a relevant field with over 12 years of experience in high-performance computing. The role requires expertise in 3D parallelism and NVIDIA GPU architecture. A competitive salary range of $272,000 to $431,250 is offered, along with equity and benefits.

Qualifications

  • 12+ years of industry experience in high-performance computing or distributed deep learning.
  • Expert knowledge of NVIDIA GPU memory hierarchy (HBM3e/HBM4, L2 cache).
  • Hands-on experience developing within Megatron-Core, DeepSpeed, or JAX/XLA.

Responsibilities

  • Define the long‑term technical roadmap for communication libraries across platforms.
  • Lead development of next‑generation communication primitives and collective algorithms.
  • Collaborate with silicon architects to influence hardware specifications.

Skills

Deep understanding of 3D parallelism
Technical proficiency with NCCL, UCX, UCC, NVSHMEM, MPI
Advanced knowledge of TensorRT-LLM, vLLM, SGLang, NVIDIA Dynamo

Education

Ph.D. or M.S. in Computer Science or related field

Tools

CUDA programming models
RDMA
RoCE
low-level InfiniBand verbs

Job description

What You'll Be Doing
  • Architecture Leadership: Define the long‑term technical roadmap for communication libraries across NVIDIA’s next‑generation platforms. Ensure seamless scaling of models to clusters comprising hundreds of thousands of nodes.
  • AI Communication Library Design: Lead development of next‑generation communication primitives and collective algorithms. Optimize for heterogeneous interconnects such as NVLink, Spectrum‑X (Ethernet), and Quantum‑X (InfiniBand).
  • Application‑Communication Library Co‑Design: Partner with application developers to architect and implement specialized communication primitives. Ensure that AI and HPC libraries—including NCCL, NIXL, NVSHMEM, UCC, and UCX—evolve to meet the requirements of trillion‑parameter and Agentic AI.
  • Hardware/Software Co‑Design: Collaborate with silicon architects and software engineers to influence hardware specifications for next‑generation networking, ensuring they meet the evolving demands of trillion‑parameter LLMs and Agentic AI.
  • Quantitative Modeling: Develop high‑fidelity analytical models and simulators to predict system behavior under emerging workloads.
What We Need To See
  • Ph.D. or M.S. in Computer Science, Electrical Engineering, or related field (or equivalent experience), with 12+ years of industry experience in high‑performance computing (HPC) or distributed deep learning.
  • Parallelism Expertise: Deep understanding of 3D parallelism (Data, Tensor, Pipeline) and advanced strategies including Context Parallelism, Expert Parallelism, and Zero Redundancy Optimizer (ZeRO) variants.
  • Technical Proficiency: Deep technical proficiency with NCCL, UCX, UCC, NVSHMEM, or MPI. Experience with RDMA, RoCE, and low‑level InfiniBand verbs is required.
  • Inference & Serving: Advanced knowledge of high‑throughput inference engines and schedulers, specifically TensorRT-LLM, vLLM, SGLang, and NVIDIA Dynamo.
  • GPU Architecture: Expert knowledge of the NVIDIA GPU memory hierarchy (HBM3e/HBM4, L2 cache) and CUDA programming models.
Ways To Stand Out
  • Framework Development: Hands‑on experience developing within Megatron‑Core, DeepSpeed, or JAX/XLA, with an understanding of how these frameworks interact with low‑level communication runtimes.
  • Significant upstream contributions to major open‑source projects (e.g., PyTorch Distributed, KServe, or Ray).
  • Proven track record of deploying and optimizing models on NVIDIA platforms or similar rack‑scale systems.
  • Strong portfolio of patents or papers in top‑tier systems/architecture venues (e.g., ISCA, ASPLOS, NeurIPS, SC).

Base salary range: 272,000 USD – 431,250 USD. You will also be eligible for equity and benefits.

Applications will be accepted until April 18, 2026.

NVIDIA is committed to fostering a diverse work environment and is an equal‑opportunity employer. We do not discriminate on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status, or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Deep Learning Communication Architect
Principal Deep Learning Communication Architect

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 272,000 - 432,000
Competitive salary
Equity options
Generous benefits package
Senior Deep Learning Framework Communications Engineer
Senior Deep Learning Framework Communications Engineer

NVIDIA • Austin (TX)

On-site
USD 152,000 - 242,000
Senior Deep Learning Communication Architect
Senior Deep Learning Communication Architect

NVIDIA • Austin (TX)

On-site
USD 184,000 - 288,000
Generous benefits package
Senior Deep Learning Framework Communications Engineer
Senior Deep Learning Framework Communications Engineer

NVIDIA • Westford (MA)

On-site
USD 152,000 - 242,000
Equity
Benefits
Senior Deep Learning Communication Architect
Senior Deep Learning Communication Architect

NVIDIA • Seattle (WA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Deep Learning Communication Architect
Senior Deep Learning Communication Architect

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior Deep Learning Communication Architect
Senior Deep Learning Communication Architect

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Comprehensive benefits package
Principal Deep Learning Communication Architect
Principal Deep Learning Communication Architect

NVIDIA Corporation • Austin (TX)

On-site
USD 272,000 - 432,000
Senior Software Architect - Deep Learning and HPC Communications
Senior Software Architect - Deep Learning and HPC Communications

NVIDIA • Austin (TX)

On-site
USD 184,000 - 288,000
Distinguished Software Architect - Deep Learning and HPC Communications
Distinguished Software Architect - Deep Learning and HPC Communications

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 320,000 - 488,750