Senior Software Engineer, NCCL

NVIDIA

California (MO)

On-site

USD 152,000 - 287,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA is seeking a researcher to design, implement and maintain highly optimized communication runtimes for deep learning frameworks and HPC interfaces on GPU clusters. You will work with CUDA and Linux, contributing to MPI/OpenSHMEM and related system software to enable GPU-to-GPU and GPU-to-component interactions.

Candidate will develop proofs-of-concept, explore new programming models, and help advance hardware features while collaborating across multinational teams in a fast-paced

Qualifications

  • M.S./Ph.D. in CS/CE or equivalent experience.
  • 5+ years of relevant experience.
  • Excellent C/C++ programming and debugging skills.
  • Strong experience with Linux.
  • Expert understanding of computer system architecture and operating systems.
  • Experience with parallel programming interfaces and communication runtimes.
  • Ability and flexibility to work and communicate effectively in a multi-national, multi-time-zone corporate environment.

Responsibilities

  • Design, implement and maintain highly-optimized communication runtimes for Deep Learning frameworks (e.g. NCCL for TensorFlow/PyTorch) and HPC programming interfaces (e.g. UCX for MPI/OpenSHMEM) on GPU clusters.
  • Participate in and contribute to parallel programming interface specifications like MPI/OpenSHMEM.
  • Design, implement and maintain system software that enables interactions among GPUs and interactions between GPUs and other system components.
  • Create proofs-of-concepts to evaluate and motivate extensions in programming models, new designs in runtimes and new features in hardware.

Skills

C/C++ programming
Linux
Parallel programming
Debugging
Communication in multinational teams
System architecture knowledge

Education

M.S./Ph.D. in CS/CE

Job description

Overview

NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. NVIDIA is looking for phenomenal people like you to help us accelerate the next wave of artificial intelligence.

Responsibilities
  • Design, implement and maintain highly-optimized communication runtimes for Deep Learning frameworks (e.g. NCCL for TensorFlow/PyTorch) and HPC programming interfaces (e.g. UCX for MPI/OpenSHMEM) on GPU clusters.
  • Participate in and contribute to parallel programming interface specifications like MPI/OpenSHMEM.
  • Design, implement and maintain system software that enables interactions among GPUs and interactions between GPUs and other system components.
  • Create proof-of-concepts to evaluate and motivate extensions in programming models, new designs in runtimes and new features in hardware.
Qualifications
  • M.S./Ph.D. in CS/CE or equivalent experience.
  • 5+ years of relevant experience.
  • Excellent C/C++ programming and debugging skills.
  • Strong experience with Linux.
  • Expert understanding of computer system architecture and operating systems.
  • Experience with parallel programming interfaces and communication runtimes.
  • Ability and flexibility to work and communicate effectively in a multi-national, multi-time-zone corporate environment.
Ways to stand out from the crowd
  • Deep understanding of technology and passion for what you do.
  • Experience with CUDA programming and NVIDIA GPUs.
  • Knowledge of high-performance networks like InfiniBand, iWARP, etc.
  • Experience with HPC applications.
  • Experience with Deep Learning Frameworks such as PyTorch, TensorFlow, etc.
  • Strong collaborative and interpersonal skills, specifically a proven ability to effectively guide and influence within a dynamic matrix environment.
Compensation and Benefits

NVIDIA offers highly competitive salaries and a comprehensive benefits package. The base salary ranges: 152,000 USD - 241,500 USD for Level 3 and 184,000 USD - 287,500 USD for Level 4, determined by location and experience. Eligible for equity and benefits.

Applications for this job will be accepted at least until July 21, 2026. This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

Equal Opportunity

NVIDIA is committed to fostering an inclusive work environment and is an equal opportunity employer. We value diversity and do not discriminate on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer, NCCL
Senior Software Engineer, NCCL

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Comprehensive benefits package
Equity opportunities
Senior Software Architect - Deep Learning and HPC Communications
Senior Software Architect - Deep Learning and HPC Communications

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior Software Architect - Deep Learning and HPC Communications
Senior Software Architect - Deep Learning and HPC Communications

NVIDIA • Westford (MA)

On-site
USD 224,000 - 357,000
Equity
Comprehensive benefits
Software Engineer, CUDA Deep Learning Systems
Software Engineer, CUDA Deep Learning Systems

NVIDIA AI • Santa Clara (TX)

On-site
USD 124,000 - 196,000
Equity
Senior Software Engineer, CUDA Deep Learning Systems
Senior Software Engineer, CUDA Deep Learning Systems

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity options
Comprehensive benefits package
Senior Software Engineer, CUDA Deep Learning Systems
Senior Software Engineer, CUDA Deep Learning Systems

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Equity and benefits
Senior Software Engineer, CUDA UMD - Graphs and GPU Sharing
Senior Software Engineer, CUDA UMD - Graphs and GPU Sharing

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Diversity and inclusion initiatives
Senior HPC AI Cluster Engineer
Senior HPC AI Cluster Engineer

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

On-site
USD 176,000 - 334,000
HPC Middleware Developer
HPC Middleware Developer

NVIDIA • Boulder (CO)

On-site
USD 152,000 - 288,000
Equity
Benefits
HPC Middleware Developer
HPC Middleware Developer

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000