Senior GPU Runtime & HPC Software Engineer

NVIDIA

California (MO)

On-site

USD 152,000 - 287,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA is seeking a researcher to design, implement and maintain highly optimized communication runtimes for deep learning frameworks and HPC interfaces on GPU clusters. You will work with CUDA and Linux, contributing to MPI/OpenSHMEM and related system software to enable GPU-to-GPU and GPU-to-component interactions.

Candidate will develop proofs-of-concept, explore new programming models, and help advance hardware features while collaborating across multinational teams in a fast-paced

Qualifications

  • M.S./Ph.D. in CS/CE or equivalent experience.
  • 5+ years of relevant experience.
  • Excellent C/C++ programming and debugging skills.
  • Strong experience with Linux.
  • Expert understanding of computer system architecture and operating systems.
  • Experience with parallel programming interfaces and communication runtimes.
  • Ability and flexibility to work and communicate effectively in a multi-national, multi-time-zone corporate environment.

Responsibilities

  • Design, implement and maintain highly-optimized communication runtimes for Deep Learning frameworks (e.g. NCCL for TensorFlow/PyTorch) and HPC programming interfaces (e.g. UCX for MPI/OpenSHMEM) on GPU clusters.
  • Participate in and contribute to parallel programming interface specifications like MPI/OpenSHMEM.
  • Design, implement and maintain system software that enables interactions among GPUs and interactions between GPUs and other system components.
  • Create proofs-of-concepts to evaluate and motivate extensions in programming models, new designs in runtimes and new features in hardware.

Skills

C/C++ programming
Linux
Parallel programming
Debugging
Communication in multinational teams
System architecture knowledge

Education

M.S./Ph.D. in CS/CE

Job description

NVIDIA is seeking a researcher to design, implement and maintain highly optimized communication runtimes for deep learning frameworks and HPC interfaces on GPU clusters. You will work with CUDA and Linux, contributing to MPI/OpenSHMEM and related system software to enable GPU-to-GPU and GPU-to-component interactions.

Candidate will develop proofs-of-concept, explore new programming models, and help advance hardware features while collaborating across multinational teams in a fast-paced

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior GPU Communications Engineer — HPC/AI
Senior GPU Communications Engineer — HPC/AI

NVIDIA AI • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Competitive salary
Benefits
Senior GPU HPC Engineer - CUDA Optimizations & Equity
Senior GPU HPC Engineer - CUDA Optimizations & Equity

NVIDIA • Arizona

On-site
USD 184,000 - 357,000
Senior Software Architect - Deep Learning and HPC Communications
Senior Software Architect - Deep Learning and HPC Communications

NVIDIA • Durham (NC)

On-site
USD 120,000 - 160,000
Senior GPU HPC Cluster Engineer — Equity Eligible
Senior GPU HPC Cluster Engineer — Equity Eligible

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Senior GPU HPC Software Engineer: MPI/UCX & NCCL
Senior GPU HPC Software Engineer: MPI/UCX & NCCL

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Comprehensive benefits package
Equity opportunities
Senior HPC Middleware Engineer — Equity
Senior HPC Middleware Engineer — Equity

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Senior GPU AI & Quant HPC Engineer
Senior GPU AI & Quant HPC Engineer

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 287,500
Equity
Benefits
Senior DL/HPC Communications Architect – GPU Networking
Senior DL/HPC Communications Architect – GPU Networking

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity
Benefits
Senior HPC Middleware Engineer - MPI/RDMA & Performance
Senior HPC Middleware Engineer - MPI/RDMA & Performance

NVIDIA • Boulder (CO)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior System Software Engineer – GPU & HPC, Equity
Senior System Software Engineer – GPU & HPC, Equity

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 357,000