NVIDIA is seeking a Performance Engineer in Indiana, United States, to enhance communication libraries crucial for HPC. This role involves performance characterization on multi-GPU clusters and evaluating proof-of-concepts while collaborating with dynamic teams. The ideal candidate has an M.S. or PhD in Computer Science, 3+ years in parallel programming, and a strong understanding of computer systems. NVIDIA is committed to diversity, offering competitive salaries and a supportive work environment.
Qualifications
3+ years of experience with parallel programming and at least one communication runtime.
Experience conducting benchmarking on large scale HPC clusters.
Good understanding of computer system architecture and HW-SW interactions.
Responsibilities
Conduct performance characterization on multi-GPU and multi-node clusters.
Study interaction of libraries with hardware and software components.
Evaluate proof-of-concepts and conduct trade-off analysis.
Skills
Parallel programming
Performance benchmarking
Debugging across HW/SW stack
Proficient in Python
Familiar with Kubernetes
Education
M.S. or PhD in Computer Science or related field
Tools
MPI
NCCL
Docker
SLURM
Job description
NVIDIA is seeking a Performance Engineer in Indiana, United States, to enhance communication libraries crucial for HPC. This role involves performance characterization on multi-GPU clusters and evaluating proof-of-concepts while collaborating with dynamic teams. The ideal candidate has an M.S. or PhD in Computer Science, 3+ years in parallel programming, and a strong understanding of computer systems. NVIDIA is committed to diversity, offering competitive salaries and a supportive work environment.