AI Compute Infra Engineer — Linux & HPC

NVIDIA Corporation

Santa Clara, Northern (CA, KY)

Hybrid

USD 124,000 - 196,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

NVIDIA in Santa Clara, CA is seeking an AI Compute Engineer to join its Infrastructure Specialists team. You will deploy, manage and validate AI Compute/HPC infrastructure in Linux environments for new and existing customers, with occasional travel up to 20%.

As the domain expert, you will participate in planning calls through implementation, hand over documentation, and share feedback with internal teams to improve workflows.

Qualifications

  • 4+ years providing in-depth support and deployment services; solving problems for hardware and software products.

Responsibilities

  • Deploying, managing, and validating AI Compute/HPC infrastructure in Linux-based environments for new and existing customers.
  • Be the domain expert with customers during planning calls through implementation.
  • Handover-related documentation and knowledge transfers required to support customers as they begin rolling out some of the most sophisticated systems in the world.
  • Provide feedback to internal teams such as opening bugs, documenting workarounds, and suggesting improvements.

Skills

Linux system administration
Cluster management
Scripting: Bash, Python, Ansible
Networking & HPC concepts
Customer-facing / interpersonal skills
Problem solving & troubleshooting
Scheduling systems (SLURM, LSF, UGE)
GPU/HPC concepts

Education

Bachelor's degree in Computer Science / Electrical Engineering or equivalent

Tools

BCM (Base Command Manager)
Kubernetes
InfiniBand
MPI
GPFS/Lustre storage
NVIDIA GPU platforms

Job description

NVIDIA in Santa Clara, CA is seeking an AI Compute Engineer to join its Infrastructure Specialists team. You will deploy, manage and validate AI Compute/HPC infrastructure in Linux environments for new and existing customers, with occasional travel up to 20%.

As the domain expert, you will participate in planning calls through implementation, hand over documentation, and share feedback with internal teams to improve workflows.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Compute Systems Engineer – HPC & Linux Infra
AI Compute Systems Engineer – HPC & Linux Infra

NVIDIA • Santa Clara (CA)

On-site
USD 124,000 - 196,000
Senior AI Compute Engineer - HPC Infra Architect
Senior AI Compute Engineer - HPC Infra Architect

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 184,000 - 288,000
Equity and benefits
Senior AI Compute Engineer - HPC Infra & Linux
Senior AI Compute Engineer - HPC Infra & Linux

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 148,000 - 287,500
Senior AI Compute Engineer — AI Infrastructure Lead
Senior AI Compute Engineer — AI Infrastructure Lead

NVIDIA • United States

Remote
USD 148,000 - 236,000
Senior AI Compute Engineer - Field Deployments & HPC
Senior AI Compute Engineer - Field Deployments & HPC

NVIDIA • Georgia

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior AI Compute Field Engineer
Senior AI Compute Field Engineer

NVIDIA • California (MO)

Hybrid
USD 184,000 - 357,000
Equity
Benefits
Senior AI Compute Deployment Engineer
Senior AI Compute Deployment Engineer

NVIDIA • Washington

On-site
USD 184,000 - 357,000
Equity and Benefits
Senior AI Compute Engineer - Field Deployment Lead
Senior AI Compute Engineer - Field Deployment Lead

NVIDIA • New Jersey

On-site
USD 184,000 - 357,000
Equity
Benefits
AI Infrastructure Engineer — GPU Compute Lead
AI Infrastructure Engineer — GPU Compute Lead

NVIDIA Corporation • Town of Italy (NY)

Hybrid
USD 74,000 - 168,000
Senior AI Compute Engineer — Forward-Deployed HPC
Senior AI Compute Engineer — Forward-Deployed HPC

NVIDIA Gruppe • California (MO)

On-site
USD 184,000 - 357,000
Equity
Benefits