Senior AI Infrastructure Engineer – Scale & Equity

Nvidia Corporation

Santa Clara (CA)

On-site

USD 184,000 - 356,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA is seeking an NCX Senior Engineer to join our AI Accelerator team, partnering with strategic customers to implement and enhance groundbreaking AI workloads. You will deliver hands-on technical assistance for advanced AI deployments, distributed systems, and ensure customers realize efficient performance from NVIDIA's AI platform across varied environments.

You will lead deployments on NCP/Neo Cloud, optimize training and inference, and guide partner teams using Kubernetes, containers, and

Qualifications

  • BS/MS/PhD in CS/EE or related field, or equivalent experience.
  • 8+ years in customer-facing technical roles such as Solutions Engineering, DevOps, SRE, or ML Infra, ideally in large cloud/service provider environments.
  • Strong expertise in Linux systems, distributed computing, Kubernetes, containers, and GPU scheduling on multi-tenant platforms.
  • Demonstrated AI/ML experience supporting large-scale training and inference workloads (LLMs, generative models) in production.
  • Solid programming skills in Python/Go, with hands-on experience using PyTorch or TensorFlow.

Responsibilities

  • Build and deploy custom AI solutions on NCP and Neo Cloud platforms, including distributed training, inference optimization, and MLOps pipelines on NVIDIA reference architectures.
  • Act as main technical contact for strategic NCPs, provide remote and on-site support, troubleshoot complex production problems, guide partner engineering teams.
  • Deploy and manage AI workloads across DGX Cloud, NCP data centers, and major CSP environments using Kubernetes, containers, and GPU scheduling systems aligned to NCP builds.
  • Profile and tune large-scale training and inference workloads on NCP platforms; implement observability and SLO/SLA monitoring; lead efforts to reduce latency, cost, and risk.
  • Implement and expand NVIDIA reference architectures on partner platforms, develop integrations with partner control planes and customer environments; ensure smooth API and data pipeline connectivity.
  • Build detailed implementation guides, runbooks, and post-mortem documentation codifying standard methodologies for NVIDIA AI workloads at scale on NCP platforms.

Skills

Linux systems
Distributed computing
Python
Go
Communication
Collaboration

Education

BS/MS/PhD in CS/EE or related field

Tools

Kubernetes
Docker
GPU scheduling
PyTorch
TensorFlow

Job description

NVIDIA is seeking an NCX Senior Engineer to join our AI Accelerator team, partnering with strategic customers to implement and enhance groundbreaking AI workloads. You will deliver hands-on technical assistance for advanced AI deployments, distributed systems, and ensure customers realize efficient performance from NVIDIA's AI platform across varied environments.

You will lead deployments on NCP/Neo Cloud, optimize training and inference, and guide partner teams using Kubernetes, containers, and

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Infrastructure Engineer — Equity Eligible
Senior AI Infrastructure Engineer — Equity Eligible

2100 NVIDIA USA • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits package
Competitive salaries
Senior AI Infrastructure Engineer - Kubernetes & Scale
Senior AI Infrastructure Engineer - Kubernetes & Scale

NVIDIA • Seattle (WA)

Hybrid
USD 184,000 - 357,000
Equity
Benefits
Senior MLOps Engineer — AI Infra & Scale Leader
Senior MLOps Engineer — AI Infra & Scale Leader

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits package
Senior AI Infrastructure & Applied ML Engineer
Senior AI Infrastructure & Applied ML Engineer

NVIDIA AI • Santa Clara (CA)

On-site
USD 200,000 - 322,000
NCX Senior Engineer
NCX Senior Engineer

2100 NVIDIA USA • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits package
Competitive salaries
Senior Full-Stack Engineer — AI Infra & GPU Cloud (Equity)
Senior Full-Stack Engineer — AI Infra & GPU Cloud (Equity)

Nvidia Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior AI Infra Systems Engineer - Equity
Senior AI Infra Systems Engineer - Equity

NVIDIA AI • Santa Clara (CA)

On-site
USD 180,000 - 260,000
Equity
Senior AI Infra Architect: Kubernetes at Scale (Equity)
Senior AI Infra Architect: Kubernetes at Scale (Equity)

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 224,000 - 357,000
Equity
Benefits
NCX Senior Engineer
NCX Senior Engineer

Nvidia Corporation • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Equity
Benefits
Senior MLOps Engineer - AI Infrastructure & Cloud
Senior MLOps Engineer - AI Infrastructure & Cloud

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits