Senior Kubernetes Runtime Lead - GPU AI Platform

NVIDIA Corporation

Santa Clara (CA)

On-site

USD 272,000 - 431,000

Full time

8 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Equity
Comprehensive benefits

Job summary

NVIDIA Corporation is seeking a technical leader to head the Runtime Engineering team for the NVIDIA Kubernetes Engine (NKE). You will drive the lifecycle of tenant workload clusters, ensure reliability, scalability, and security across the platform, and coordinate with cross-functional teams to deliver a production-grade Kubernetes runtime.

You will manage architecture decisions for networking, storage, and GPU resource partitioning, and contribute to open sources.

Qualifications

  • 12+ years in designing and delivering large-scale distributed software systems.
  • 5+ years of people-management leading software teams.
  • Experience bridging runtime, networking, and security across orgs.

Responsibilities

  • Oversee build, implementation and reliability of cluster configurations for NKE tenant workloads.
  • Lead a team coordinating the container runtime stack: AICR, GPU management operator, DCGM.
  • Drive architecture decisions for cluster networking, storage, and GPU resource partitioning.
  • Define cluster hardening standards, RBAC models, and multi-tenancy boundaries.
  • Collaborate with platform, infra, and cybersecurity teams to integrate capabilities.
  • Build tooling for AICR lifecycle management: provisioning, upgrades, drift detection.

Skills

Leadership
Kubernetes
Distributed systems
Security & compliance
RBAC & pod security
API design
Cross-org collaboration

Education

BS/MS in Computer Science or related field

Tools

Cluster API
kubeadm
NVIDIA DCGM/AI Container Runtime (AICR)

Job description

NVIDIA Corporation is seeking a technical leader to head the Runtime Engineering team for the NVIDIA Kubernetes Engine (NKE). You will drive the lifecycle of tenant workload clusters, ensure reliability, scalability, and security across the platform, and coordinate with cross-functional teams to deliver a production-grade Kubernetes runtime.

You will manage architecture decisions for networking, storage, and GPU resource partitioning, and contribute to open sources.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Kubernetes Runtime Engineering Lead — Multi-Tenant GPU Platform
Kubernetes Runtime Engineering Lead — Multi-Tenant GPU Platform

NVIDIA • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Senior Manager, Kubernetes Runtime Engineering
Senior Manager, Kubernetes Runtime Engineering

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Equity
Comprehensive benefits
Senior Manager, Kubernetes Runtime Engineering
Senior Manager, Kubernetes Runtime Engineering

NVIDIA • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Senior Kubernetes Runtime & Release Engineer - GPU Cloud
Senior Kubernetes Runtime & Release Engineer - GPU Cloud

NVIDIA Corporation • California (MO), Northern (KY)

Hybrid
USD 184,000 - 357,000
Equity
Benefits
Senior Kubernetes Runtime & Release Engineer
Senior Kubernetes Runtime & Release Engineer

NVIDIA • United States

Remote
USD 184,000 - 357,000
Equity
Benefits
Lead, GPU Kubernetes Runtime & Open-Source Strategy
Lead, GPU Kubernetes Runtime & Open-Source Strategy

NVIDIA Gruppe • Santa Clara (CA)

Hybrid
USD 208,000 - 380,000
Equity
Benefits package
Senior Kubernetes Runtime & Release Engineer — GPU Cloud
Senior Kubernetes Runtime & Release Engineer — GPU Cloud

NVIDIA Corporation • Kansas

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Systems Software Engineer - GPU-Driven Kubernetes
Senior Systems Software Engineer - GPU-Driven Kubernetes

NVIDIA AI • Seattle (WA)

On-site
USD 180,000 - 260,000
Equity
Benefits
Senior Kubernetes Node Lifecycle Architect
Senior Kubernetes Node Lifecycle Architect

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Principal Kubernetes & AI Infrastructure Engineer
Principal Kubernetes & AI Infrastructure Engineer

NVIDIA • Durham (NC)

On-site
USD 272,000 - 431,000
Equity
Benefits