Senior Cluster Networking Architect for GPU AI HPC

NVIDIA

Austin (TX)

On-site

USD 184,000 - 357,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA in Austin, TX is seeking a Senior Networking Engineer to own the Kubernetes networking architecture for GPU clusters at multi-thousand-node scale. You will design, operate, and scale overlay networks and gateways across clouds.

Bring 6+ years in systems or networks, deep Kubernetes/CNI expertise (Calico preferred), and strong Go/Python/C skills. You will debug complex, distributed networks and collaborate across time zones. Equity and benefits are included.

Qualifications

  • 6+ years of professional experience in systems, network or infrastructure software engineering.
  • Deep command of Kubernetes networking architecture and CNI standards, with production experience operating Calico strongly preferred.
  • Proficiency designing and maintaining modern mesh and VPN networking topologies - Tailscale, WireGuard or equivalent.
  • Strong Linux networking fundamentals: routing, netfilter and iptables/nftables, packet marking, network namespaces, and how these interact with container runtimes.
  • Demonstrated ability to debug distributed network problems at scale - packet capture, tracing, and correlating behaviour across many hosts to find a single root cause.

Responsibilities

  • Own and evolve the Kubernetes networking architecture for GPU clusters running at multi-thousand-node scale.
  • Design, operate and scale the overlay network - CNI, mesh and VPN topologies (Tailscale, WireGuard), and the gateways that connect control and data planes.
  • Design, operate and scale the L7 gateways/load balancers/tunnels (Envoy, Cloudflare).
  • Find and eliminate scale ceilings: packet loss under load, control-plane saturation, IP address management exhaustion, and failure modes visible only at scale.
  • Build scale-test environments and validation suites to catch networking regressions before they reach production.

Skills

Kubernetes networking
CNI
Calico
Tailscale
WireGuard
Go
Python
C
Linux networking
Troubleshooting distributed networks

Education

BS/MS in Computer Science or Electrical Engineering

Tools

Envoy
Cloudflare
Calico
Cilium

Job description

NVIDIA in Austin, TX is seeking a Senior Networking Engineer to own the Kubernetes networking architecture for GPU clusters at multi-thousand-node scale. You will design, operate, and scale overlay networks and gateways across clouds.

Bring 6+ years in systems or networks, deep Kubernetes/CNI expertise (Calico preferred), and strong Go/Python/C skills. You will debug complex, distributed networks and collaborate across time zones. Equity and benefits are included.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cluster Networking Engineer for GPU AI Superclusters
Senior Cluster Networking Engineer for GPU AI Superclusters

NVIDIA • Westford (MA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Cluster Networking Architect for GPU Superclusters
Senior Cluster Networking Architect for GPU Superclusters

NVIDIA • Durham (NC)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior GPU Cluster Networking Architect — Multi-Cloud Scale
Senior GPU Cluster Networking Architect — Multi-Cloud Scale

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior GPU Cluster Networking Engineer
Senior GPU Cluster Networking Engineer

Socket.dev • North Carolina

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Kubernetes Networking Engineer for GPU Clusters
Senior Kubernetes Networking Engineer for GPU Clusters

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior GPU Cluster Networking Architect
Senior GPU Cluster Networking Architect

NVIDIA • United States

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Software Engineer - Cluster Networking
Senior Software Engineer - Cluster Networking

NVIDIA AI • Durham (NC)

On-site
USD 180,000 - 240,000
Equity
Benefits
Senior Network Architect for 10k+ GPU HPC Clusters
Senior Network Architect for 10k+ GPU HPC Clusters

AMD • San Jose (CA)

Hybrid
USD 150,000 - 190,000
Hybrid work model
AMD benefits
Senior Software Engineer - Cluster Networking
Senior Software Engineer - Cluster Networking

NVIDIA • Durham (NC)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Software Engineer - Cluster Networking
Senior Software Engineer - Cluster Networking

NVIDIA • Westford (MA)

On-site
USD 184,000 - 357,000
Equity
Benefits