Senior Network Architect – HPC GPU Data Center

Advanced Micro Devices

San Jose (CA)

Hybrid

USD 150,000 - 190,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Advanced Micro Devices (AMD) seeks a Senior Network Engineer to architect and operate high‑performance backend networks for large GPU clusters, including 10,000+ GPUs. You will own the fabric from GPU servers to leaf-spine, optimize RoCEv2 Ethernet fabrics, and model capacity for AI/HPC workloads.

You will collaborate with AI, data center, storage, and security teams, lead incident response, and mentor engineers while driving scalable, observable networking solutions.

Qualifications

  • Significant experience designing, deploying, and operating production data center networks for AI/GPU/HPC environments.
  • Experience building backend networks for GPU clusters containing ~10,000+ GPUs or similar hyperscale systems.
  • Deep knowledge of data center networking fundamentals and modern leaf-spine architectures.
  • Hands-on experience with RDMA, RoCEv2, PFC, DCQCN, QoS, and related technologies.

Responsibilities

  • Architect, deploy, and operate high-performance back-end networks for large GPU clusters.
  • Own end-to-end network path from GPU servers/NICs through leaf-spine fabric.
  • Design RoCEv2 Ethernet fabrics and scalable topologies; model capacity and plan growth.
  • Lead incident response, root-cause analysis, and optimization for network workloads.
  • Collaborate across AI, data center, storage, and security teams to prevent bottlenecks.

Skills

RDMA
RoCEv2
Data center networking
Leaf-spine
GPU clusters
BGP/ECMP
EVPN/VXLAN
Prometheus
Grafana
Junos OS
NVIDIA/AMD ROCm RCCL

Education

Bachelor's or Master's in Computer Engineering

Tools

Juniper Junos OS
Prometheus
Grafana
RDMA tooling

Job description

Advanced Micro Devices (AMD) seeks a Senior Network Engineer to architect and operate high‑performance backend networks for large GPU clusters, including 10,000+ GPUs. You will own the fabric from GPU servers to leaf-spine, optimize RoCEv2 Ethernet fabrics, and model capacity for AI/HPC workloads.

You will collaborate with AI, data center, storage, and security teams, lead incident response, and mentor engineers while driving scalable, observable networking solutions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Network Architect for 10k+ GPU HPC Clusters
Senior Network Architect for 10k+ GPU HPC Clusters

AMD • San Jose (CA)

Hybrid
USD 150,000 - 190,000
Hybrid work model
AMD benefits
Senior Network Engineer – GPU Cluster Networking
Senior Network Engineer – GPU Cluster Networking

AMD • San Jose (CA)

Hybrid
USD 150,000 - 190,000
Hybrid work model
AMD benefits
Senior Network Engineer – GPU Cluster Networking
Senior Network Engineer – GPU Cluster Networking

Advanced Micro Devices • San Jose (CA)

Hybrid
USD 150,000 - 190,000
Senior Data Center AI GPU Architect (Hybrid)
Senior Data Center AI GPU Architect (Hybrid)

Advanced Micro Devices • Northern (KY)

Hybrid
USD 150,000 - 210,000
Senior Data Center Architect – Cloud & AI Solutions
Senior Data Center Architect – Cloud & AI Solutions

Advanced Micro Devices • Austin (TX)

On-site
USD 180,000 - 240,000
Datacenter GPU Platform Systems Engineer (HPC/AI)
Datacenter GPU Platform Systems Engineer (HPC/AI)

Advanced Micro Devices • Santa Clara (CA)

On-site
USD 130,000 - 180,000
AI/HPC Cluster Architect - Scalable Data Center Design
AI/HPC Cluster Architect - Scalable Data Center Design

Advanced Micro Devices, Inc. • Austin (TX)

On-site
USD 140,000 - 190,000
AMD benefits
Equal opportunity employer
Visa sponsorship not available
Senior Network Architect — Global HPC Data Center Networks
Senior Network Architect — Global HPC Data Center Networks

Together AI • San Francisco (CA)

On-site
USD 190,000 - 270,000
Equity
Health insurance
Competitive pay
Senior Data Center Networking Architect - Ethernet & HPC AI
Senior Data Center Networking Architect - Ethernet & HPC AI

NVIDIA • Pennsylvania

On-site
USD 148,000 - 288,000
Equity
Benefits
Senior Manager, Data Center GPU Applications
Senior Manager, Data Center GPU Applications

Advanced Micro Devices • Santa Clara (CA)

On-site
USD 180,000 - 240,000