GPU Network Architect for AI/HPC Clusters

CyberCoders

Santa Clara (CA)

Hybrid

USD 200,000 - 250,000

Full time

7 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Health Benefits
401k
Relocation assistance

Job summary

CyberCoders in Sunnyvale, CA is seeking a hands-on GPU Network Engineer to design, build and operate the high-speed fabrics that connect our GPU clusters. This hybrid role requires on-site collaboration and a strong focus on reliability and scalability.

You will deploy InfiniBand (NDR/XDR) and RoCEv2 interconnects, design CLOS/ECMP networks, and work with Kubernetes and bare-metal environments to move data efficiently at scale.

Qualifications

  • 2+ years in data center or HPC networking with GPU/AI clusters.
  • Hands-on GPU cluster networking design and operations.
  • Deep experience with InfiniBand (NDR/XDR) and RoCEv2.
  • Experience with Kubernetes and multi-host networking.
  • Familiarity with Netbox, Netconf, and IaC tooling.

Responsibilities

  • Design and deploy east-west network fabrics for GPU-to-GPU, rack-to-rack, and cluster-to-cluster communication.
  • Build and operate InfiniBand (NDR/XDR) and RoCEv2 interconnect fabrics as the primary transport for GPU workloads.
  • Implement CLOS/ECMP architectures for high-bandwidth, low-latency data movement across GPU clusters.

Skills

Data center networking
HPC networking
GPU cluster environments
InfiniBand (NDR/XDR)
RoCEv2 interconnects
Multi-host networking

Tools

Netbox
Netconf
Ansible
Terraform
KVM

Job description

CyberCoders in Sunnyvale, CA is seeking a hands-on GPU Network Engineer to design, build and operate the high-speed fabrics that connect our GPU clusters. This hybrid role requires on-site collaboration and a strong focus on reliability and scalability.

You will deploy InfiniBand (NDR/XDR) and RoCEv2 interconnects, design CLOS/ECMP networks, and work with Kubernetes and bare-metal environments to move data efficiently at scale.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Network Engineer
GPU Network Engineer

CyberCoders • Santa Clara (CA)

Hybrid
USD 200,000 - 250,000
Health Benefits
401k
Relocation assistance
GPU Network Engineer
GPU Network Engineer

Blue Signal Search • Santa Clara (CA)

On-site
USD <240,000
Senior Network Architect for 10k+ GPU HPC Clusters
Senior Network Architect for 10k+ GPU HPC Clusters

AMD • San Jose (CA)

Hybrid
USD 150,000 - 190,000
Hybrid work model
AMD benefits
Senior GPU Network Architect for AI Clusters
Senior GPU Network Architect for AI Clusters

Blue Signal Search • Santa Clara (CA)

On-site
USD <240,000
Cluster Design
Cluster Design

Blue Signal Search • San Francisco (CA)

On-site
USD 150,000 - 230,000
Senior Network Architect – HPC GPU Data Center
Senior Network Architect – HPC GPU Data Center

Advanced Micro Devices • San Jose (CA)

Hybrid
USD 150,000 - 190,000
Senior GPU Networking Architect for AI & HPC
Senior GPU Networking Architect for AI & HPC

Experis • Morrisville (NC)

On-site
USD 117,000 - 131,000
Medical Plans
Vision Plan
HSA
+5
GPU Networking Engineer for Large-Scale AI Fabric
GPU Networking Engineer for Large-Scale AI Fabric

Thinking Machines Lab Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 350,000 - 475,000
Health, dental, and vision benefits
Unlimited PTO
Parental leave
+1
Senior Network Engineer — AI Infra & HPC Fabric Expert
Senior Network Engineer — AI Infra & HPC Fabric Expert

Nscale • Houston (TX)

On-site
USD 150,000 - 210,000
Competitive benefits package
Flexible paid time off
Parental leave
+1
Hybrid HPC Network Architect for AI Cloud
Hybrid HPC Network Architect for AI Cloud

Lambda • San Francisco (CA)

On-site
USD 180,000 - 240,000
Health, dental, and vision coverage
401k with company match (USA)
Wellness stipend
+2