Lead HPC Network Architect for AI Cloud

Socket.dev

San Jose (CA)

Hybrid

USD 180,000 - 260,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Health, dental, and vision coverage
Equity compensation
401k with company match
Flexible paid time off
Wellness and commuter stipends

Job summary

Lambda, The Superintelligence Cloud is seeking an experienced Senior Network Architect to design and optimize high-performance data center networks for AI workloads. You will define topologies for large GPU clusters, multi-tenant storage, and multi-site environments, while driving automation and fault tolerance.

You will lead architectural decisions, mentor teams, and collaborate across compute, storage, and networking groups to ensure ultra-low latency, high bandwidth, and reliable operation in

Qualifications

  • 7+ years of experience in architecting high-performance data center networks, preferably for HPC/AI/cloud.
  • Expertise with InfiniBand HDR/NDR and advanced Ethernet fabrics including RoCE and RDMA.
  • Strong understanding of data center switching architectures, congestion control, QoS, and VXLAN/EVPN systems.
  • Experience with BGP-based fabric design including MP-BGP EVPN, ECMP, and convergence.
  • Proficiency in SDN, APIs, automation, and network telemetry.
  • Cloud-scale, multi-site, fault-tolerant design with high availability.
  • Excellent communication and cross-functional leadership abilities.

Responsibilities

  • Architect high-performance networking solutions for cloud platforms with ultra-low latency and high bandwidth.
  • Define topology and patterns for large GPU clusters, storage backends, and multi-tenant environments.
  • Evaluate and select next-generation network technologies to meet AI workload requirements.
  • Develop standards, reference designs, and roadmaps for scalability across sites.
  • Collaborate with compute/storage teams to ensure end-to-end data flow and fault tolerance.
  • Lead network automation strategies and tooling for provisioning, telemetry, and visibility.
  • Mentor engineers and cross-functional teams on advanced network concepts and best practices.

Skills

InfiniBand HDR/NDR
RoCE
RDMA
eBGP/MP-BGP EVPN
VXLAN
Data center architectures
Networking performance optimization
Network automation
Telemetry
Networking leadership

Tools

NVIDIA BlueField
Broad and flexible network hardware

Job description

Lambda, The Superintelligence Cloud is seeking an experienced Senior Network Architect to design and optimize high-performance data center networks for AI workloads. You will define topologies for large GPU clusters, multi-tenant storage, and multi-site environments, while driving automation and fault tolerance.

You will lead architectural decisions, mentor teams, and collaborate across compute, storage, and networking groups to ensure ultra-low latency, high bandwidth, and reliable operation in

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Hybrid HPC Network Architect for AI Cloud
Hybrid HPC Network Architect for AI Cloud

Lambda • San Francisco (CA)

On-site
USD 180,000 - 240,000
Health, dental, and vision coverage
401k with company match (USA)
Wellness stipend
+2
Senior HPC Architect for AI Compute Platforms
Senior HPC Architect for AI Compute Platforms

Lambda • San Jose (CA)

On-site
USD 180,000 - 260,000
Health, dental, and vision coverage
Equity compensation
401k with 2% company match
+3
Hybrid HPC Systems Architect - GPU Cloud for AI
Hybrid HPC Systems Architect - GPU Cloud for AI

The Consensus • San Jose (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Cash compensation
Equity compensation
Health, dental and vision coverage
+1
Principal Network Architect, AI Cloud Backbone
Principal Network Architect, AI Cloud Backbone

Lambda • San Francisco (CA)

Hybrid
USD 180,000 - 290,000
Health, dental, and vision coverage
Wellness stipends
Commuter stipends
+2
Senior Network Architect for AI Cloud
Senior Network Architect for AI Cloud

Lambda Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 190,000 - 230,000
Health, dental, vision coverage
Wellness and commuter stipends
401k Plan with company match
+1
Senior Network Architect — AI Cloud Backbone
Senior Network Architect — AI Cloud Backbone

Lambda • United States

Hybrid
USD 180,000 - 210,000
Senior HPC Compute Architect for AI Cloud
Senior HPC Compute Architect for AI Cloud

Socket.dev • San Jose (CA)

On-site
USD 180,000 - 240,000
Senior HPC Architect: GPU Clusters & Liquid Cooling
Senior HPC Architect: GPU Clusters & Liquid Cooling

Neura Market • San Jose (CA)

Hybrid
USD 180,000 - 240,000
Health, dental, and vision
401k with company match
Wellness stipend
+1
Senior AI Cloud Solutions Engineer
Senior AI Cloud Solutions Engineer

Lambda Labs • United States

On-site
USD 180,000 - 260,000
Health, dental, and vision coverage
401k Plan with 2% company match
Wellness and commuter stipends
+1
Senior HPC Validation Engineer – AI Cloud Infrastructure
Senior HPC Validation Engineer – AI Cloud Infrastructure

Lambda • San Jose (CA)

On-site
USD 150,000 - 210,000
Health coverage
Dental coverage
Vision coverage
+4