Lead HPC Network Architect for AI Cloud Infra

Lambda

Santo Niño 1st

On-site

PHP 11,314,000 - 16,342,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health, dental, and vision coverage
401k with 2% company match
Wellness and commuter stipends
Flexible PTO

Job summary

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure delivering scalable, low-latency networking for tens of thousands of customers. This senior role focuses on building high-performance data center networks for AI/ML workloads and large-scale cloud platforms.

You will architect InfiniBand and Ethernet fabrics, define topology, mentor teams, and drive standards across multi-site environments while collaborating with compute and storage teams.

Qualifications

  • Proven experience (7+ years) architecting high-performance data center networks, preferably for HPC, AI/ML, or large-scale cloud infrastructure.
  • Deep expertise with InfiniBand (HDR/NDR) and advanced Ethernet fabrics, including RoCE and RDMA protocols.
  • Strong understanding of data center switching architectures, congestion control (PFC, ECN), QoS, and network virtualization technologies such as VXLAN and EVPN.
  • Skilled in designing for low-latency and high-throughput data paths, including GPU-to-GPU and storage traffic optimization.
  • Expertise with BGP-based fabric design, including eBGP underlays, MP-BGP EVPN, ECMP, route policy, and convergence.
  • Experience with SDN, APIs, automation, and network telemetry.
  • Strong optical networking competency, including transceiver technologies, fiber types and topologies, link budgets, breakout architectures, and WDM.
  • Experience building resilient, fault-tolerant high-performance network architectures with redundancy, failover, and high availability.
  • Excellent communication and leadership skills, capable of influencing technical decisions across diverse teams.
  • Strong ownership and can do attitude, self-starter who feels comfortable working in ambiguity.

Responsibilities

  • Architect high-performance networking solutions that power cloud platforms, with a focus on ultra-low-latency and high-bandwidth connectivity.
  • Define the network topology and architectural patterns for large-scale GPU clusters, storage backends, and multi-tenant environments.
  • Evaluate, benchmark, and select next-generation network technologies (e.g., InfiniBand NDR/XDR, RoCE, 400G-1.6T Ethernet, Ultra Ethernet, ESUN, OCI, etc) to meet AI workload requirements.
  • Develop and maintain network architecture standards, reference designs, and scalability roadmaps for multi-site and hybrid environments.
  • Partner with compute and storage architects to ensure seamless end-to-end data flow and fault tolerance.
  • Guide network automation strategies and tooling to enable efficient provisioning, telemetry, and operational visibility.
  • Mentor engineers and cross-functional teams on advanced network concepts, troubleshooting, and architectural best practices.

Skills

Leadership
Communication
Strategic planning
Ambiguity handling

Tools

InfiniBand HDR/NDR
RoCE
RDMA
VXLAN EVPN
BGP-based fabric
SDN & automation
Telemetry
WDM/optical tech

Job description

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure delivering scalable, low-latency networking for tens of thousands of customers. This senior role focuses on building high-performance data center networks for AI/ML workloads and large-scale cloud platforms.

You will architect InfiniBand and Ethernet fabrics, define topology, mentor teams, and drive standards across multi-site environments while collaborating with compute and storage teams.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior GPU HPC Systems Architect for AI Cloud
Senior GPU HPC Systems Architect for AI Cloud

Lambda • Santo Niño 1st

Hybrid
PHP 11,942,000 - 15,085,000
Health, dental, vision coverage foryou
Wellness and commuter stipends
401k Plan with 2% company match (USA)
+1
Staff HPC Network Architect
Staff HPC Network Architect

Lambda • Santo Niño 1st

On-site
PHP 11,314,000 - 16,342,000
Health, dental, and vision coverage
401k with 2% company match
Wellness and commuter stipends
+1
Staff HPC Systems Architect
Staff HPC Systems Architect

Lambda • Santo Niño 1st

On-site
PHP 11,942,000 - 15,085,000
Health, dental, vision coverage foryou
Wellness and commuter stipends
401k Plan with 2% company match (USA)
+1
Senior Cloud and AI Engineer
Senior Cloud and AI Engineer

WHR Global Consulting • Quezon City

On-site
PHP 1,339,000 - 2,678,000
Lead Cloud Deployment Engineer - AWS Infra & Automation
Lead Cloud Deployment Engineer - AWS Infra & Automation

mylo1 • Hinoba-an

On-site
PHP 1,327,000 - 2,123,000
AI-Driven Network Engineer - Cloud & On-Prem (Hybrid)
AI-Driven Network Engineer - Cloud & On-Prem (Hybrid)

WideNet Consulting Group • Boston

Hybrid
PHP 5,166,000 - 5,855,000
Health benefits
401K
Employee Assistance Program
+1
Remote Lead AI-First Cloud Architect
Remote Lead AI-First Cloud Architect

Caylent • Philippines

Remote
MXN 1,200,000 - 1,800,000
Medical Insurance
Generous holidays and PTO
Equipment & Office Stipend
+1
Senior HPC & GPU Cluster Engineer — Remote EU
Senior HPC & GPU Cluster Engineer — Remote EU

Verda • España

On-site
PHP 6,536,000 - 9,441,000
Equity included
Healthcare
Lunch
+1
GPU Architect
GPU Architect

NVIDIA Corporation • Hinoba-an

On-site
PHP 1,200,000 - 1,800,000
Data Center Engineer/Lead
Data Center Engineer/Lead

Atomic Recruitment SEA • Metro Manila

On-site
PHP 391,000 - 670,000