Senior HPC Architect for AI Compute Platforms

Lambda

San Jose (CA)

On-site

USD 180,000 - 260,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health, dental, and vision coverage
Equity compensation
401k with 2% company match
Wellness stipend
Commuter stipend
Flexible paid time off

Job summary

Lambda, The Superintelligence Cloud, is seeking a Senior Compute Platform Architect to lead the design of scalable AI/ML compute infrastructure across bare metal and cloud deployments. You will define architecture standards, drive performance tuning, and guide hardware/software tradeoffs for density and efficiency.

You will collaborate with product and engineering teams, assess CPU/GPU/accelerator technologies, and mentor cross-functional staff.

Qualifications

  • Proven track record architecting 10k+ GPU HPC or cloud compute platforms.
  • Deep knowledge of CPU/GPU architectures and memory hierarchies.
  • Experience with high-bandwidth, low-latency fabrics (NVLink, InfiniBand, RoCE).

Responsibilities

  • Architect scalable compute platforms optimized for AI/ML, simulation, and high-throughput workloads.
  • Develop compute system standards and design patterns for performance and maintainability.
  • Evaluate CPU/GPU/accelerator tech and guide tradeoffs for density, power, cooling, and TCO.
  • Collaborate with product and engineering to map workloads to platform capabilities across bare metal and cloud deployments.
  • Translate ambiguous business needs into measurable platform requirements and architecture decisions.
  • Define compute platform roadmaps and reference designs for hardware, firmware, and rack/clusters.
  • Lead validation and performance characterization during new platform introductions.
  • Mentor engineers on compute performance tuning, sizing, and architectural decisions.

Skills

CPU architectures
GPU architectures
Memory hierarchies
High-bandwidth fabrics
Resource scheduling
Thermal optimization
Power optimization
Compute lifecycle management
Orchestration layers
Analytical skills
Communication skills
Ownership / self-starter

Tools

Slurm
Kubernetes
GPU virtualization

Job description

Lambda, The Superintelligence Cloud, is seeking a Senior Compute Platform Architect to lead the design of scalable AI/ML compute infrastructure across bare metal and cloud deployments. You will define architecture standards, drive performance tuning, and guide hardware/software tradeoffs for density and efficiency.

You will collaborate with product and engineering teams, assess CPU/GPU/accelerator technologies, and mentor cross-functional staff.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead HPC Network Architect for AI Cloud
Lead HPC Network Architect for AI Cloud

Socket.dev • San Jose (CA)

Hybrid
USD 180,000 - 260,000
Health, dental, and vision coverage
Equity compensation
401k with company match
+2
Senior HPC Compute Architect for AI Cloud
Senior HPC Compute Architect for AI Cloud

Socket.dev • San Jose (CA)

On-site
USD 180,000 - 240,000
Senior HPC Validation Engineer – AI Cloud Infrastructure
Senior HPC Validation Engineer – AI Cloud Infrastructure

Lambda • San Jose (CA)

On-site
USD 150,000 - 210,000
Health coverage
Dental coverage
Vision coverage
+4
Senior HPC Architect: GPU Clusters & Liquid Cooling
Senior HPC Architect: GPU Clusters & Liquid Cooling

Neura Market • San Jose (CA)

Hybrid
USD 180,000 - 240,000
Health, dental, and vision
401k with company match
Wellness stipend
+1
Senior HPC Systems Validation Engineer – AI Cloud Infra
Senior HPC Systems Validation Engineer – AI Cloud Infra

Neura Market • San Jose (CA)

Hybrid
USD 180,000 - 230,000
Hybrid HPC Systems Architect - GPU Cloud for AI
Hybrid HPC Systems Architect - GPU Cloud for AI

The Consensus • San Jose (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Cash compensation
Equity compensation
Health, dental and vision coverage
+1
Staff Software Engineer — AI Cloud Compute Platform
Staff Software Engineer — AI Cloud Compute Platform

Lambda Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 300,000
Health, dental, vision coverage
Wellness stipend
401k with 2% company match
+2
Senior HPC Platform Hardware Engineer - Hybrid (San Jose)
Senior HPC Platform Hardware Engineer - Hybrid (San Jose)

Neura Market • San Jose (CA)

On-site
USD 180,000 - 280,000
Cash & equity compensation
Health, dental, and vision coverage
Wellness stipends
+1
Staff Compute Platform Engineer – AI Cloud (Remote)
Staff Compute Platform Engineer – AI Cloud (Remote)

Applied Methods Ltd • Bellevue (WA), Northern (KY)

Hybrid
USD 190,000 - 260,000
Health coverage
Dental coverage
Vision coverage
+2
Senior GPU HPC Architect for AI/ML Compute Platforms
Senior GPU HPC Architect for AI/ML Compute Platforms

Jobtailor • California (MO)

On-site
USD 180,000 - 260,000