Senior HPC & GPU Cluster Architect

sfcompute

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Generous equity grant
Visa Sponsorships
Retirement matching
Medical, dental & vision
Time off
Parental leave
Daily lunch
Unlimited office book budget

Job summary

San Francisco Compute seeks an experienced HPC infrastructure engineer to design and deploy high‑performance GPU clusters worldwide. You’ll participate in on‑call rotations, deploy new environments, and drive automation to scale operations.

You’ll mentor junior engineers and help shape a customer‑focused culture. Ideal candidates have 5+ years building and operating HPC or GPU compute clusters, deep hardware knowledge, and strong IaC skills.

Qualifications

  • Five+ years of hands-on designing, architecting and scaling HPC or GPU compute clusters.
  • Deep understanding of server hardware fundamentals: GPUs, NICs, PCIe, memory, thermals, power.
  • Ability to debug performance and reliability across hardware, OS, drivers, networking – full-stack.
  • Excels at automating fleet operations (provisioning, monitoring, remediation) through infrastructure-as-code.
  • Produces solid operational documentation and runbooks; mentors junior engineers.

Responsibilities

  • Architect and deploy new GPU clusters globally; keep them running smoothly.
  • Participate in on-call rotation and respond to issues; improve automation.
  • Mentor junior engineers and shape team culture as an early contributor.
  • Collaborate with customers and internal teams to design custom solutions matched to workloads.

Skills

HPC cluster design
GPU compute
Linux systems
Automation / IaC
Mentoring
Travel willingness

Tools

Slurm
Kubernetes
KVM
QEMU
libvirt

Job description

San Francisco Compute seeks an experienced HPC infrastructure engineer to design and deploy high‑performance GPU clusters worldwide. You’ll participate in on‑call rotations, deploy new environments, and drive automation to scale operations.

You’ll mentor junior engineers and help shape a customer‑focused culture. Ideal candidates have 5+ years building and operating HPC or GPU compute clusters, deep hardware knowledge, and strong IaC skills.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior HPC & GPU Cluster Architect
Senior HPC & GPU Cluster Architect

The Consensus • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Visa sponsorships
401(k) retirement matching
Medical, dental & vision insurance
+2
Senior HPC & GPU Cluster Architect — Scale & Automate
Senior HPC & GPU Cluster Architect — Scale & Automate

San Francisco Compute Company • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Generous equity grant
Competitive salary
Visa sponsorship
+6
Lead HPC/GPU Cluster Architect | Automate & Scale
Lead HPC/GPU Cluster Architect | Automate & Scale

The San Francisco Compute Company • Boston (MA)

Hybrid
USD 140,000 - 200,000
Generous equity grant
Visa Sponsorships
Retirement matching
+5
Senior GPU Infrastructure Engineer — HPC & Clusters
Senior GPU Infrastructure Engineer — HPC & Clusters

Prime Intellect AI • San Francisco (CA)

On-site
USD 150,000 - 300,000
Senior HPC & GPU Compute Cluster Architect
Senior HPC & GPU Compute Cluster Architect

Electric Capital • San Francisco (CA)

Hybrid
USD 220,000 - 300,000
Generous equity grant
401(k) matching
Comprehensive medical, dental, and vision insurance
+3
Senior HPC GPU Compute Engineer (Hybrid SF)
Senior HPC GPU Compute Engineer (Hybrid SF)

The San Francisco Compute Company • San Francisco (CA)

Hybrid
USD 180,000 - 260,000
Generous equity grant
Retirement matching
Comprehensive medical, dental, and vision insurance
+3
Senior GPU Infrastructure Architect
Senior GPU Infrastructure Architect

Primeintellect • San Francisco (CA)

On-site
USD 150,000 - 300,000
Senior HPC Architect: At-Scale GPU Deployments & Automation
Senior HPC Architect: At-Scale GPU Deployments & Automation

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Health and wellness program
Equity
Senior HPC Systems Engineer: Slurm, GPU & Hybrid Clusters
Senior HPC Systems Engineer: Slurm, GPU & Hybrid Clusters

Parallel Works • United States

Hybrid
USD 180,000 - 230,000
Medical, vision, and dental coverage
401(k) with company match
Paid vacation & sick time
+1
Senior HPC Architect - GPU Compute, Equity Eligible
Senior HPC Architect - GPU Compute, Equity Eligible

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Equity
Inclusive work environment
Comprehensive benefits