Senior HPC & GPU Cluster Architect — Scale & Automate

San Francisco Compute Company

San Francisco (CA)

Hybrid

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Generous equity grant
Competitive salary
Visa sponsorship
Retirement matching up to 4%
100% medical, dental, and vision insurance
Unlimited paid time off
Paid parental leave
Daily lunch coverage
Unlimited office book budget

Job summary

San Francisco Compute Company is seeking an experienced professional to manage and deploy GPU clusters. You'll work to ensure optimal performance, participate in on-call rotations, and help mentor junior engineers within our team.

This role requires a deep understanding of HPC and GPU systems, coupled with a passion for automation and infrastructure as code. Additional perks include competitive salary and generous equity options.

Qualifications

  • 5+ years experience with designing and scaling HPC or GPU compute clusters.
  • Deep understanding of server hardware fundamentals.
  • Comfortable debugging across hardware and software layers.
  • Excited about automating fleet operations.
  • Ability to mentor junior engineers.

Responsibilities

  • Architect and deploy new GPU clusters globally.
  • Participate in on-call rotation and fix issues as they arise.
  • Lean into automation for large-scale deployments.
  • Shape company culture and mentor junior engineers.

Skills

HPC or GPU compute cluster experience
Server hardware fundamentals
Performance debugging
Infrastructure as code
Strong operational documentation
Mentoring junior engineers

Tools

Linux systems administration
Schedulers and orchestration systems (Slurm, Kubernetes)
Virtualization technologies (KVM, QEMU)
Telemetry pipelines
InfiniBand and RoCEv2 Ethernet troubleshooting

Job description

San Francisco Compute Company is seeking an experienced professional to manage and deploy GPU clusters. You'll work to ensure optimal performance, participate in on-call rotations, and help mentor junior engineers within our team.

This role requires a deep understanding of HPC and GPU systems, coupled with a passion for automation and infrastructure as code. Additional perks include competitive salary and generous equity options.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior HPC & GPU Cluster Architect
Senior HPC & GPU Cluster Architect

The Consensus • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Visa sponsorships
401(k) retirement matching
Medical, dental & vision insurance
+2
Senior HPC GPU Compute Engineer (Hybrid SF)
Senior HPC GPU Compute Engineer (Hybrid SF)

The San Francisco Compute Company • San Francisco (CA)

Hybrid
USD 180,000 - 260,000
Generous equity grant
Retirement matching
Comprehensive medical, dental, and vision insurance
+3
Senior HPC & GPU Compute Cluster Architect
Senior HPC & GPU Compute Cluster Architect

Electric Capital • San Francisco (CA)

Hybrid
USD 220,000 - 300,000
Generous equity grant
401(k) matching
Comprehensive medical, dental, and vision insurance
+3
Senior HPC Architect — Lead GPU Compute & Scale
Senior HPC Architect — Lead GPU Compute & Scale

NVIDIA AI • Illinois

On-site
USD 184,000 - 357,000
Senior HPC Architect: At-Scale GPU Deployment & Automation
Senior HPC Architect: At-Scale GPU Deployment & Automation

NVIDIA • New Mexico

On-site
USD 184,000 - 288,000
Senior GPU Compute Infra Architect - Onsite SF
Senior GPU Compute Infra Architect - Onsite SF

Hamilton Barnes Associates Limited • San Francisco (CA)

On-site
USD 233,000 - 316,000
Founding-level ownership and visible价值
Direct access to founders
Onsite role in San Francisco
Senior HPC AI Cluster Architect — Equity Eligible
Senior HPC AI Cluster Architect — Equity Eligible

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 176,000 - 334,000
Staff GPU Cluster Automation Engineer
Staff GPU Cluster Automation Engineer

Cssmerge • San Francisco (CA)

On-site
USD 224,000 - 284,000
Medical, Dental, Vision, Disability, and Life Insurance
401(k)
Unlimited Flexible Time Off
Senior HPC Engineer
Senior HPC Engineer

Hamilton Barnes ? • San Francisco (CA)

On-site
USD 180,000 - 250,000
Relocation provided
Founding-level ownership
Senior HPC Architect - GPU Compute, Equity Eligible
Senior HPC Architect - GPU Compute, Equity Eligible

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Equity
Inclusive work environment
Comprehensive benefits