Senior HPC & GPU Compute Cluster Architect

Electric Capital

San Francisco (CA)

Hybrid

USD 220,000 - 300,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Generous equity grant
401(k) matching
Comprehensive medical, dental, and vision insurance
Unlimited paid time off
Daily lunch provision
Unlimited office book budget

Job summary

Electric Capital is seeking an experienced engineer to manage GPU clusters in San Francisco. You’ll be responsible for deploying and maintaining clusters, contributing to a small and ambitious team. Ideal candidates will have over 5 years of relevant experience and be comfortable with both networking and debugging across layers. The position offers a competitive salary along with equity, unlimited paid time off, medical, dental, vision, and various other benefits.

Qualifications

  • 5+ years of experience with operating, supporting, and scaling GPU compute clusters.
  • Deep understanding of server hardware fundamentals.
  • Experience troubleshooting high-speed fabrics.

Responsibilities

  • Participate in on-call rotation and deploy new environments.
  • Fix issues and enable deployments at scale.
  • Contribute to culture and mentor junior engineers.

Skills

Operating and supporting HPC or GPU compute clusters
Understanding server hardware fundamentals
Troubleshooting high-speed fabrics (InfiniBand, RoCEv2)
Debugging across hardware and OS layers
Automating fleet operations
Strong operational documentation skills
Mentoring junior engineers
Comfortable with partial onsite presence
Willingness for domestic travel

Tools

Linux
Slurm
Kubernetes

Job description

Electric Capital is seeking an experienced engineer to manage GPU clusters in San Francisco. You’ll be responsible for deploying and maintaining clusters, contributing to a small and ambitious team. Ideal candidates will have over 5 years of relevant experience and be comfortable with both networking and debugging across layers. The position offers a competitive salary along with equity, unlimited paid time off, medical, dental, vision, and various other benefits.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior HPC GPU Compute Engineer (Hybrid SF)
Senior HPC GPU Compute Engineer (Hybrid SF)

The San Francisco Compute Company • San Francisco (CA)

Hybrid
USD 180,000 - 260,000
Generous equity grant
Retirement matching
Comprehensive medical, dental, and vision insurance
+3
Senior HPC & GPU Cluster Architect — Scale & Automate
Senior HPC & GPU Cluster Architect — Scale & Automate

San Francisco Compute Company • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Generous equity grant
Competitive salary
Visa sponsorship
+6
Senior HPC & GPU Cluster Architect
Senior HPC & GPU Cluster Architect

The Consensus • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Visa sponsorships
401(k) retirement matching
Medical, dental & vision insurance
+2
Senior HPC Architect - GPU Compute, Equity Eligible
Senior HPC Architect - GPU Compute, Equity Eligible

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Equity
Inclusive work environment
Comprehensive benefits
Lead HPC Cluster Engineer for GPU AI Compute
Lead HPC Cluster Engineer for GPU AI Compute

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Equity
Benefits
Senior HPC Network Architect for GPU Compute
Senior HPC Network Architect for GPU Compute

Cssmerge • San Francisco (CA)

On-site
USD 224,000 - 284,000
Medical, Dental, Vision, Disability, and Life Insurance
401(k)
Unlimited Flexible Time Off and Paid Holidays
+1
Senior GPU HPC Cluster Engineer — Equity Eligible
Senior GPU HPC Cluster Engineer — Equity Eligible

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Senior GPU Compute Infra Architect - Onsite SF
Senior GPU Compute Infra Architect - Onsite SF

Hamilton Barnes Associates Limited • San Francisco (CA)

On-site
USD 233,000 - 316,000
Founding-level ownership and visible价值
Direct access to founders
Onsite role in San Francisco
GPU HPC Cluster Engineer for EDA & AI Compute
GPU HPC Cluster Engineer for EDA & AI Compute

NVIDIA Corporation • Austin (TX)

On-site
USD 152,000 - 242,000
Equity participation
Diverse work environment
Senior Data Center Network Engineer – GPU Clusters
Senior Data Center Network Engineer – GPU Clusters

Baseten • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation, including meaningful equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
+4