Compute Platform Lead — Multi-Cloud & GPU Scale

B Capital

San Francisco (CA)

On-site

USD 240,000 - 360,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Top-tier compensation
Stock options
Health & wellness

Job summary

Reflection's Compute Platform team leads a Kubernetes-based, multi-cloud compute layer across neo-clouds. As Compute Platform Lead you will mentor a team of systems engineers, guide architectural decisions, and work with training teams on fault tolerance and remediation.

You will stay hands-on to contribute and ensure the compute fleet supports large-scale training runs. This role also manages vendor relationships, drives capacity planning, and prepares the fleet for next-gen GPUs and

Qualifications

  • Experience building, mentoring, and growing systems or infrastructure teams while staying technically hands-on.
  • Deep systems-level engineering with a focus on cluster-wide behavior and maintenance.
  • Strong coding ability and credibility to earn technical trust of a team.
  • Depth in orchestration, storage, or GPU hardware with ability to learn the rest.

Responsibilities

  • Build, mentor, and grow a high-performing team of systems engineers.
  • Provide front-line leadership to keep the compute fleet reliable and highly available.
  • Stay hands-on to contribute as an individual in the team's stack.
  • Manage day-to-day execution and project priorities in a fast-paced environment.
  • Guide architectural decisions emphasizing scalability, robustness, and reliability.

Skills

Team leadership
Kubernetes
Multi-cloud
GPU deployment
Systems engineering
Vendor management
Architectural decisions
Coding ability

Tools

NCCL
CUDA
Prometheus

Job description

Reflection's Compute Platform team leads a Kubernetes-based, multi-cloud compute layer across neo-clouds. As Compute Platform Lead you will mentor a team of systems engineers, guide architectural decisions, and work with training teams on fault tolerance and remediation.

You will stay hands-on to contribute and ensure the compute fleet supports large-scale training runs. This role also manages vendor relationships, drives capacity planning, and prepares the fleet for next-gen GPUs and

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Compute Platform Lead: Multi-Cloud Kubernetes & GPU
Compute Platform Lead: Multi-Cloud Kubernetes & GPU

Socket.dev • San Francisco (CA)

On-site
USD 210,000 - 320,000
Top-tier compensation and equity
Stock options
Health & wellness
+5
Compute Platform Engineering Lead (Multi-Cloud)
Compute Platform Engineering Lead (Multi-Cloud)

Doist • San Francisco (CA), New York (NY)

On-site
USD 180,000 - 240,000
Stock options
Health & wellness
Meals provided in office
+1
Compute Platform Lead: Multi-Cloud, GPU Ops, & Mentorship
Compute Platform Lead: Multi-Cloud, GPU Ops, & Mentorship

Reflection AI • New York (NY), San Francisco (CA)

On-site
USD 180,000 - 280,000
Top-tier compensation
Stock options
Health benefits
+5
Compute Platform Lead: Multi-Cloud, GPU & Kubernetes
Compute Platform Lead: Multi-Cloud, GPU & Kubernetes

B Capital • San Francisco (CA)

On-site
USD 210,000 - 290,000
Top-tier compensation
Stock options
Comprehensive health/dental/vision
+5
Staff Engineer, Compute Platform & GPU Infra
Staff Engineer, Compute Platform & GPU Infra

Visa Hunt • New York (NY)

On-site
USD 180,000 - 240,000
Top-tier compensation
Stock options
Health & wellness
+5
Compute Platform Engineer - GPU & Multi-Cloud Infra
Compute Platform Engineer - GPU & Multi-Cloud Infra

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive health benefits
Paid parental leave
+2
Platform Lead, Multi-Cloud GPU Inference & Training
Platform Lead, Multi-Cloud GPU Inference & Training

Perplexity • New York (NY)

On-site
USD 250,000 - 485,000
Compute Platform Engineer – HPC & GPU Systems
Compute Platform Engineer – HPC & GPU Systems

NorthMark Compute & Cloud • Dallas (TX)

On-site
USD 120,000 - 180,000
Platform Engineering Lead — GPU & Kubernetes
Platform Engineering Lead — GPU & Kubernetes

Volta • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Equity
Health benefits
Retirement plan
+2
GPU Platform Engineer — Self-Serve ML Compute
GPU Platform Engineer — Self-Serve ML Compute

B Capital • United States

On-site
USD 180,000 - 230,000