Compute Platform Lead: Multi-Cloud Systems & GPU

reflectionai

San Francisco, New York (CA, NY)

On-site

USD 200,000 - 280,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Top-tier compensation
Stock options
Health & wellness
Meals provided
Parental leave
Unlimited PTO
Visa sponsorship

Job summary

Reflection is seeking a Compute Platform Lead to build and mentor a high-performing systems team responsible for a multi-cloud Kubernetes-based compute platform. You will guide architectural decisions, balance hands-on work with leadership, and work closely with training teams to ensure fault tolerance and scalable GPU deployments.

The role emphasizes growth, vendor management, and long-term fleet readiness.

Qualifications

  • Experience building and growing systems teams while staying technically hands-on.
  • Deep systems engineering with a focus on cluster-wide behavior and maintenance.
  • Strong coding ability and the credibility to earn the technical trust of a strong team.
  • Depth in orchestration, storage, or GPU hardware with ability to learn the rest.
  • Alignment with a Kubernetes-first architecture.

Responsibilities

  • Build, mentor, and grow a high-performing team of systems engineers.
  • Provide front-line leadership to keep the compute fleet reliable and highly available across multi-cloud environments.
  • Stay hands-on to contribute as an individual contributor.
  • Manage day-to-day execution and project prioritization in a fast-paced setting.
  • Guide architectural decisions emphasizing scalability, robustness, and reliability.
  • Collaborate with training teams and manage important vendor relationships.
  • Prepare the fleet for next-gen GPUs and larger clusters.

Skills

Team leadership
Mentoring
Multi-cloud
GPU hardware
Cluster management
Fault tolerance
Vendor management
Kubernetes-first
Distributed systems

Tools

Kubernetes
NCCL

Job description

Reflection is seeking a Compute Platform Lead to build and mentor a high-performing systems team responsible for a multi-cloud Kubernetes-based compute platform. You will guide architectural decisions, balance hands-on work with leadership, and work closely with training teams to ensure fault tolerance and scalable GPU deployments.

The role emphasizes growth, vendor management, and long-term fleet readiness.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Compute Platform Lead: Multi-Cloud, GPU & Kubernetes
Compute Platform Lead: Multi-Cloud, GPU & Kubernetes

B Capital • San Francisco (CA)

On-site
USD 210,000 - 290,000
Top-tier compensation
Stock options
Comprehensive health/dental/vision
+5
Compute Platform Lead — Multi-Cloud & GPU Scale
Compute Platform Lead — Multi-Cloud & GPU Scale

B Capital • San Francisco (CA)

On-site
USD 240,000 - 360,000
Top-tier compensation
Stock options
Health & wellness
Platform Engineer: Multi-Cloud GPU Clusters & K8s
Platform Engineer: Multi-Cloud GPU Clusters & K8s

reflectionai • New York (NY)

On-site
USD 180,000 - 250,000
Top-tier compensation
Stock options
Health & wellness benefits
+4
Compute Platform Engineer - GPU & Multi-Cloud Infra
Compute Platform Engineer - GPU & Multi-Cloud Infra

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive health benefits
Paid parental leave
+2
Senior GPU Platform Architect (GPUaaS)
Senior GPU Platform Architect (GPUaaS)

Unify Technologies Ltd • Plano (TX)

On-site
USD 180,000 - 240,000
Staff Software Engineer, Compute Control Plane (GPU/CPU)
Staff Software Engineer, Compute Control Plane (GPU/CPU)

Cloudjobs • San Jose (CA)

On-site
USD 180,000 - 280,000
Cash compensation
Equity compensation
Health insurance
+6
Platform Engineering Lead — GPU & Kubernetes
Platform Engineering Lead — GPU & Kubernetes

Volta • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Equity
Health benefits
Retirement plan
+2
Senior Compute Platform Lead | Kubernetes & Bare-Metal
Senior Compute Platform Lead | Kubernetes & Bare-Metal

Autonomai Recruitment • Chicago (IL)

On-site
USD 180,000 - 280,000
Platform Engineer
Platform Engineer

Harrison Clarke • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Cloud Platform Engineer, GPU Core & Lifecycle
Senior Cloud Platform Engineer, GPU Core & Lifecycle

Lambda Labs • United States

Hybrid
USD 180,000 - 240,000
Health insurance
Dental insurance
Vision insurance
+5