Compute Platform Lead: Multi-Cloud & GPU Systems

Reflection AI Ltd

New York (NY)

On-site

USD 180,000 - 300,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Top-tier compensation
Stock options
Health & wellness benefits
Meals provided in office
Parental leave

Job summary

Reflection AI Ltd. is seeking a Compute Platform Lead to guide a multi-cloud, Kubernetes-based compute layer across our neo-cloud footprint.

You will mentor a team of systems engineers, shape architecture for multi-cloud scheduling, cluster management, and next‑generation GPU deployments, while contributing as an IC when required. You will manage vendor relationships, drive scalable, reliable operations, and partner with training teams to ensure fault tolerance and performance.

Qualifications

  • Proven experience leading systems or infrastructure teams with hands-on engineering.
  • Strong background in cluster-wide behavior, maintenance, and reliability.
  • Ability to earn technical trust and guide architectural decisions.

Responsibilities

  • Build, mentor, and grow a high-performing systems team.
  • Provide front-line leadership for multi-cloud scheduling and cluster management.
  • Stay hands-on to contribute technically when needed.
  • Prioritize work and manage projects in a fast-paced environment.
  • Shape strategy and execution for scalable, robust compute fleets.
  • Collaborate with training teams and manage important vendor deals.

Skills

Team leadership
Hands-on systems engineering
Multi-cloud scheduling
GPU deployment expertise
Vendor management
Strategic execution

Tools

Kubernetes
NCCL

Job description

Reflection AI Ltd. is seeking a Compute Platform Lead to guide a multi-cloud, Kubernetes-based compute layer across our neo-cloud footprint.

You will mentor a team of systems engineers, shape architecture for multi-cloud scheduling, cluster management, and next‑generation GPU deployments, while contributing as an IC when required. You will manage vendor relationships, drive scalable, reliable operations, and partner with training teams to ensure fault tolerance and performance.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Compute Platform Lead — Multi-Cloud & GPU Scale
Compute Platform Lead — Multi-Cloud & GPU Scale

B Capital • San Francisco (CA)

On-site
USD 240,000 - 360,000
Top-tier compensation
Stock options
Health & wellness
Compute Platform Lead: Multi-Cloud, GPU & Kubernetes
Compute Platform Lead: Multi-Cloud, GPU & Kubernetes

B Capital • San Francisco (CA)

On-site
USD 210,000 - 290,000
Top-tier compensation
Stock options
Comprehensive health/dental/vision
+5
Staff Platform Engineer - Multi-Cloud GPU & Kubernetes
Staff Platform Engineer - Multi-Cloud GPU & Kubernetes

Reflection AI Ltd • New York (NY)

On-site
USD 180,000 - 240,000
Top-tier compensation
Stock options
Health & wellness
+5
Compute Platform Engineer - GPU & Multi-Cloud Infra
Compute Platform Engineer - GPU & Multi-Cloud Infra

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive health benefits
Paid parental leave
+2
Multi-Cloud HPC Platform Architect (Kubernetes & GPUs)
Multi-Cloud HPC Platform Architect (Kubernetes & GPUs)

EPAM Systems Inc • United States

Remote
USD 140,000 - 230,000
Senior Compute Platform Lead | Kubernetes & Bare-Metal
Senior Compute Platform Lead | Kubernetes & Bare-Metal

Autonomai Recruitment • Chicago (IL)

On-site
USD 180,000 - 280,000
Senior GPU Infra Lead: Slurm, Kubernetes & Platform
Senior GPU Infra Lead: Slurm, Kubernetes & Platform

Jobgether SRL • United States

Remote
USD 170,000 - 250,000
Staff Software Engineer, Compute Control Plane (GPU/CPU)
Staff Software Engineer, Compute Control Plane (GPU/CPU)

Cloudjobs • San Jose (CA)

On-site
USD 180,000 - 280,000
Cash compensation
Equity compensation
Health insurance
+6
Staff Compute Platform Engineer – AI Cloud (Remote)
Staff Compute Platform Engineer – AI Cloud (Remote)

Applied Methods Ltd • Bellevue (WA), Northern (KY)

Hybrid
USD 190,000 - 260,000
Health coverage
Dental coverage
Vision coverage
+2
Staff Cloud Compute Engineer - GPU/CPU Lifecycle
Staff Cloud Compute Engineer - GPU/CPU Lifecycle

Lambda • Bellevue (KY)

On-site
USD 230,000 - 340,000
Health coverage
401k with 2% match
Flexible paid time off
+1