Distributed Systems Engineer — Scale GPU Cloud Platforms

Beam

New York (NY)

On-site

USD 120,000 - 170,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Competitive salary
Equity
Health, dental, vision
Learning budget
Fitness stipend

Job summary

Beam is seeking a Platform Engineer to advance our ultrafast AI inference platform. You will work on low‑level systems, container runtimes, OCI image formats, and lazy‑loading large files from content addressable storage, optimizing GPU utilization across clouds.

Our backend is primarily Go, with Python components. You’ll collaborate closely with customers, scale workloads to thousands of GPUs, and contribute to cutting‑edge features like GPU checkpoint restore and CRIU.

Qualifications

  • 3–5 years of experience working with a large distributed system.
  • Proficient with Kubernetes and cloud-native architectures.
  • Experience writing a statically typed language (Go or Rust).
  • Familiar with gRPC, Helm/Kustomize, and Terraform.
  • Comfortable collaborating with customers and open-source communities.

Responsibilities

  • Work on low-level systems, container runtimes, OCI image formats, and lazy-loading large files from content-addressable storage.
  • Efficiently pack thousands of workloads into GPUs across multiple clouds.
  • Contribute to cutting-edge GPU features like checkpoint/restore and CRIU.

Skills

Distributed systems
Kubernetes
Go
Rust
gRPC
Helm/Kustomize
Terraform
Customer collaboration
Cloud native
Open source

Tools

Container runtimes

Job description

Beam is seeking a Platform Engineer to advance our ultrafast AI inference platform. You will work on low‑level systems, container runtimes, OCI image formats, and lazy‑loading large files from content addressable storage, optimizing GPU utilization across clouds.

Our backend is primarily Go, with Python components. You’ll collaborate closely with customers, scale workloads to thousands of GPUs, and contribute to cutting‑edge features like GPU checkpoint restore and CRIU.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Distributed Systems Engineer
Distributed Systems Engineer

Beam • San Francisco (CA)

On-site
USD 140,000 - 200,000
Competitive compensation
Equity
Health, dental, vision benefits
+3
Distributed Systems Engineer
Distributed Systems Engineer

Beam • New York (NY)

On-site
USD 120,000 - 170,000
Competitive salary
Equity
Health, dental, vision
+2
GPU Platform Reliability Engineer
GPU Platform Reliability Engineer

Beam • San Francisco (CA)

On-site
USD 140,000 - 180,000
Competitive salary
Meaningful equity
Health, dental, vision benefits
+3
AI Infra Engineer — GPU Cloud, Kubernetes/Slurm
AI Infra Engineer — GPU Cloud, Kubernetes/Slurm

Blue Signal Search • San Francisco (CA)

On-site
USD 180,000 - 240,000
Annual bonus
Equity participation
Comprehensive benefits
+1
GPU Cluster Infrastructure Engineer
GPU Cluster Infrastructure Engineer

Smartshare, Inc. • New York (NY), Northern (KY)

Hybrid
USD 165,000 - 303,000
Competitive salary and meaningfulequiy
Health, dental, vision benefits
Fitness stipend and learning budget
+2
GPU Cloud Platform Lead for AI Inference
GPU Cloud Platform Lead for AI Inference

Blue Signal LLC. • United States

On-site
USD 250,000 - 350,000
Senior GPU Infra Engineer for Scalable AI Platform
Senior GPU Infra Engineer for Scalable AI Platform

Ll Oefentherapie • Nashville (TN)

On-site
USD 130,000 - 190,000
Senior GPU Cluster Engineer — HPC Infra, InfiniBand
Senior GPU Cluster Engineer — HPC Infra, InfiniBand

Smartshare, Inc. • New York (NY), Northern (KY)

Hybrid
USD 165,000 - 303,000
Competitive salary and meaningfulequiy
Health, dental, vision benefits
Fitness stipend and learning budget
+2
Lead GPU Systems & Fabric Architect for AI Cloud
Lead GPU Systems & Fabric Architect for AI Cloud

Bitdeer Technologies Group • San Jose (CA)

On-site
USD 180,000 - 240,000
Site Reliability Engineer
Site Reliability Engineer

Beam • San Francisco (CA)

On-site
USD 140,000 - 180,000
Competitive salary
Meaningful equity
Health, dental, vision benefits
+3