Dataloading Systems Engineer - Multimodal AI (4d/w)

Eventual

San Francisco (CA)

On-site

USD 150,000 - 250,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

In‑person, tight-knit team — 4 days/wk
Competitive compensation and startup‑s
Catered lunches and dinners
Commuter benefit
Team-building events
Health, vision, and dental
401(k) plan with match

Job summary

Eventual in San Francisco is seeking a Systems Engineer for the Dataloading team to turn multi-petabyte video corpora into tensors on the GPU, driving the data path from object storage through NVMe, cache, and into device memory.

You will optimize for peak bandwidth on NVMe and memory hierarchies, align with Vera Rubin roadmap, and collaborate with labs and partners to push MFU end-to-end using CUDA, Rust/C++, and SLURM.

Qualifications

  • Strong systems-level performance focus and flamegraph mindset.
  • Proficiency with Rust, C++, or C and modern OS concepts.
  • Experience with IO and storage subsystems (NVMe, memory, network).
  • Familiarity with GPU data pipelines and HPC tooling (SLURM, CUDA).
  • Contributions to open-source systems in Rust or C++.
  • Experience with GV/FP architectures and high-throughput data paths.

Responsibilities

  • Design and build the video-native dataloader, returning tensors to the GPU.
  • Profile and optimize the data path from object store to device RAM.
  • Saturate latest hardware (B200, GB200, NVL72) on real training jobs.
  • Own performance benchmarks against baselines and prior numbers.
  • Collaborate with labs to land the loader in training stacks.
  • Work across Storage Infrastructure and model-output ingestion paths.

Skills

Rust
C++
C
Operating systems
Memory hierarchies

Tools

SLURM
Kubernetes
CUDA
PyTorch DataLoader

Job description

Eventual in San Francisco is seeking a Systems Engineer for the Dataloading team to turn multi-petabyte video corpora into tensors on the GPU, driving the data path from object storage through NVMe, cache, and into device memory.

You will optimize for peak bandwidth on NVMe and memory hierarchies, align with Vera Rubin roadmap, and collaborate with labs and partners to push MFU end-to-end using CUDA, Rust/C++, and SLURM.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC DataLoader Systems Engineer - 4 Day Week
HPC DataLoader Systems Engineer - 4 Day Week

Eventual Inc. • San Francisco (CA)

On-site
USD 180,000 - 240,000
In-person, tight-knit team — 4 days a‑
Software Engineer, High Performance Computing
Software Engineer, High Performance Computing

Eventual • United States

On-site
USD 100,000 - 130,000
Catered lunches and dinners
Flexible PTO
Health, vision, and dental coverage
+1
Dataloading Systems Engineer — GPU-Scale Data Pipelines
Dataloading Systems Engineer — GPU-Scale Data Pipelines

Eventual • United States

On-site
Software Engineer, High Performance Computing
Software Engineer, High Performance Computing

Eventual • San Francisco (CA)

On-site
USD 150,000 - 250,000
In‑person, tight-knit team — 4 days/wk
Competitive compensation and startup‑s
Catered lunches and dinners
+4
Software Engineer, High Performance Computing
Software Engineer, High Performance Computing

Eventual Inc. • San Francisco (CA)

On-site
USD 180,000 - 240,000
In-person, tight-knit team — 4 days a‑
Multimodal Data Engineer for AI Research
Multimodal Data Engineer for AI Research

Metamorphic • Palo Alto (CA)

On-site
USD 175,000 - 250,000
Senior Storage Systems Engineer for AI Data Infrastructure
Senior Storage Systems Engineer for AI Data Infrastructure

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity compensation
Benefits
Lead ML Systems Engineer — Distributed GPU Training & Infra
Lead ML Systems Engineer — Distributed GPU Training & Infra

Nvidia Corporation • Santa Clara (CA)

On-site
USD 224,000 - 431,000
Equity
Benefits
GPU Systems Engineer — Distributed Training & Inference
GPU Systems Engineer — Distributed Training & Inference

TensorScale AI • San Francisco (CA)

On-site
USD 180,000 - 240,000
AI Training Systems Architect (Distributed)
AI Training Systems Architect (Distributed)

Unconventional AI • Palo Alto (CA)

On-site
USD 210,000 - 320,000
Health benefits
401k matching
Unlimited PTO
+1