Get more replies from employers
Send a job-specific resume in minutes.
Cumulus Labs (YC W26) is building the systems layer for AI infrastructure in San Francisco. We seek an ML Platforms Engineer to build and run the orchestration layer that schedules workloads, allocates GPUs, and keeps a heterogeneous, multi-cloud fleet running efficiently.
You'll own the GPU orchestrator end-to-end, implement multi-tenant primitives, and drive observability with metrics, logs, and traces at scale.
Cumulus Labs builds the software that turns raw GPU capacity into fast, cheap, production AI. We're looking for an ML Platforms Engineer to help build and run the orchestration layer underneath our inference and agent products, the system that schedules workloads, allocates GPUs, and keeps a heterogeneous, multi-cloud fleet running at high utilization.
We're small, early, and building the systems layer for the next generation of AI infrastructure. You'll own real infrastructure from day one, not tickets in a backlog.