Infrastructure-as-a-Service Engineer

fluidstack

San Francisco (CA)

On-site

USD 140,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Fluidstack is building civilization-scale infrastructure for AI with a focus on scalable data center operations. The role involves creating the IaaS layer, automating bare-metal workflows, and debugging complex systems across compute, network, and storage.

You will contribute to production-grade automation in Go or Python, work with IPMI/Redfish, and potentially Kubernetes integrations to reduce toil and improve reliability.

Qualifications

  • Built infrastructure automation in production using Go or Python against real systems.
  • Experience with bare-metal provisioning tooling and knowledge of its location.
  • Deep debugging of Linux: kernel, drivers, networking.

Responsibilities

  • Develop and maintain the IaaS provisioning automation layer.
  • Automate bare-metal workflows end-to-end: image, boot, configure, validate, hand over.
  • Debug across compute, network, and storage when tenant workloads misbehave; reduce toil by writing software.

Skills

Go
Python
Bare-metal provisioning
Linux debugging

Tools

IPMI
Redfish
Kubernetes
Storage systems

Job description

About Fluidstack

We exist to make humanity more free. For most of human history, you farmed or you starved. Technology gave people more time for the things they wanted to do, instead of things they had to do. Powerful AI will be the biggest lever for human choice we've ever built - but only if models are aligned with what humanity actually wants. There are groups building AI who don't share these goals. Whoever deploys frontier compute infrastructure fastest will decide whether AI expands human freedom or shrinks it.

We're singularly focused on delivering 10 to 100s of GWs of compute faster than anyone else, rethinking every layer of the stack. We acquire power, design and build data centers, and operate them - with teams spanning hardware and software. Speed and scale are our key differentiators. Come be a part of building civilization-scale infrastructure for AI.

How We Operate

Extreme ownership. Full autonomy. Own things end to end often taking on scope outside your core role without being asked to get things done.

Velocity. We drive everything forward as fast as possible.

First principles. Challenge every assumption. Zero analogy thinking, no egos, the best idea wins.

Love of the game. The frontier of AI is the most interesting problem of our time. We put in long hours at high intensity to push the frontier forward.

The Data Center Operations Team

Operate at the scale of a nation, not a building. The fleet you run will draw more power than some countries, on the way to 100 GW.

Fly the plane while it's being built. Sites come online in pieces, and you keep the live ones running flawlessly while construction continues around them.

Write the playbook, don't inherit it. No prior operations org has run at this speed and scale, so the standards you set become the standard.

Role Scope

Build the IaaS layer: provisioning automation, tenant isolation, and lifecycle management for GPU fleets.

Automate bare‑metal workflows end to end: image, boot, configure, validate, hand over.

Debug across compute, network, and storage when tenant workloads misbehave.

Write software that removes toil: every manual runbook is a backlog item with your name on it.

What We're Looking For
  • Built infrastructure automation running in production, in Go or Python against real systems.
  • Worked with bare‑metal provisioning tooling and know where it lies to you.
  • Debug Linux deeply: kernel, drivers, networking.
  • Treat operational pain as a software bug, and fix it.
  • Bonus: GPU systems. Kubernetes internals. IPMI and Redfish. Storage systems.

Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans' status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Infrastructure-as-a-Service Lead
Infrastructure-as-a-Service Lead

Fluidstack • New York (NY)

On-site
USD 120,000 - 190,000
Production Engineer, IaaS Team Lead
Production Engineer, IaaS Team Lead

Fluidstack • San Francisco (CA)

On-site
USD 225,000 - 284,000
Production Engineer, Compute Team Lead
Production Engineer, Compute Team Lead

Fluidstack • Seattle (WA), New York (NY), Austin (TX), San Francisco (CA)

On-site
USD 180,000 - 280,000
Production Engineer, IaaS Team Lead
Production Engineer, IaaS Team Lead

Fluidstack • San Francisco (CA)

On-site
USD 180,000 - 260,000
Compute Deployment Engineer
Compute Deployment Engineer

Fluidstack • New York (NY)

On-site
USD 120,000 - 190,000
Kubernetes-based bare-metal Provisiong
Accelerator platform bring-up
Burn-in and stress harness design
Infrastructure Delivery Program Lead
Infrastructure Delivery Program Lead

Fluidstack • San Francisco (CA)

On-site
USD 253,000 - 295,000
Facilities Technician
Facilities Technician

PVH (Tommy Hilfiger/Calvin Klein) • Lubbock (TX)

On-site
USD 55,000 - 85,000
Compute Engineer, Deployment
Compute Engineer, Deployment

Fluidstack • San Francisco (CA)

On-site
USD 150,000 - 250,000
Infrastructure Delivery Program Lead
Infrastructure Delivery Program Lead

FluidStack • New York (NY)

On-site
USD 110,000 - 170,000
Data center exposure
GPU cluster bring-up experience
Facilities Technician
Facilities Technician

PVH (Tommy Hilfiger/Calvin Klein) • Buffalo (NY)

On-site
USD 55,000 - 75,000