Principal Data Center Reliability Engineer

Fluidstack

United States

On-site

USD 220,000 - 260,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Fluidstack is seeking a Senior Reliability Engineer to own the fleet reliability program for its data-center fleet in the United States. You will set availability targets and measure them to close gaps across multiple sites, ensuring high uptime and safe, scalable operations.

You will analyze incidents using Weibull, Pareto, and FMEA, drive corrective actions across teams, and shape the maintenance strategy with CMMS analytics and reliability-centered approaches to minimize future failures.

Qualifications

  • Experience in reliability engineering for critical infrastructure.
  • Strong root cause analysis skills and data-driven decision making.
  • Experience with Weibull, Pareto, and FMEA methods.
  • Ability to drive cross-team corrective actions.

Responsibilities

  • Own fleet reliability engineering: define availability targets, measure them honestly, and close the gap.
  • Run root cause analysis on the fleet's worst incidents and drive corrective actions to done across every site.
  • Build the failure data pipeline, facility and hardware both, that turns incident history into engineering priorities.
  • Set the maintenance strategy (reliability-centered, condition-based) so the fleet spends effort where the failure data says to.

Skills

Reliability engineering
Root cause analysis
Weibull analysis
Pareto analysis
FMEA
Cross-team coordination
Maintenance planning

Tools

CMMS

Job description

Fluidstack is seeking a Senior Reliability Engineer to own the fleet reliability program for its data-center fleet in the United States. You will set availability targets and measure them to close gaps across multiple sites, ensuring high uptime and safe, scalable operations.

You will analyze incidents using Weibull, Pareto, and FMEA, drive corrective actions across teams, and shape the maintenance strategy with CMMS analytics and reliability-centered approaches to minimize future failures.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Center Reliability Engineer
Senior Data Center Reliability Engineer

Fluidstack • Austin (TX)

On-site
USD 220,000 - 260,000
Reliability Engineer - AI Infra & Data Centers
Reliability Engineer - AI Infra & Data Centers

Fluidstack • Seattle (WA), New York (NY), San Francisco (CA), Austin (TX)

On-site
USD 120,000 - 180,000
Data Center Reliability Engineer — Design & Modeling (Equity)
Data Center Reliability Engineer — Design & Modeling (Equity)

FluidStack • United States

On-site
USD 200,000 - 250,000
Stock options
Reliability Engineer - Large-Scale AI & HPC Systems
Reliability Engineer - Large-Scale AI & HPC Systems

PVH (Tommy Hilfiger/Calvin Klein) • San Francisco (CA)

On-site
USD 120,000 - 180,000
Network Reliability Engineer — Automation & AI Tooling
Network Reliability Engineer — Automation & AI Tooling

Fluidstack • Austin (TX)

On-site
USD 208,000 - 269,000
Salary growth potential
Premium health benefits
Reliability Engineer, Data Center Design
Reliability Engineer, Data Center Design

Fluidstack • Austin (TX)

On-site
USD 200,000 - 250,000
Stock options
Data Center Operations Lead
Data Center Operations Lead

Fluidstack • Lubbock (TX)

On-site
USD 120,000 - 180,000
Health, dental, and vision insurance
Retirement plan
Generous PTO policy
+1
Reliability Engineer, Data Center Design
Reliability Engineer, Data Center Design

FluidStack • United States

On-site
USD 200,000 - 250,000
Stock options
Global Enterprise Asset & Maintenance Leader
Global Enterprise Asset & Maintenance Leader

Fluidstack • New York (NY)

On-site
USD 240,000 - 310,000
Senior AI Compute Reliability Engineer
Senior AI Compute Reliability Engineer

Fluidstack • New York (NY)

On-site
USD 173,000 - 224,000