Software Engineer, Facility Software Automation

Fluidstack

Austin (TX)

On-site

USD 250,000 - 300,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity in the form of stock options

Job summary

Fluidstack, located in Austin, Texas, is seeking a dedicated professional to build and own production services for their telemetry and automation stack. You will be responsible for designing the data architecture that ingests real-time telemetry from their vast compute fleet.

The ideal candidate has expertise in high-throughput messaging systems and telemetry pipeline integrity, and appreciates the importance of observability tools in production environments. Compensation ranges from $250,000 to $300,000 per year, offering equity options as well.

Qualifications

  • Experience in building and owning production services.
  • Focus on telemetry pipeline health and data integrity.
  • Proficient in messaging systems like NATS or Kafka.
  • Experience designing time-series data models.
  • Use of observability tools for production services.
  • Ability to troubleshoot pipelines rapidly.
  • Bonus for Kubernetes and ArgoCD experience.

Responsibilities

  • Own production services for telemetry collectors and automation stack.
  • Design the messaging and time-series data layer.
  • Oversee the software infrastructure and data lake.
  • Create APIs and dashboards for real-time telemetry.
  • Integrate data streams with the automation API layer.

Skills

Production services ownership
Telemetry pipeline health
High-throughput messaging systems
Time-series data models design
Observability tooling instrumentation
Signal chain expertise
Pipeline troubleshooting
Kubernetes and ArgoCD knowledge

Tools

ClickHouse
Kafka
Prometheus
Grafana

Job description

About Fluidstack

We exist to make humanity more free. For most of human history, you farmed or you starved. Technology gave people more time for the things they wanted to do, instead of things they had to do. Powerful AI will be the biggest lever for human choice we've ever built - but only if models are aligned with what humanity actually wants. There are groups building AI who don't share these goals. Whoever deploys frontier compute infrastructure fastest will decide whether AI expands human freedom or shrinks it.

We're singularly focused on delivering 10 to 100s of GWs of compute faster than anyone else, rethinking every layer of the stack. We acquire power, design and build data centers, and operate them - with teams spanning hardware and software. Speed and scale are our key differentiators. Come be a part of building civilization-scale infrastructure for AI.

We hire people who care deeply about this problem space. If that is you, please apply!

How We Operate
  • Extreme ownership. Full autonomy. Own things end to end often taking on scope outside your core role without being asked to get things done.

  • Velocity. We drive everything forward as fast as possible.

  • First principles. Challenge every assumption. Zero analogy thinking, no egos, the best idea wins.

  • Love of the game. The frontier of AI is the most interesting problem of our time. We put in long hours at high intensity to push the frontier forward.

The Facility Software Automation Team

Examples of key problems the team is working on

  • We do not just watch infrastructure run. We use the signals we collect to build automations that act on it -- compressing the time between a facility coming online and compute being in customers' hands. Detect, decide, act. If we fail to do any one of those three, we lose the trust of the customers who depend on us to keep the frontier moving.

  • Bad telemetry is a trust failure at gigawatt scale. Every facility Fluidstack operates runs on the data we collect. A missed signal is a missed alert. A missed alert is a customer running frontier AI workloads on infrastructure that cannot see itself. At the scale we are building, that is not a monitoring gap -- it is an outage.

  • The platform has to scale as fast as the build does. Fluidstack delivers gigawatts of compute in six months -- a fraction of the industry's 18 to 24 month timeline -- and every new site adds more devices, more signals, and more automation decisions that depend on clean data. The telemetry platform either grows ahead of the business or it becomes the constraint that slows it down.

Role Scope
  • Build and own production services across the telemetry collectors, ingesters, and automation stack that every downstream Automation and Operations team depends on.

  • Design and evolve the messaging and time‑series data layer that ingests real‑time telemetry from device signals across the global compute fleet.

  • Own the software infrastructure and data lake underpinning the platform, setting the data contracts and engineering standards other teams build against.

  • Build the APIs and dashboards that surface telemetry to commissioning, operations, and leadership teams in real time.

  • Drive the integration between device-level data streams and the automation API layer as the platform scales to new sites.

What We're Looking For

The below is a starting point. We always make space for exceptional people, so if you don’t fit this role exactly, tell us where you would.

  • You’ve built and owned production services that other teams depend on, and you don't ship code you wouldn't want to be paged for at 3am.

  • You treat telemetry pipeline health and data integrity as non-negotiable, because the automation decisions that deliver compute to customers are only as good as the signal underneath them.

  • You’ve worked with high-throughput messaging systems (NATS, Kafka, or equivalent) at real scale and know where the failure modes show up before they hit production.

  • You’ve designed time‑series data models in ClickHouse, TimescaleDB, or a comparable system, and can walk through the tradeoffs between write throughput, query performance, and schema evolution.

  • You’ve instrumented production services with observability tooling (Prometheus, Grafana, or equivalent) and treat metrics and alerting as part of shipping, not an afterthought.

  • You’ve worked across the full signal chain from device to database, and you design pipelines backwards from device‑level constraints so they’re correct by design.

  • You move toward a broken pipeline in production with urgency and without drama.

  • Bonus: Kubernetes and ArgoCD for production service deployment. ClickHouse schema design for time‑series workloads. Compute or critical infrastructure telemetry environments. Real-time dash‑boarding with Grafana or equivalent.

Compensation: $250,000 - $300,000 per year, depending on experience, skills, qualifications, and location. Offers equity in the form of stock options.

We are committed to pay equity and transparency.

Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Cloud Infrastructure
Software Engineer, Cloud Infrastructure

Fluidstack • San Francisco (CA)

On-site
USD 175,000 - 300,000
Health, dental, and vision insurance
Generous PTO policy
Retirement or pension plan
Distributed Systems Engineer
Distributed Systems Engineer

Fluidstack • San Francisco (CA)

On-site
USD 175,000 - 300,000
Health, dental, and vision insurance
Retirement or pension plan
Generous PTO policy
Software Engineer, Cloud Infrastructure
Software Engineer, Cloud Infrastructure

Fluidstack • New York (NY)

On-site
USD 175,000 - 300,000
Health, dental, and vision insurance
Retirement or pension plan
Generous PTO policy
Software Engineer, Cloud Infrastructure
Software Engineer, Cloud Infrastructure

Fluidstack • Seattle (WA)

On-site
USD 175,000 - 300,000
Health, dental, and vision insurance
Retirement or pension plan
Generous PTO policy
Development Engineer, Facility Software Automation
Development Engineer, Facility Software Automation

Fluidstack • Austin (TX)

On-site
USD 250,000 - 300,000
Stock options
Software Engineer, Facilities Pipeline
Software Engineer, Facilities Pipeline

Fluidstack • Seattle (WA), New York (NY), San Francisco (CA), Austin (TX)

On-site
USD 120,000 - 210,000
Staff Software Engineer
Staff Software Engineer

Fluidstack • San Francisco (CA)

On-site
USD 150,000 - 250,000
Health insurance
Generous PTO policy
Retirement plan
Production Engineer, IaaS
Production Engineer, IaaS

FluidStack • San Francisco (CA)

On-site
USD 175,000 - 300,000
Health, dental, and vision insurance
Generous PTO policy
Retirement or pension plan
Distributed Systems Engineer
Distributed Systems Engineer

Fluidstack • New York (NY)

On-site
USD 150,000 - 210,000
Technical Lead, Facilities Telemetry Platform
Technical Lead, Facilities Telemetry Platform

Fluidstack • New York (NY)

On-site
USD 150,000 - 210,000