Staff Software Engineer

Crusoe

Dublin

On-site

EUR 90,000 - 120,000

Full time

48 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Pension contributions
Private health and dental insurance
Income protection
Life assurance

Job summary

Crusoe is hiring a Software Engineer for the Cloud Availability Platform to build Conductor, a self-driving control plane that optimizes power, cost, and compute in real time. You will own greenfield infrastructure, tackle energy-aware scheduling, and partner with hardware, energy management, and data center teams to deliver scalable solutions.

The role requires 2+ years in production backend or infrastructure systems, strong systems programming skills in Go/Rust/C++.

Qualifications

  • 2+ years of production backend or infrastructure experience.
  • Mastery of state reconciliation, retries, idempotency, and automation.
  • Proficiency in Go, Rust, or C++ for systems programming.

Responsibilities

  • Design, build, and operate greenfield distributed services and control planes.
  • Collaborate with staff and principal engineers to review architectures and safety rails.
  • Work with cross-functional teams across hardware, energy management, and data center construction.
  • Develop features like unified observability pipelines and energy-aware compute scheduling.

Skills

Production systems
Distributed systems
Systems programming
Go / Rust / C++
Ownership & shipping mindset

Education

Bachelor's degree or equivalent

Tools

Kubernetes
Slurm

Job description

Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world's most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.

We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that — with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.

We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved — people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.

If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.

About This Role:

Crusoe is on a mission to accelerate the abundance of energy and intelligence by building the world’s only vertically integrated AI infrastructure company from the ground up. As a Software Engineer on the Cloud Availability Platform team, you will help build Conductor—our self-driving control plane designed to predict, decide, and act autonomously to keep tens of thousands of accelerators running at peak efficiency. In this role, you will play a direct part in solving the AI compute energy bottleneck by treating power as a control input rather than a static constraint. Working alongside staff and principal engineers, you will have true ownership over greenfield services, building distributed systems that optimize power, cost, and useful compute in real time.

This is a full-time position tailored for a problem-solving engineer who thrives in ambiguity and has a strong bias toward shipping. You will tackle rare infrastructure challenges—such as energy-aware compute scheduling, closed-loop remediation, and failure prediction on noisy hardware telemetry—transforming complex operational signals into a single, programmable control plane. If you are excited to build foundational infrastructure that turns tens of thousands of accelerators across sites into one cohesive, logical system, this role offers an unprecedented opportunity to do the most meaningful work of your career.

Key Responsibilities:
  • Core Responsibilities: Design, build, and operate greenfield distributed services, control planes, and closed-loop remediation systems that automatically drain, checkpoint, replace, and resume live AI workloads without manual intervention.

  • Communication & Collaboration: Partner closely with staff and principal engineers to review architectures, establish safety rails for autonomous operations, and share technical knowledge across the cloud services team.

  • Daily Interactions: Collaborate daily with cross-functional partners spanning hardware engineering, energy management, data center construction, and customer support to align system design with physical infrastructure.

  • Key Contributions: Own the development of core platform features such as unified observability pipelines, energy-aware compute scheduling, straggler detection, and self-qualifying hardware pipelines to maximize overall system goodput.

What You’ll Bring to the Team:
  • Production Systems Experience: Minimum of 2+ years of experience building, shipping, and operating backend or infrastructure systems in production (e.g., distributed services, control planes, schedulers, or observability pipelines).

  • Distributed Systems Foundations: Demonstrated mastery of state reconciliation, retries and idempotency, consistency tradeoffs, and autonomous automation on live infrastructure.

  • Systems Programming Proficiency: Strong engineering fundamentals using a modern systems language such as Go, Rust, or C++.

  • Problem-Solving & Ownership: High comfort level working through ambiguous problem statements, proposing clear designs, and driving features from concept to production.

  • Shipping Mindset: A strong bias toward action and iterative delivery, favoring practical production software over over-engineered documentation.

  • Educational Background: Bachelor's degree or equivalent practical experience in Computer Science, Engineering, or a related technical discipline.

Bonus Points:
  • Hardware & HPC Familiarity: Hands-on experience or deep familiarity with GPU health telemetry, NVLink/InfiniBand/RoCE networking fabrics, or hardware thermal and power behavior.

  • Observability Architecture: Prior experience designing scalable observability and telemetry pipelines that span compute, storage, and networking layers.

  • Scheduler Internals: Familiarity or direct exposure to Kubernetes internals, Slurm, or other distributed orchestrator ecosystems.

  • Statistical Methods & ML: Experience applying statistical methods, anomaly detection, or time-series forecasting to noisy operational data for predictive maintenance.

  • Zero-Trust Security: Knowledge of zero-trust architectures, policy-based access systems, or automated multi-tenant audit streaming.

Benefits:
  • pension contributions
  • private health and dental insurance
  • income protection
  • life assurance
  • and more.
Compensation:

Compensation will be paid as a salary or hourly. Compensation to be determined by the applicant’s education, experience, knowledge, skills, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Support Engineer
Senior Cloud Support Engineer

Crusoe Energy Systems LLC • Dublin

On-site
EUR 70,000 - 105,000
Senior Cloud Support Engineer
Senior Cloud Support Engineer

Crusoe • Dublin

On-site
EUR 60,000 - 90,000
Pension contributions
Private health insurance
Dental insurance
+2
Production Engineer (Kubernetes)
Production Engineer (Kubernetes)

Crusoe • Dublin

On-site
EUR 55,000 - 85,000
Hybrid work schedule
Competitive Paid Time Off
Healthcare benefits
+2
Senior Production Engineer
Senior Production Engineer

Crusoe Energy Systems LLC • Dublin

On-site
EUR 90,000 - 130,000
Pension contributions
Private health insurance
Life assurance
+1
Technical Project Manager, Data Center Operations
Technical Project Manager, Data Center Operations

Crusoe • Dublin

On-site
EUR 90,000 - 120,000
Pension contributions
Private health and dental insurance
Solutions Engineer (Dublin)
Solutions Engineer (Dublin)

Crusoe • Dublin

On-site
EUR 70,000 - 90,000
Pension contributions
Private health insurance
Life assurance
Data Center Deployment Engineer
Data Center Deployment Engineer

Dormont Manufacturing Co • Dublin

On-site
EUR 70,000 - 90,000
Pension contributions
Private health and dental insurance
Income protection
+1
Staff Production Engineer
Staff Production Engineer

crusoe • Dublin

On-site
EUR 90,000 - 150,000
Social security coverage
Pension funds
Private health insurance
+3
Staff Production Engineer
Staff Production Engineer

Linuxconfig • Dublin

Hybrid
EUR 90,000 - 120,000
Social security coverage
Generous leave policies
Private health insurance
+2
Production Engineer (Kubernetes)
Production Engineer (Kubernetes)

Crusoe Energy Systems LLC • Dublin

On-site
EUR 70,000 - 110,000
Pension contributions
Private health insurance
Income protection
+1