Senior Systems Software Engineer: Node & Cluster Management

Acceler8 Talent

Mountain View (CA)

Hybrid

USD 225,000 - 275,000

Full time

9 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Acceler8 Talent is sourcing a System Software Engineer for Node & Cluster Management in Mountain View, CA. This hybrid role requires in-person presence Tue–Thu and offers a $250k+ base with RSUs. You will shape how a first-generation AI platform is operated from a single node to a full cluster, interfacing with host software and OpenBMC firmware.

The role emphasizes low-level Linux development, node health, telemetry, PCIe devices, and firmware interactions across hardware boundaries.

Qualifications

  • 8+ years of systems-software experience.
  • Strong Linux systems development and low-level userspace experience.
  • Experience building REST APIs and CLI tools for hardware or infrastructure management.

Responsibilities

  • Build node-level management services for health, telemetry, inventory, and control.
  • Design cluster-management and failover to minimize downtime.
  • Develop REST APIs and CLI tools for diagnostics, firmware updates, and device recovery.
  • Create unified management across host software and OpenBMC firmware.
  • Extend management from nodes to racks and clusters.
  • Debug across APIs, daemons, kernel drivers, firmware and hardware.
  • Build provisioning, test automation, and monitoring tools for lab systems.

Skills

C
Go
Rust
C++
Python
Linux
REST APIs
CLI tools
Device drivers
BMC

Tools

Redfish
OpenBMC
gNMI
IPMI
PCIe
Telemetry

Job description

Acceler8 Talent is sourcing a System Software Engineer for Node & Cluster Management in Mountain View, CA. This hybrid role requires in-person presence Tue–Thu and offers a $250k+ base with RSUs. You will shape how a first-generation AI platform is operated from a single node to a full cluster, interfacing with host software and OpenBMC firmware.

The role emphasizes low-level Linux development, node health, telemetry, PCIe devices, and firmware interactions across hardware boundaries.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior System Software Engineer
Senior System Software Engineer

Acceler8 Talent • Mountain View (CA)

Hybrid
USD 225,000 - 275,000
Systems Engineer: Node & Cluster Orchestration
Systems Engineer: Node & Cluster Orchestration

MatX • Mountain View (CA)

On-site
USD 120,000 - 250,000
Health insurance
Dental insurance
Vision insurance
+4
Senior Node Infra Engineer — Scale Clusters Flexible
Senior Node Infra Engineer — Scale Clusters Flexible

Anthropic • San Francisco (CA)

Hybrid
USD 405,000 - 485,000
Competitive compensation
Generous vacation
Flexible working hours
Senior Linux Kernel Driver Engineer for AI Accelerators
Senior Linux Kernel Driver Engineer for AI Accelerators

Acceler8 Talent • Mountain View (CA)

Hybrid
USD 140,000 - 210,000
System Software Engineer
System Software Engineer

Acceler8 Talent • Mountain View (CA)

Hybrid
USD 140,000 - 210,000
Senior Systems Engineer, High-Performance Compute (Remote)
Senior Systems Engineer, High-Performance Compute (Remote)

Andromeda Cluster, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 250,000
Competitive compensation
Equity
Healthcare
+4
Senior GPU/HPC Systems Engineer - Onsite Fremont
Senior GPU/HPC Systems Engineer - Onsite Fremont

Acceler8 Talent • Fremont (CA), Northern (KY)

Hybrid
USD 135,000 - 165,000
Staff Node Infra Engineer - Scalable AI Clusters
Staff Node Infra Engineer - Scalable AI Clusters

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 405,000
Senior AI Cluster Network Engineer
Senior AI Cluster Network Engineer

NVIDIA • Seattle (WA)

On-site
USD 168,000 - 322,000
AI Systems Engineer: HPC & GPU Clusters
AI Systems Engineer: HPC & GPU Clusters

AMD • San Jose (CA)

On-site
USD 180,000 - 260,000
AMD benefits