AI Infrastructure Engineer

Acceler8 Talent

San Francisco, Northern (CA, KY)

On-site

USD 180,000 - 300,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Cash bonus
Founding engineer equity
Benefits

Job summary

Acceler8 Talent partners with an early-stage AI infrastructure company to hire a Founding AI Infrastructure Engineer in San Francisco. You’ll build and operate production GPU clusters across Kubernetes and Slurm, shape the hardware/software stack, and own performance, reliability and incident response for active workloads.

The role offers founding-engineer equity, competitive salary and benefits, with direct collaboration with the founders to deliver scalable AI compute and power-aware

Qualifications

  • Hands-on experience building and operating Kubernetes or Slurm clusters from scratch.
  • Strong knowledge of GPU infrastructure, distributed storage and high-performance networking.
  • Experience owning production systems and resolving complex infrastructure issues.
  • Experience with inference serving, distributed training or power-aware computing is highly valuable.

Responsibilities

  • Building and operating production GPU clusters across Kubernetes and Slurm.
  • Developing GPU orchestration, scheduling and lifecycle automation.
  • Designing distributed storage, NVMe and high-bandwidth networking systems.
  • Owning observability, reliability, workload performance and incident response.
  • Working directly with customers to translate workload needs into infrastructure.

Tools

Kubernetes
Slurm
GPU Infrastructure
Distributed Storage
High-Performance Networking
Observability
Power-Aware Computing

Job description

Founding AI Infrastructure Engineer

Founding Infrastructure Engineer – GPU Neocloud & AI Infrastructure

San Francisco, CA – On-site

I’m working with an early-stage infrastructure company building the first flex-power neocloud.

The company is building the hardware and software stack to keep GPU workloads reliable while managing compute around power availability. This could unlock approximately 70GW of existing capacity and bring AI compute online in months rather than waiting five or more years for new grid infrastructure.

This is a true founding engineering opportunity, working directly with the founders to build and operate production infrastructure for active customer workloads.

You’ll work on:
  • Building and operating production GPU clusters across Kubernetes and Slurm
  • Developing GPU orchestration, scheduling and lifecycle automation
  • Designing distributed storage, NVMe and high-bandwidth networking systems
  • Owning observability, reliability, workload performance and incident response
  • Working directly with customers to translate workload needs into infrastructure
Requirements:
  • Hands-on experience building and operating Kubernetes or Slurm clusters from scratch
  • Strong knowledge of GPU infrastructure, distributed storage and high-performance networking
  • Experience owning production systems and resolving complex infrastructure issues
  • Experience with inference serving, distributed training or power-aware computing is highly valuable

The company was founded by experienced second-time (successful!) entrepreneurs from AI infrastructure, software, power, real estate and data centres.

Despite being at the founding-team stage, the business already has approximately $5M in annual contract value under contract, a further $40M in pipeline and ongoing partnership discussions with NVIDIA and AMD.

Package: Competitive salary, cash bonus, benefits and meaningful founding-engineer equity.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC Engineer - AI Infrastructure
HPC Engineer - AI Infrastructure

Hamilton Barnes Associates Limited • San Francisco (CA)

On-site
USD 235,000 - 315,000
Founding engineer equity
Full benefits package
Founding AI Infra Engineer: GPU Clusters, Kubernetes
Founding AI Infra Engineer: GPU Clusters, Kubernetes

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 300,000
Cash bonus
Founding engineer equity
Benefits
Software Engineer - AI Infrastructure
Software Engineer - AI Infrastructure

Hamilton Barnes Associates Limited • San Francisco (CA)

Hybrid
USD 300,000 - 500,000
Early-stage equity
Founding engineer role
Equity package
Senior GPU Infrastructure Engineer - AI Infrastructure
Senior GPU Infrastructure Engineer - AI Infrastructure

Hamilton Barnes Associates Limited • Town of Texas (WI)

On-site
USD 120,000 - 160,000
Potential equity/bonus
AI Infra/HPC Engineer
AI Infra/HPC Engineer

Blue Signal Search • San Francisco (CA)

On-site
USD 180,000 - 240,000
Annual bonus
Equity participation
Comprehensive benefits
+1
Member of Technical Staff- Distributed Systems
Member of Technical Staff- Distributed Systems

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 350,000
Equity
Head of AI Data Center Infrastructure Platforms and Software
Head of AI Data Center Infrastructure Platforms and Software

Summit Group Solutions, LLC • United States

On-site
USD 150,000 - 350,000
Senior Solution Architect - AI Infrastructure
Senior Solution Architect - AI Infrastructure

Hamilton Barnes Associates Limited • San Francisco (CA)

On-site
USD 233,000 - 316,000
Founding-level ownership and visible价值
Direct access to founders
Onsite role in San Francisco
Senior Site Reliability Engineer (SRE) - AI Inftastructure
Senior Site Reliability Engineer (SRE) - AI Inftastructure

Hamilton Barnes Associates Limited • San Francisco (CA)

On-site
USD 270,000 - 330,000
Equity
Member of Technical Staff - GPU Infrastructure
Member of Technical Staff - GPU Infrastructure

Prime Intellect • San Francisco (CA)

On-site
USD 150,000 - 300,000