Senior AI Infrastructure Engineer | Scale GPU Clusters

Fuel Talent LLC

Seattle (WA)

Hybrid

USD 126,000 - 189,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Fuel Talent LLC in Seattle, WA is seeking a Senior Software Engineer to design and operate the infrastructure powering large-scale AI training. You will build and optimize orchestration and resource scheduling across GPU clusters, collaborating with researchers and engineers to scale performance and reliability.

The role combines hands-on software development with systems engineering, requiring Go and/or Python expertise, strong Linux knowledge, and experience with containers.

Qualifications

  • 8+ years of experience building business-critical software and running large-scale compute infrastructure.
  • Proficient in Go and/or Python and Linux systems.
  • Experience with container technologies such as Docker and orchestration tools like Kubernetes.
  • Strong distributed systems design, debugging, and performance optimization.

Responsibilities

  • Design and deliver infrastructure for large-scale AI training workloads.
  • Build and improve workload scheduling, orchestration, and execution systems.
  • Automate infrastructure management and reduce manual overhead.
  • Collaborate with researchers and engineers to optimize GPU resource utilization.

Skills

Go
Python
Linux
Docker
Distributed systems
GPU compute

Education

Bachelor's degree in Computer Science or equivalent

Tools

Kubernetes
Slurm
NCCL
InfiniBand
WEKA
Ceph

Job description

Fuel Talent LLC in Seattle, WA is seeking a Senior Software Engineer to design and operate the infrastructure powering large-scale AI training. You will build and optimize orchestration and resource scheduling across GPU clusters, collaborating with researchers and engineers to scale performance and reliability.

The role combines hands-on software development with systems engineering, requiring Go and/or Python expertise, strong Linux knowledge, and experience with containers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Infrastructure Engineer — Scale GPU Clusters Remote
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 280,000 - 420,000
Equity options
Health, vision, dental benefits
Unlimited PTO
+2
Senior AI Training Infra Engineer - Scale GPU Clusters
Senior AI Training Infra Engineer - Scale GPU Clusters

Designworks Talent • Bellevue (WA)

Hybrid
USD 180,000 - 240,000
Medical insurance
401(k) with company match
Paid holidays
Senior AI Infrastructure Engineer — GPU Clusters
Senior AI Infrastructure Engineer — GPU Clusters

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior AI Infrastructure Engineer — Scale GPU Clusters
Senior AI Infrastructure Engineer — Scale GPU Clusters

AI Breaking Wire • San Francisco (CA)

On-site
USD 280,000 - 400,000
Equity
Medical, dental, and vision benefits
Unlimited PTO
+2
Senior AI Infrastructure Engineer — Scalable GPU Clusters
Senior AI Infrastructure Engineer — Scalable GPU Clusters

NVIDIA AI • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Senior AI Infra Engineer: Scalable GPU Orchestration
Senior AI Infra Engineer: Scalable GPU Orchestration

Allen Institute for Artificial Intelligence • Seattle (WA)

On-site
USD 126,000 - 189,000
Medical, dental, and vision insurance
401k plan
Paid vacation and sick leave
+2
AI Infra & Cluster Engineer — Scale GPU/CPU Orchestration
AI Infra & Cluster Engineer — Scale GPU/CPU Orchestration

Linuxcareers • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior ML Infra Engineer - Scale GPU Clusters, Remote
Senior ML Infra Engineer - Scale GPU Clusters, Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 500,000
Equity
Medical/Dental/Vision coverage
Unlimited PTO
+1
Senior AI Infra Engineer — HPC & Scheduler
Senior AI Infra Engineer — HPC & Scheduler

Ai2 • Seattle (WA)

On-site
USD 126,000 - 189,000
Medical, dental, and vision insurance
401(k) plan enrollment
Monthly stipends for commuting and fitness
+1
Senior AI Infrastructure Lead - GPU Clusters & Model Serving
Senior AI Infrastructure Lead - GPU Clusters & Model Serving

Outsourceit • San Francisco (CA)

On-site
USD 120,000 - 170,000