AI Infrastructure Engineer

Fuel Talent

Seattle (WA)

On-site

USD 180,000 - 210,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Fuel Talent in Seattle is seeking an AI Infrastructure Engineer to lead the 0-to-1 build of the core compute and orchestration layer for ML workloads. You will own GPU orchestration and resource management, shaping engineering standards from day one.

You will design researcher-friendly SDKs/APIs, operationalize ML pipelines, and partner with customers to improve model performance and observability, while guiding the tech roadmap and security posture.

Qualifications

  • 5+ years of software engineering experience focused on ML infrastructure or backend systems.
  • Proficiency in Golang, Python, or Typescript.
  • Deep Kubernetes experience for GPU workloads on AWS/GCP.
  • Experience deploying ML training or inference pipelines in production (PyTorch, Hugging Face).

Responsibilities

  • Architect the GPU compute fabric and orchestration layer for ML workloads.
  • Design SDKs and APIs to streamline ML workflows for researchers.
  • Operationalize end-to-end ML pipelines from data ingestion to model serving.
  • Collaborate with customers to debug jobs and build observability tools.
  • Lead the technical roadmap and influence infra and security decisions.

Skills

ML infra
Golang
Python
Typescript
Kubernetes
System design

Tools

AWS
GCP
PyTorch
Hugging Face

Job description

Compensation: $180K–210K + 0.1%–0.5% equity

Visa Support: NO; due to the sensitive nature of this role, we are only able to accommodate those with US Citizenship or Green Card Holder status

We are partnered with a seed-stage startup building a platform for scientific ML.

As an AI Infrastructure Engineer, you will report to the CTO and lead the 0-to-1 build of the core compute and orchestration layer, holding significant ownership and shaping the technical direction from day one.

What you'll do
  • Architect the GPU compute fabric: build and manage the orchestration layer for GPU workloads, ensuring efficient resource allocation and cost management across training, fine-tuning, and inference
  • Design developer-centric SDKs and APIs that turn complex ML workflows into intuitive experiences for researchers and data scientists
  • Operationalize the ML lifecycle with robust end-to-end pipelines, from data ingestion and preprocessing to secure model serving and monitoring
  • Partner with customers to debug fine-tuning jobs and build observability tools for tracking model performance and resource health in real time
  • Lead the technical roadmap, making critical build-vs-buy decisions on infrastructure and security while shaping engineering standards and hiring
What we're looking for
  • 5+ years of software engineering experience, focused on ML infrastructure or backend systems supporting ML workloads
  • Programming expertise in Golang, Python, or Typescript
  • Deep Kubernetes experience, ideally for GPU workloads on AWS/GCP
  • Experience deploying and operating ML/DL training or inference pipelines in production (PyTorch, Hugging Face, or similar)
  • Strong CS fundamentals and system design skills
  • Ability to thrive in fast-paced, ambiguous environments
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Infrastructure Engineer
ML Infrastructure Engineer

Strativ Group • Menlo Park (CA)

On-site
USD 250,000 - 320,000
AI Infrastructure / ML Infrastructure Engineer
AI Infrastructure / ML Infrastructure Engineer

DeWinter Group • Campbell (CA)

On-site
AI/ML Infra Engineer - Hosting
AI/ML Infra Engineer - Hosting

Hamilton Barnes Associates Limited • San Francisco (CA)

On-site
USD 225,000 - 275,000
Stock options
Solution Architect - AI Infrastructure
Solution Architect - AI Infrastructure

Hamilton Barnes Associates Limited • Town of Texas (WI)

On-site
USD 283,500 - 346,500
Equity (RSUs)
Software Engineer, AI Infrastructure
Software Engineer, AI Infrastructure

Harell Data • Palo Alto (CA)

On-site
USD 180,000 - 260,000
Principal ML Infrastructure Engineer (Relocation Available)
Principal ML Infrastructure Engineer (Relocation Available)

Franklin Fitch • Dallas (TX)

On-site
USD 100,000 - 140,000
Staff Software Engineer (AI Infrastructure)
Staff Software Engineer (AI Infrastructure)

DeepRec.ai • Palo Alto (CA)

On-site
USD 180,000 - 320,000
Software Engineer, AI Infra
Software Engineer, AI Infra

Makers Fund • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Equity
Health benefits
Monthly stipends
+1
ML Infrastructure Engineer
ML Infrastructure Engineer

Lattice, Inc. • San Francisco (CA)

Hybrid
USD 200,000 - 280,000
Competitive salary
Premium health, dental, and vision insurance
Unlimited PTO
+2
Customer Solution Architect - Systems Integrator
Customer Solution Architect - Systems Integrator

Hamilton Barnes Associates Limited • New York (NY)

On-site
USD 225,000 - 275,000
RSU equity
20% bonus