GPU Infra Solutions Architect for Large-Scale AI Clusters

Prime Intellect

San Francisco (CA)

On-site

USD 150,000 - 300,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Prime Intellect in San Francisco seeks a Solutions Architect for GPU Infrastructure who will transform client requirements into robust systems capable of training advanced AI models. Responsibilities include designing GPU cluster architectures, deploying orchestration systems, and supporting clients in optimizing their infrastructure.

The ideal candidate will have extensive experience with SLURM, Kubernetes, and NVIDIA architectures. A competitive cash compensation range of $150-300k plus equity incentives is provided.

Qualifications

  • 3+ years hands-on experience managing GPU clusters in HPC environments.
  • Deep expertise with orchestration tools SLURM and Kubernetes.
  • Strong understanding of NVIDIA architectures and CUDA ecosystems.

Responsibilities

  • Design GPU cluster architectures to meet client workload requirements.
  • Deploy orchestration systems for distributed workloads.
  • Act as technical escalation point for customer issues in infrastructure.

Skills

GPU clusters and HPC environments
SLURM
Kubernetes
InfiniBand configuration
NVIDIA GPU architecture
Python
Bash

Tools

Ansible
Terraform
Docker
Containerd

Job description

Prime Intellect in San Francisco seeks a Solutions Architect for GPU Infrastructure who will transform client requirements into robust systems capable of training advanced AI models. Responsibilities include designing GPU cluster architectures, deploying orchestration systems, and supporting clients in optimizing their infrastructure.

The ideal candidate will have extensive experience with SLURM, Kubernetes, and NVIDIA architectures. A competitive cash compensation range of $150-300k plus equity incentives is provided.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Compute Infra Engineer - GPU & AI Systems
Staff Compute Infra Engineer - GPU & AI Systems

xAI • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Lead AI Infrastructure Architect for Large-Scale GPU Clusters
Lead AI Infrastructure Architect for Large-Scale GPU Clusters

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity and benefits
Head of AI Data Center Infrastructure Platforms and Software
Head of AI Data Center Infrastructure Platforms and Software

Summit Group Solutions, LLC • United States

On-site
USD 150,000 - 350,000
Customer Solution Architect - Systems Integrator
Customer Solution Architect - Systems Integrator

Hamilton Barnes Associates Limited • New York (NY)

On-site
USD 225,000 - 275,000
RSU equity
20% bonus
Senior AI GPU Cluster Architect
Senior AI GPU Cluster Architect

STN Inc • San Francisco (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff - GPU Infrastructure
Member of Technical Staff - GPU Infrastructure

Prime Intellect • San Francisco (CA)

On-site
USD 150,000 - 300,000
AI Infra Solutions Architect — GPU & Cloud Deployments
AI Infra Solutions Architect — GPU & Cloud Deployments

NVIDIA Corporation • Austin (TX)

On-site
USD 152,000 - 288,000
Lead Large-Scale GPU Cluster Engineer for AI Research
Lead Large-Scale GPU Cluster Engineer for AI Research

Linuxcareers • San Francisco (CA)

On-site
USD 120,000 - 180,000
GPU Systems Engineer for AI Training Clusters
GPU Systems Engineer for AI Training Clusters

Thinkingmachines • San Francisco (CA)

On-site
USD 350,000 - 475,000
Generous health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
GPU Infrastructure Solutions Architect for Frontier AI
GPU Infrastructure Solutions Architect for Frontier AI

Prime Intellect • United States

On-site
USD 120,000 - 150,000