Staff Compute Infra Engineer - GPU & AI Systems

xAI

Palo Alto (CA)

On-site

USD 180,000 - 440,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Comprehensive medical, vision, and dental coverage
Access to a 401(k) retirement plan
Short & long-term disability insurance
Life insurance
Various discounts and perks

Job summary

xAI is seeking a talented individual to join their Compute Infrastructure team, focused on building one of the world’s largest AI supercomputers. You will design and optimize massive GPU clusters, ensuring fast and reliable AI training. Ideal candidates will possess deep programming skills, GPU kernel optimization experience, and a strong grasp of large-scale distributed systems.

This role offers a competitive salary range of $180,000 - $440,000, along with equity and comprehensive benefits including medical coverage and a 401(k) retirement plan.

Qualifications

  • Deep low-level systems programming skills in C/C++ or Rust.
  • Experience with large-scale distributed compute infrastructure.
  • Hands-on experience with GPU kernel optimization techniques.

Responsibilities

  • Design and build GPU clusters for extreme-scale training.
  • Develop and optimize low-level CUDA kernels.
  • Work on Linux kernel internals and resource management.

Skills

Deep low-level systems programming (C/C++ or Rust)
Experience building high performance exabyte scale storage systems
Strong experience with large-scale GPU clusters
Hands-on work with GPU kernel optimization
Experience with Linux kernel internals
Ability to reason from first principles

Job description

xAI is seeking a talented individual to join their Compute Infrastructure team, focused on building one of the world’s largest AI supercomputers. You will design and optimize massive GPU clusters, ensuring fast and reliable AI training. Ideal candidates will possess deep programming skills, GPU kernel optimization experience, and a strong grasp of large-scale distributed systems.

This role offers a competitive salary range of $180,000 - $440,000, along with equity and comprehensive benefits including medical coverage and a 401(k) retirement plan.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer - Kernels/CUDA (C++)
Software Engineer - Kernels/CUDA (C++)

Pantera Capital • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Comprehensive medical, vision, and dental coverage
401(k) retirement plan
Equity options
+2
Member of Technical Staff - Compute Infrastructure
Member of Technical Staff - Compute Infrastructure

xAI • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Equity
Comprehensive medical, vision, and dental coverage
Access to a 401(k) retirement plan
+3
Head of AI Data Center Infrastructure Platforms and Software
Head of AI Data Center Infrastructure Platforms and Software

Summit Group Solutions, LLC • United States

On-site
USD 150,000 - 350,000
GPU Infra Solutions Architect for Large-Scale AI Clusters
GPU Infra Solutions Architect for Large-Scale AI Clusters

Prime Intellect • San Francisco (CA)

On-site
USD 150,000 - 300,000
Senior AI Infrastructure Engineer - GPU Compute
Senior AI Infrastructure Engineer - GPU Compute

Unchain Data • United States

On-site
USD 120,000 - 160,000
Senior AI Infra Engineer: Large-Scale GPU & HPC
Senior AI Infra Engineer: Large-Scale GPU & HPC

Anduril Industries • Costa Mesa (CA)

On-site
USD 166,000 - 220,000
Equity grants
Benefits package
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 280,000 - 420,000
Equity options
Health, vision, dental benefits
Unlimited PTO
+2
Senior GPU Infra Architect for Scalable AI Compute
Senior GPU Infra Architect for Scalable AI Compute

AI Chopping Block • Costa Mesa (CA), Northern (KY)

Hybrid
USD 166,000 - 220,000
Senior AI Infra Engineer — HPC & Scheduler
Senior AI Infra Engineer — HPC & Scheduler

Ai2 • Seattle (WA)

On-site
USD 126,000 - 189,000
Medical, dental, and vision insurance
401(k) plan enrollment
Monthly stipends for commuting and fitness
+1
Senior AI Infrastructure Lead - GPU Clusters & Model Serving
Senior AI Infrastructure Lead - GPU Clusters & Model Serving

Outsourceit • San Francisco (CA)

On-site
USD 120,000 - 170,000