AI Inference Engineer — Scalable Model Serving

InvestedintheMission

Palo Alto (CA)

On-site

USD 135,000 - 210,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

SpaceX in Palo Alto is seeking a Software Engineer, Inference (AI Data Engineering) to design and optimize high-throughput model serving systems across SpaceX's internal AI infrastructure. You will own from distributed infrastructure to low-level optimizations, delivering reliable inference for mission-critical applications.

Aerospace experience is not required; the role emphasizes collaboration, problem solving, and impact on space exploration.

Qualifications

  • Bachelor's degree in computer science, engineering, math, or a scientific discipline or 2+ years of professional software experience in lieu of a degree.
  • Experience designing, implementing, and maintaining reliable, horizontally scalable distributed systems.
  • 1+ years of backend or full-stack production systems experience.

Responsibilities

  • Develop highly reliable, high-throughput inference systems for internal models.
  • Architect scalable distributed infrastructure for model serving including load balancing and auto-scaling.
  • Optimize latency and throughput under production workloads with GPU kernels and quantization.
  • Build high-concurrency serving systems with strong observability.
  • Own end-to-end components like request routing, SDKs, rate limiting, and scaling.
  • Benchmark and accelerate inference engines (SGLang, vLLM, TensorRT-LLM).
  • Develop tooling for tracing, replay, and issue resolution across the stack.

Skills

Distributed systems
Rust
C++
Python
Go
gRPC
Docker
Kubernetes
LLM inference engines

Education

Bachelor's degree in computer science, engineering, math, or scientific discipline

Tools

Docker
Kubernetes
PostgreSQL
ClickHouse
MongoDB
SGLang
vLLM
TensorRT-LLM
Triton

Job description

SpaceX in Palo Alto is seeking a Software Engineer, Inference (AI Data Engineering) to design and optimize high-throughput model serving systems across SpaceX's internal AI infrastructure. You will own from distributed infrastructure to low-level optimizations, delivering reliable inference for mission-critical applications.

Aerospace experience is not required; the role emphasizes collaboration, problem solving, and impact on space exploration.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Inference Engineer - Scalable, Low-Latency Systems
AI Inference Engineer - Scalable, Low-Latency Systems

SPACE EXPLORATION TECHNOLOGIES CORP • Palo Alto (CA), Northern (KY)

Hybrid
USD 135,000 - 210,000
401(k)
Medical, vision and dental coverage
Paid parental leave
+4
AI Inference Systems Engineer (High-Throughput, Low-Latency)
AI Inference Systems Engineer (High-Throughput, Low-Latency)

SpaceX • Palo Alto (CA)

On-site
USD 135,000 - 210,000
Stock options
Excellent medical coverage
401(k) plan
Staff Engineer - Large-Scale Model Inference & Systems
Staff Engineer - Large-Scale Model Inference & Systems

Xai • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Software Engineer, Inference (AI Data Engineering)
Software Engineer, Inference (AI Data Engineering)

SPACE EXPLORATION TECHNOLOGIES CORP • Palo Alto (CA), Northern (KY)

On-site
USD 135,000 - 210,000
401(k)
Medical, vision and dental coverage
Paid parental leave
+4
Software Engineer - Training/Inference (C++)
Software Engineer - Training/Inference (C++)

Xai • Palo Alto (CA)

On-site
USD 180,000 - 440,000
AI Data Engineer for Training & Evaluation
AI Data Engineer for Training & Evaluation

Socket.dev • Palo Alto (CA)

On-site
USD 144,000 - 270,000
Equity
Comprehensive medical, vision, anddent
401(k) retirement plan
+3
Inference Systems Engineer, High-Throughput AI Serving
Inference Systems Engineer, High-Throughput AI Serving

Future Ventures • Palo Alto (CA)

On-site
USD 135,000 - 160,000
Comprehensive medical, vision, dental coverage
401(k) retirement plan
Paid parental leave
+1
AI Data Steward & Evaluation Lead
AI Data Steward & Evaluation Lead

x • Palo Alto (CA), Northern (KY)

Hybrid
USD 100,000 - 186,000
Equity
Medical coverage
Vision coverage
+5
Senior AI Platform Engineer
Senior AI Platform Engineer

jobs.frontdoordefense.com - Jobboard • Starbase (TX)

On-site
USD 120,000 - 150,000
Inference Infra Engineer: Scale Low-Latency AI Serving
Inference Infra Engineer: Scale Low-Latency AI Serving

Elorian • Palo Alto (CA)

On-site
USD 200,000 - 400,000
Health, dental, and vision benefits
Unlimited PTO
Parental leave
+1