AI Inference Engineer - Scalable, Low-Latency Systems

SPACE EXPLORATION TECHNOLOGIES CORP

Palo Alto, Northern (CA, KY)

Hybrid

USD 135,000 - 210,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

401(k)
Medical, vision and dental coverage
Paid parental leave
Paid vacation and holidays
Employee stock purchase plan
Disability insurance
Life insurance

Job summary

SpaceX in Palo Alto seeks a Software Engineer, Inference (AI Data Engineering) to design and optimize high-throughput AI model serving systems. You will own distributed infrastructure from routing to batching and work with SpaceX AI teams to deliver reliable, scalable inference that powers mission-critical applications.

You will collaborate across teams while enhancing performance with GPU kernels, quantization, and low-level optimizations, reporting through SpaceX Internal AI Infrastructure.

Qualifications

  • Bachelor's degree or 2+ years of software experience in lieu of degree.
  • Experience designing and maintaining reliable, horizontally scalable distributed systems.
  • 1+ years full stack or backend development with production systems.
  • 1+ years of Rust or C++ programming.

Responsibilities

  • Develop high-throughput inference systems across SpaceX internals.
  • Architect scalable distributed infrastructure for model serving (load balancing, auto-scaling, batching).
  • Optimize latency and throughput under production workloads with GPU work and quantization.
  • Build reliable, high-concurrency serving systems with observability.
  • Own end-to-end components for internal SpaceX AI inference platforms.
  • Benchmark and accelerate inference engines (SGLang, vLLM, TensorRT-LLM).
  • Create tools for tracing, replaying issues across the stack.
  • Develop CI/CD for endpoint deployment and updates.
  • Collaborate with SpaceX AI teams to integrate inference into workflows.

Skills

Distributed systems design
Full stack/backend development
Rust or C++
LLM inference engines
Docker/Kubernetes
gRPC
Python/Go
Performance profiling
High-concurrency systems

Education

Bachelor's degree in CS/Engineering/Math

Tools

Docker
Kubernetes
PostgreSQL
ClickHouse
MongoDB
SGLang
vLLM
TensorRT-LLM
GPU kernels

Job description

SpaceX in Palo Alto seeks a Software Engineer, Inference (AI Data Engineering) to design and optimize high-throughput AI model serving systems. You will own distributed infrastructure from routing to batching and work with SpaceX AI teams to deliver reliable, scalable inference that powers mission-critical applications.

You will collaborate across teams while enhancing performance with GPU kernels, quantization, and low-level optimizations, reporting through SpaceX Internal AI Infrastructure.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Inference Systems Engineer (High-Throughput, Low-Latency)
AI Inference Systems Engineer (High-Throughput, Low-Latency)

SpaceX • Palo Alto (CA)

On-site
USD 135,000 - 210,000
Stock options
Excellent medical coverage
401(k) plan
Staff Engineer - Large-Scale Model Inference & Systems
Staff Engineer - Large-Scale Model Inference & Systems

Xai • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Software Engineer - Training/Inference (C++)
Software Engineer - Training/Inference (C++)

Xai • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Software Engineer, Inference (AI Data Engineering)
Software Engineer, Inference (AI Data Engineering)

SPACE EXPLORATION TECHNOLOGIES CORP • Palo Alto (CA), Northern (KY)

Hybrid
USD 135,000 - 210,000
401(k)
Medical, vision and dental coverage
Paid parental leave
+4
Inference Infra Engineer: Scale Low-Latency AI Serving
Inference Infra Engineer: Scale Low-Latency AI Serving

Elorian • Palo Alto (CA)

On-site
USD 200,000 - 400,000
Health, dental, and vision benefits
Unlimited PTO
Parental leave
+1
AI Supercomputer Network Engineer – HPC Infrastructure
AI Supercomputer Network Engineer – HPC Infrastructure

xAI • Memphis (TN)

Hybrid
USD 120,000 - 190,000
Platform Infra Engineer (Rust/C++, Kubernetes)
Platform Infra Engineer (Rust/C++, Kubernetes)

SpaceXAI • Bellevue (WA)

On-site
USD 180,000 - 440,000
Senior AI Platform Engineer
Senior AI Platform Engineer

jobs.frontdoordefense.com - Jobboard • Starbase (TX)

On-site
USD 120,000 - 150,000
Software Engineer, Inference (AI Data Engineering)
Software Engineer, Inference (AI Data Engineering)

SpaceX • Palo Alto (CA)

On-site
USD 135,000 - 210,000
Stock options
Excellent medical coverage
401(k) plan
Applied AI Software Engineer: Production-Grade AI
Applied AI Software Engineer: Production-Grade AI

SpaceX • Redmond (WA)

On-site
USD 125,000 - 200,000
Stock options
401(k) plan
Comprehensive medical, vision and dent
+1