Staff, Developer Technology - GPU AI Inference & Training

RadixArk

Palo Alto (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Flexible work arrangements
Competitive benefits

Job summary

RadixArk is seeking a Member of Technical Staff, Developer Technology (DevTech) to accelerate LLM inference and training on modern GPU hardware. Our engines power trillions of tokens daily and enable researchers and partners to scale AI workloads.

You will profile, optimize, and extend SGLang and Miles across GPUs, collaborate with ecosystem partners, and ship day-0 model support while shaping the roadmap.

Qualifications

  • 4+ years of experience in GPU systems, LLM infrastructure, or performance engineering.
  • Strong profiling and debugging skills: able to root‑cause performance and correctness issues across the stack.
  • Hands‑on GPU programming experience in at least one of CUDA, ROCm, or Triton, and willingness to work across platforms.
  • Strong programming skills in Python plus C++ or CUDA.
  • Comfortable making progress on hard, ambiguous problems with little context to start from, and fast to ramp into unfamiliar systems, codebases, and domains.
  • Ability to translate ambiguous asks into clear technical plans, verified cookbooks, and actionable recommendations, and to communicate credibly with expert engineering audiences.

Responsibilities

  • Accelerate AI workloads: Profile and optimize GPU performance for real production workloads on current and next-generation hardware, root‑causing bottlenecks from kernels to distributed multi-node systems.
  • Go deep in one or two focus areas. The team covers the full stack; each engineer specializes in one or two tracks: Inference performance, Kernels and model/hardware enablement, Training systems.
  • Partner directly with the ecosystem. Turn ambiguous, high‑stakes problems from expert engineers at our key partners into concrete wins, clear technical guidance, and reproducible cookbooks.
  • Enhance SGLang and Miles. Feed user‑driven improvements back into our open‑source systems and roadmap, so every win compounds across the ecosystem.

Skills

GPU systems
LLM infrastructure
Performance engineering
Profiling
Debugging
Python
C++
CUDA
Cross-functional communication
Ambiguity handling

Tools

CUDA
ROCm
Triton

Job description

RadixArk is seeking a Member of Technical Staff, Developer Technology (DevTech) to accelerate LLM inference and training on modern GPU hardware. Our engines power trillions of tokens daily and enable researchers and partners to scale AI workloads.

You will profile, optimize, and extend SGLang and Miles across GPUs, collaborate with ecosystem partners, and ship day-0 model support while shaping the roadmap.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff DevTech: Accelerate AI Inference on GPUs
Staff DevTech: Accelerate AI Inference on GPUs

RadixArk • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff — Developer Technology
Member of Technical Staff — Developer Technology

RadixArk • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Equity
Flexible work arrangements
Competitive benefits
Member of Technical Staff — Developer TechnologyPalo Alto, CA
Member of Technical Staff — Developer TechnologyPalo Alto, CA

RadixArk • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff — Inference
Member of Technical Staff — Inference

Dormont Manufacturing Co • Palo Alto (CA)

On-site
USD 190,000 - 260,000
Competitive compensation
Meaningful equity
Comprehensive benefits
+1
Staff Accelerator Systems Engineer
Staff Accelerator Systems Engineer

RadixArk • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Competitive compensation
Equity
Flexible work arrangements
+1
Member of Technical Staff — Inference-Multi-HardwarePalo Alto, CA
Member of Technical Staff — Inference-Multi-HardwarePalo Alto, CA

RadixArk • Palo Alto (CA)

On-site
USD 210,000 - 310,000
Member of Technical Staff — Performance
Member of Technical Staff — Performance

RadixArk • Palo Alto (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
Comprehensive health benefits
Flexible work arrangements
Member of Technical Staff — Heterogenous Hardware
Member of Technical Staff — Heterogenous Hardware

RadixArk • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Competitive compensation
Equity
Flexible work arrangements
+1
Member of Technical Staff — Inference-Kernel, Compiler & Communication
Member of Technical Staff — Inference-Kernel, Compiler & Communication

RadixArk • Palo Alto (CA)

On-site
USD 210,000 - 290,000
Competitive compensation
Comprehensive benefits
Flexible work arrangements
Member of Technical Staff — Training
Member of Technical Staff — Training

RadixArk • Palo Alto (CA)

On-site
USD 120,000 - 160,000
Comprehensive benefits
Flexible work arrangements