Performance Engineer: AI Inference & Systems (LLM)

RadixArk

Palo Alto (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation
Comprehensive health benefits
Flexible work arrangements

Job summary

RadixArk is seeking a Performance Engineer in Palo Alto, CA to improve performance in AI workloads. You'll work on optimizing various aspects of AI systems, ensuring they are reliable and efficient.

Ideal candidates have a strong systems engineering background, experience with GPU systems, and the ability to debug performance issues across multiple layers. RadixArk offers competitive compensation and flexible working arrangements.

Qualifications

  • Experience in performance-critical software engineering.
  • Familiarity with GPU systems and cloud environments.
  • Strong communication skills to translate complex issues.

Responsibilities

  • Optimize performance across production deployments.
  • Benchmark workloads for AI models.
  • Investigate performance regressions in real environments.

Skills

Performance-critical software
GPU systems
Distributed systems
Machine learning runtimes
Python
C++
Performance debugging
Profiling tools

Tools

CUDA
Triton
Pallas
ROCm
XLA

Job description

RadixArk is seeking a Performance Engineer in Palo Alto, CA to improve performance in AI workloads. You'll work on optimizing various aspects of AI systems, ensuring they are reliable and efficient.

Ideal candidates have a strong systems engineering background, experience with GPU systems, and the ability to debug performance issues across multiple layers. RadixArk offers competitive compensation and flexible working arrangements.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff — Performance
Member of Technical Staff — Performance

RadixArk • Palo Alto (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
Comprehensive health benefits
Flexible work arrangements
Member of Technical Staff — Inference-Kernel, Compiler & CommunicationPalo Alto, CA
Member of Technical Staff — Inference-Kernel, Compiler & CommunicationPalo Alto, CA

RadixArk • Palo Alto (CA)

On-site
USD 190,000 - 270,000
Member of Technical Staff — Inference
Member of Technical Staff — Inference

Dormont Manufacturing Co • Palo Alto (CA)

On-site
USD 190,000 - 260,000
Competitive compensation
Meaningful equity
Comprehensive benefits
+1
Performance Engineer — AI Inference Systems
Performance Engineer — AI Inference Systems

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy
Member of Technical Staff — Inference-Kernel, Compiler & Communication
Member of Technical Staff — Inference-Kernel, Compiler & Communication

RadixArk • Palo Alto (CA)

On-site
USD 210,000 - 290,000
Competitive compensation
Comprehensive benefits
Flexible work arrangements
Member of Technical Staff — Inference-Multi-HardwarePalo Alto, CA
Member of Technical Staff — Inference-Multi-HardwarePalo Alto, CA

RadixArk • Palo Alto (CA)

On-site
USD 210,000 - 310,000
Member of Technical Staff — Heterogenous Hardware
Member of Technical Staff — Heterogenous Hardware

The Consensus • Palo Alto (CA), Northern (KY)

Hybrid
USD 200,000 - 400,000
Member of Technical Staff — Heterogenous Hardware
Member of Technical Staff — Heterogenous Hardware

RadixArk • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Competitive compensation
Equity
Flexible work arrangements
+1
Member of Technical Staff — Developer TechnologyPalo Alto, CA
Member of Technical Staff — Developer TechnologyPalo Alto, CA

RadixArk • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff — TrainingPalo Alto, CA
Member of Technical Staff — TrainingPalo Alto, CA

RadixArk • Palo Alto (CA)

On-site
USD 240,000 - 320,000
Equity
Flexible work arrangements
Comprehensive benefits