AI Inference Systems Performance Engineer

Tensordyne

Sunnyvale (CA)

On-site

USD 140,000 - 210,000

Full time

9 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Comprehensive benefits
Flexible spending options
Recognition programs

Job summary

Tensordyne in Sunnyvale, CA, seeks a Systems Performance Modeling Engineer to build models and tools predicting AI inference workloads on our systems from single accelerators to rack-scale deployments. This hands-on role involves writing simulator code, running experiments, and analyzing discrepancies between predictions and measurements.

You'll collaborate across architecture, silicon, hardware, and software teams to extend models of silicon, interconnects, and fabric, capture workload

Qualifications

  • Hands-on experience building performance models, simulators, or analytical tools for ML workloads.
  • Strong programming skills in C++ and Python with clean, testable code.
  • Experience comparing model predictions against real measurements and debugging divergences.
  • Understanding of distributed ML execution and parallelism strategies.

Responsibilities

  • Build and extend simulation-based performance models for multimodal AI inference at rack, pod, and cluster scale.
  • Model batching, KV-cache management and disaggregation effects on latency and throughput.
  • Create trace-capture and replay tooling for hardware configurations.
  • Run calibration experiments on Tensordyne hardware and compare to model predictions.
  • Provide design-space analyses and performance projections for decisions.

Skills

C++
Python
Distributed systems
Performance modeling
Debugging
Communication

Education

MS or higher in CS/CE/EE or related field

Tools

NCCL
Python testing
Simulation tools
Profiling ML workloads

Job description

Tensordyne in Sunnyvale, CA, seeks a Systems Performance Modeling Engineer to build models and tools predicting AI inference workloads on our systems from single accelerators to rack-scale deployments. This hands-on role involves writing simulator code, running experiments, and analyzing discrepancies between predictions and measurements.

You'll collaborate across architecture, silicon, hardware, and software teams to extend models of silicon, interconnects, and fabric, capture workload

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead AI Inference Performance Architect
Lead AI Inference Performance Architect

Acceler8 Talent • San Francisco (CA)

On-site
USD 120,000 - 160,000
Opportunity to shape next-generation AI inference infrastructure
High-impact technical ownership
Work in a fast-moving engineering environment
Performance Modeling Engineer - AI Systems & Infrastructure
Performance Modeling Engineer - AI Systems & Infrastructure

OpenAI • Seattle (WA)

Hybrid
USD 293,000 - 385,000
Relocation assistance
Hybrid work model
Performance Modeling Engineer for AI Infrastructure
Performance Modeling Engineer for AI Infrastructure

OpenAI • United States

Hybrid
USD 140,000 - 210,000
Relocation assistance
Hybrid work model
Principal AI Accelerator Performance Architect
Principal AI Accelerator Performance Architect

Cerebras • Sunnyvale (CA)

On-site
USD 175,000 - 275,000
Senior Hardware Design Engineer, AI Inference Systems
Senior Hardware Design Engineer, AI Inference Systems

Tensordyne • Sunnyvale (CA)

On-site
USD 120,000 - 150,000
Comprehensive benefits
Competitive compensation
Flexible spending options
Senior ASIC Design Engineer — AI Inference Hardware
Senior ASIC Design Engineer — AI Inference Hardware

Tensordyne • Sunnyvale (CA)

On-site
USD 180,000 - 260,000
Meals
Snacks & drinks
Unlimited PTO
+1
Lead AI Accelerator Performance Architect
Lead AI Accelerator Performance Architect

New Grad 2026 @ Cerebras Systems • Sunnyvale (CA)

On-site
USD 175,000 - 275,000
AI Performance Architect & Modeling Lead
AI Performance Architect & Modeling Lead

D-Matrix Corp. • Santa Clara (CA), Northern (KY)

Hybrid
USD 150,000 - 230,000
AI Systems Performance Engineer
AI Systems Performance Engineer

OpenAI • Los Angeles (CA)

Hybrid
USD 100,000 - 150,000
Relocation assistance
Hybrid work model
AI Hardware Performance Engineer – SoC & Simulator
AI Hardware Performance Engineer – SoC & Simulator

River AI Inc. • Palo Alto (CA), Austin (TX)

On-site
USD 200,000 - 420,000
Health benefits
Unlimited PTO
Relocation assistance