Lead AI Accelerator Performance Architect

Morr0

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

19 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Morr0 is seeking a Principal Performance Modeling Architect in the Bay Area (onsite). You will own the performance analysis framework used to evaluate proposed silicon and system architectures before they are built.

You'll collaborate with senior GPU/AI architects to shape accelerator compute, memory, interconnect, and cluster-level systems, developing production-quality modeling tools in Python and C++. This hands-on IC role will influence RTL decisions and architecture choices, ensuring

Qualifications

  • Experience in GPU/TPU/NPU performance modeling.
  • Background in architecture simulation or analytical modeling for AI accelerators.
  • Experience with LLM training/inference performance.
  • Knowledge of memory hierarchy and bandwidth modeling.
  • Familiarity with NoC, fabric, and interconnect performance.
  • Experience with Python and C++ modeling frameworks.
  • Understanding roofline and bottleneck analysis.

Responsibilities

  • Build end-to-end performance models for AI accelerator platforms.
  • Model workloads across silicon, memory, interconnect, and multi-accelerator systems.
  • Translate proposals into projections of performance, power, and cost.
  • Identify bottlenecks and quantify architectural trade-offs before RTL/silicon.
  • Model LLM inference workloads including batching and KV-cache management.
  • Develop production-quality analytical frameworks in Python/C++.

Skills

Performance modeling
GPU/TPU/NPU modeling
Architecture analysis
LLM inference performance
System-level optimization

Tools

Python
C++

Job description

Morr0 is seeking a Principal Performance Modeling Architect in the Bay Area (onsite). You will own the performance analysis framework used to evaluate proposed silicon and system architectures before they are built.

You'll collaborate with senior GPU/AI architects to shape accelerator compute, memory, interconnect, and cluster-level systems, developing production-quality modeling tools in Python and C++. This hands-on IC role will influence RTL decisions and architecture choices, ensuring

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Performance Modeling Architect
Performance Modeling Architect

Morr0 • San Francisco (CA)

On-site
USD 180,000 - 240,000
Lead AI Accelerator Performance Architect
Lead AI Accelerator Performance Architect

New Grad 2026 @ Cerebras Systems • Sunnyvale (CA)

On-site
USD 175,000 - 275,000
AI Performance Architect & Modeling Lead
AI Performance Architect & Modeling Lead

D-Matrix Corp. • Santa Clara (CA), Northern (KY)

Hybrid
USD 150,000 - 230,000
Performance Modeling Lead
Performance Modeling Lead

OpenAI • Los Angeles (CA)

On-site
USD 130,000 - 180,000
Relocation assistance
Hybrid work model
Principal AI Accelerator Performance Architect
Principal AI Accelerator Performance Architect

Cerebras • Sunnyvale (CA)

On-site
USD 175,000 - 275,000
Lead Architect, AI Performance & Modeling (Hybrid)
Lead Architect, AI Performance & Modeling (Hybrid)

d-Matrix inc. • Santa Clara (CA)

Hybrid
USD 150,000 - 200,000
Lead AI Inference Performance Architect
Lead AI Inference Performance Architect

Acceler8 Talent • San Francisco (CA)

On-site
USD 120,000 - 160,000
Opportunity to shape next-generation AI inference infrastructure
High-impact technical ownership
Work in a fast-moving engineering environment
AI Accelerator Performance Modeling Engineer
AI Accelerator Performance Modeling Engineer

MatX • Mountain View (CA)

On-site
USD 160,000 - 600,000
4 weeks PTO
Remote work up to 3 weeks
Health insurance
+9
AI Hardware Performance Modeling Architect (Hybrid/Remote)
AI Hardware Performance Modeling Architect (Hybrid/Remote)

d-Matrix • Santa Clara (CA)

Hybrid
USD 170,000 - 210,000
Hardware Performance Engineer, AI Accelerators
Hardware Performance Engineer, AI Accelerators

River AI • Palo Alto (CA)

On-site
USD 200,000 - 420,000
Health, dental, vision benefits
Unlimited PTO
Relocation support