Principal Performance Modeling Architect – AI Accelerators

Oho Group

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Oho Group in the San Francisco Bay Area is seeking a Principal Performance Modeling Architect to lead a system-level performance modeling platform that informs silicon architecture decisions before tape-out.

You will build and extend analytical models, evaluate AI hardware architectures, and drive design choices with simulation-backed analysis, collaborating with architecture teams and mentoring junior engineers.

Qualifications

  • 5+ years' experience in performance modeling, computer architecture, or systems performance engineering.
  • Strong understanding of AI accelerators, GPUs, HPC systems, memory hierarchies, and interconnects.
  • Experience developing analytical or first-principles performance models rather than purely benchmarking hardware.
  • Knowledge of AI inference and training workloads and how they utilise compute, memory, and networking resources.
  • Strong Python development skills with experience building maintainable engineering tools.
  • Ability to analyse complex simulation results and influence architecture decisions through data-driven insights.
  • Excellent communication skills with the ability to explain technical concepts clearly.

Responsibilities

  • Own and develop a system-level performance modeling platform for AI accelerator and cluster architectures.
  • Build analytical models to predict performance, power, energy efficiency, and cost across different hardware designs.
  • Expand support beyond LLM inference to training, multimodal, and vision workloads.
  • Work closely with architecture teams to evaluate new silicon concepts and guide design decisions using simulation-backed analysis.
  • Validate models against real-world AI frameworks and continuously improve accuracy.
  • Improve platform scalability, automation, and maintainability.
  • Provide technical guidance to junior engineers and collaborate with customers and partners where required.

Skills

Performance modeling
Computer architecture
Python
Data analysis
Communication

Job description

Oho Group in the San Francisco Bay Area is seeking a Principal Performance Modeling Architect to lead a system-level performance modeling platform that informs silicon architecture decisions before tape-out.

You will build and extend analytical models, evaluate AI hardware architectures, and drive design choices with simulation-backed analysis, collaborating with architecture teams and mentoring junior engineers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Performance Modeling Engineer
Principal Performance Modeling Engineer

Oho Group • San Francisco (CA)

On-site
USD 180,000 - 240,000
Performance Modeling Lead
Performance Modeling Lead

OpenAI • Los Angeles (CA)

Hybrid
USD 130,000 - 180,000
Relocation assistance
Hybrid work model
Principal AI Performance Architect & Modeling Lead
Principal AI Performance Architect & Modeling Lead

d-Matrix • Seattle (WA)

Hybrid
USD 180,000 - 240,000
AI Systems Performance Modeling Architect
AI Systems Performance Modeling Architect

Velaura • Santa Clara (CA)

On-site
USD 200,000 - 500,000
Lead Architect, AI Performance & Modeling (Hybrid)
Lead Architect, AI Performance & Modeling (Hybrid)

d-Matrix inc. • Santa Clara (CA)

Hybrid
USD 150,000 - 200,000
Performance Modeling Lead
Performance Modeling Lead

OpenAI • Seattle (WA)

On-site
USD 342,000 - 555,000
Senior AI System Performance Architect
Senior AI System Performance Architect

Oxmiq Labs • Campbell (CA)

On-site
USD 180,000 - 240,000
Senior Performance Modeling Architect - C++ AI Simulations
Senior Performance Modeling Architect - C++ AI Simulations

Apple • Austin (TX)

On-site
USD 180,000 - 240,000
Lead AI Inference Performance Architect
Lead AI Inference Performance Architect

Acceler8 Talent • San Francisco (CA)

On-site
USD 120,000 - 160,000
Opportunity to shape next-generation AI inference infrastructure
High-impact technical ownership
Work in a fast-moving engineering environment
Performance Modeling Engineer ~2
Performance Modeling Engineer ~2

OpenAI • Los Angeles (CA)

Hybrid
USD 266,000 - 445,000