Principal Performance Modeling Engineer

Oho Group

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

19 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Oho Group in the San Francisco Bay Area is seeking a Principal Performance Modeling Architect to lead a system-level performance modeling platform that informs silicon architecture decisions before tape-out.

You will build and extend analytical models, evaluate AI hardware architectures, and drive design choices with simulation-backed analysis, collaborating with architecture teams and mentoring junior engineers.

Qualifications

  • 5+ years' experience in performance modeling, computer architecture, or systems performance engineering.
  • Strong understanding of AI accelerators, GPUs, HPC systems, memory hierarchies, and interconnects.
  • Experience developing analytical or first-principles performance models rather than purely benchmarking hardware.
  • Knowledge of AI inference and training workloads and how they utilise compute, memory, and networking resources.
  • Strong Python development skills with experience building maintainable engineering tools.
  • Ability to analyse complex simulation results and influence architecture decisions through data-driven insights.
  • Excellent communication skills with the ability to explain technical concepts clearly.

Responsibilities

  • Own and develop a system-level performance modeling platform for AI accelerator and cluster architectures.
  • Build analytical models to predict performance, power, energy efficiency, and cost across different hardware designs.
  • Expand support beyond LLM inference to training, multimodal, and vision workloads.
  • Work closely with architecture teams to evaluate new silicon concepts and guide design decisions using simulation-backed analysis.
  • Validate models against real-world AI frameworks and continuously improve accuracy.
  • Improve platform scalability, automation, and maintainability.
  • Provide technical guidance to junior engineers and collaborate with customers and partners where required.

Skills

Performance modeling
Computer architecture
Python
Data analysis
Communication

Job description

Principal Performance Modeling Architect - San Francisco Bay Area - AI Accelerator Startup

About the Role

An innovative AI hardware startup developing next-generation AI accelerator technology are looking for a Principal Performance Modeling Architect to lead the development of a system-level performance modeling platform that influences silicon architecture decisions before tape-out.

This is a highly technical, hands-on individual contributor role where you'll build and extend analytical performance models, evaluate AI hardware architectures, and help shape the future of advanced AI compute systems.

What You'll Be Doing

  • Own and develop a system-level performance modeling platform for AI accelerator and cluster architectures.
  • Build analytical models to predict performance, power, energy efficiency, and cost across different hardware designs.
  • Expand support beyond LLM inference to training, multimodal, and vision workloads.
  • Work closely with architecture teams to evaluate new silicon concepts and guide design decisions using simulation-backed analysis.
  • Validate models against real-world AI frameworks and continuously improve accuracy.
  • Improve platform scalability, automation, and maintainability.
  • Provide technical guidance to junior engineers and collaborate with customers and partners where required.

What We're Looking For

  • 5+ years' experience in performance modeling, computer architecture, or systems performance engineering.
  • Strong understanding of AI accelerators, GPUs, HPC systems, memory hierarchies, and interconnects.
  • Experience developing analytical or first-principles performance models rather than purely benchmarking hardware.
  • Knowledge of AI inference and training workloads and how they utilise compute, memory, and networking resources.
  • Strong Python development skills with experience building maintainable engineering tools.
  • Ability to analyse complex simulation results and influence architecture decisions through data-driven insights.
  • Excellent communication skills with the ability to explain technical concepts clearly.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Performance Modeling Lead
Performance Modeling Lead

OpenAI • Los Angeles (CA)

Hybrid
USD 130,000 - 180,000
Relocation assistance
Hybrid work model
Principal Performance Modeling Architect – AI Accelerators
Principal Performance Modeling Architect – AI Accelerators

Oho Group • San Francisco (CA)

On-site
USD 180,000 - 240,000
Performance Modeling Lead
Performance Modeling Lead

OpenAI • Seattle (WA)

On-site
USD 342,000 - 555,000
Performance Modeling Engineer ~2
Performance Modeling Engineer ~2

OpenAI • Los Angeles (CA)

Hybrid
USD 266,000 - 445,000
Performance Modeling Lead
Performance Modeling Lead

OpenAI • San Francisco (CA)

Hybrid
USD 342,000 - 555,000
Performance Architect
Performance Architect

Acceler8 Talent • San Francisco (CA)

On-site
USD 120,000 - 160,000
Opportunity to shape next-generation AI inference infrastructure
High-impact technical ownership
Work in a fast-moving engineering environment
Performance Modeling Engineer
Performance Modeling Engineer

OpenAI • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Relocation assistance
System Performance & AI Architect
System Performance & AI Architect

Majestic Labs ai • Los Altos (CA)

On-site
USD 140,000 - 190,000
Performance Modeling Engineer ~2
Performance Modeling Engineer ~2

OpenAI • San Francisco (CA)

Hybrid
USD 266,000 - 445,000
Performance Modeling Engineer ~2
Performance Modeling Engineer ~2

OpenAI • Seattle (WA)

Hybrid
USD 266,000 - 445,000
Relocation assistance