Performance Modeling Architect for AI Hardware

Oho Group

San Francisco (CA)

On-site

USD 180,000 - 260,000

Full time

31 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Oho Group in San Francisco seeks a senior performance-modeling engineer to build analytical and simulation-based models that guide architectural decisions for AI workloads before hardware exists.

Your work will span processors, accelerators, memory hierarchies and interconnects, analyzing transformer inference across single-device and multi-accelerator systems, and evaluating latency, throughput, energy and cost trade-offs.

Qualifications

  • Strong Python and/or C++ development.
  • Experience building performance models.
  • Deep computer-architecture and microarchitecture knowledge.
  • Understanding of compute pipelines, memory systems and data movement.
  • Ability to convert model results into concrete architecture decisions.

Responsibilities

  • Build analytical and simulation-based performance models.
  • Model processors, accelerators, memory hierarchies and interconnects.
  • Analyze transformer inference across single-device and multi-accelerator systems.
  • Evaluate latency, throughput, utilization, energy and cost trade-offs.
  • Characterize workload behavior using traces and representative benchmarks.
  • Identify architectural bottlenecks and propose measurable improvements.
  • Partner with hardware, compiler, runtime and inference teams.
  • Correlate models against simulation, emulation or silicon as the platform matures.

Skills

Python
C++
Performance modeling
Computer architecture

Job description

Oho Group in San Francisco seeks a senior performance-modeling engineer to build analytical and simulation-based models that guide architectural decisions for AI workloads before hardware exists.

Your work will span processors, accelerators, memory hierarchies and interconnects, analyzing transformer inference across single-device and multi-accelerator systems, and evaluating latency, throughput, energy and cost trade-offs.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Performance Modeling Engineer
Performance Modeling Engineer

Oho Group • San Francisco (CA)

On-site
USD 180,000 - 260,000
Performance Modeling Engineer - AI Systems & Infrastructure
Performance Modeling Engineer - AI Systems & Infrastructure

OpenAI • Seattle (WA)

Hybrid
USD 293,000 - 385,000
Relocation assistance
Hybrid work model
Performance Modeling Lead
Performance Modeling Lead

OpenAI • Los Angeles (CA)

Hybrid
USD 130,000 - 180,000
Relocation assistance
Hybrid work model
Performance Modeling Engineer ~2
Performance Modeling Engineer ~2

OpenAI • Los Angeles (CA)

Hybrid
USD 266,000 - 445,000
AI Systems Performance Architect
AI Systems Performance Architect

Actalent • Raleigh (NC)

Hybrid
USD 200,000 - 500,000
Performance Modeling Lead
Performance Modeling Lead

OpenAI • Seattle (WA)

On-site
USD 342,000 - 555,000
Performance Modeling Engineer
Performance Modeling Engineer

OpenAI • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Relocation assistance
Senior Performance Modeling Engineer - AI Hardware
Senior Performance Modeling Engineer - AI Hardware

Socket.dev • Santa Clara (CA)

On-site
USD 170,000 - 250,000
Health, Dental, Vision coverage
401(k) match
Equity grant
+2
Performance Modeling Engineer
Performance Modeling Engineer

OpenAI • Los Angeles (CA)

Hybrid
USD 100,000 - 150,000
Relocation assistance
Hybrid work model
Performance Modeling Engineer - AI Memory Systems
Performance Modeling Engineer - AI Memory Systems

Netpreme • Santa Clara (CA)

On-site
USD 180,000 - 260,000
Health, Dental, and Vision coverage
Early-stage equity grant
401(k) match
+2