AI Performance Architect & Modeling Lead

D-Matrix Corp.

Santa Clara, Northern (CA, KY)

Hybrid

USD 150,000 - 230,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

d-Matrix is seeking a Principal Software Engineer to advance performance analysis and modeling across its AI inference accelerators. You will analyze ML workloads, build analytical models, and extend simulators, collaborating with hardware design, compiler, inference server, kernels and product teams.

Based at our Santa Clara HQ with a hybrid onsite schedule (3 days/week) and possible remote options across the US, you will translate workload analysis into modeling inputs and identify HW/SW

Qualifications

  • BSEE with 6+ years of industry experience, or MSEE with 4+ years.
  • Working knowledge of computer architecture, HW/SW co-design, performance modeling, and ML fundamentals (DNNs).
  • Programming fluency in C/C++ or Python.
  • Experience building analytical performance models or architecture simulators.
  • Self-motivated and collaborative, across hardware and software teams.

Responsibilities

  • Analyze emerging ML workloads and multi-modal LLMs to identify performance-relevant properties.
  • Build and maintain analytical performance models projecting behavior on current/future hardware generations.
  • Develop and extend architecture simulators for performance analysis of HW/SW features.
  • Collaborate with hardware design, compiler, inference server, kernel, and product teams to validate modeling assumptions.
  • Track ML architecture and algorithm research and incorporate findings into modeling work.
  • Propose HW/SW feature improvements based on modeling results and workload analysis.
  • Document modeling methodology and findings for reuse across the team.

Skills

Computer architecture
HW/SW co-design
Performance modeling
ML fundamentals
C/C++ or Python
Analytical performance models
Collaboration

Education

BSEE with 6+ years OR MSEE with 4+ years

Job description

d-Matrix is seeking a Principal Software Engineer to advance performance analysis and modeling across its AI inference accelerators. You will analyze ML workloads, build analytical models, and extend simulators, collaborating with hardware design, compiler, inference server, kernels and product teams.

Based at our Santa Clara HQ with a hybrid onsite schedule (3 days/week) and possible remote options across the US, you will translate workload analysis into modeling inputs and identify HW/SW

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Hardware Performance Modeling Architect (Hybrid/Remote)
AI Hardware Performance Modeling Architect (Hybrid/Remote)

d-Matrix • Santa Clara (CA)

Hybrid
USD 170,000 - 210,000
Lead Architect, AI Performance & Modeling (Hybrid)
Lead Architect, AI Performance & Modeling (Hybrid)

d-Matrix inc. • Santa Clara (CA)

Hybrid
USD 150,000 - 200,000
Principal Architect, Performance Analysis and Modeling
Principal Architect, Performance Analysis and Modeling

d-Matrix • Santa Clara (CA)

Hybrid
USD 170,000 - 210,000
Principal Architect, Performance Analysis and Modeling
Principal Architect, Performance Analysis and Modeling

D-Matrix Corp. • Santa Clara (CA), Northern (KY)

Hybrid
USD 150,000 - 230,000
Principal Architect, Performance Analysis and Modeling
Principal Architect, Performance Analysis and Modeling

d-Matrix inc. • Santa Clara (CA)

Hybrid
USD 150,000 - 200,000
Principal AI Kernel Engineer — HW/SW Co-Design (Hybrid)
Principal AI Kernel Engineer — HW/SW Co-Design (Hybrid)

D-Matrix Corp. • Santa Clara (CA), Northern (KY)

Hybrid
USD 250,000 - 350,000
Senior AI Inference Systems Engineer - Hybrid, Santa Clara
Senior AI Inference Systems Engineer - Hybrid, Santa Clara

Entrada Ventures • Santa Clara (CA)

Hybrid
USD 180,000 - 260,000
Senior AI Systems Performance Modeling Architect
Senior AI Systems Performance Modeling Architect

Oho Group • San Francisco (CA)

On-site
USD 180,000 - 300,000
Performance Modeling Lead
Performance Modeling Lead

OpenAI • Los Angeles (CA)

Hybrid
USD 130,000 - 180,000
Relocation assistance
Hybrid work model
Performance Modeling Architect
Performance Modeling Architect

Oho Group • San Francisco (CA)

On-site
USD 180,000 - 300,000