Performance Modeling Engineer - AI Memory Systems

Netpreme

Santa Clara (CA)

On-site

USD 180,000 - 260,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health, Dental, and Vision coverage
Early-stage equity grant
401(k) match
Daily lunch stipend
On-site parking with EV charging

Job summary

Netpreme is seeking a Member of Technical Staff, Functional and Performance Modeling to develop functional and performance models for our scale-up memory expansion device for AI accelerators. You’ll work with the silicon architecture team to explore design tradeoffs and validate performance assumptions.

This onsite role is based in Santa Clara, CA or Boston, MA, and involves modeling system- and chip-level data movement, identifying bottlenecks, and communicating results to stakeholders across

Qualifications

  • Bachelor’s or Master’s degree in Electrical Engineering, Computer Engineering, or a closely related field.
  • 5–10+ years of experience in performance modeling for data movement devices: NICs, memory expansion cards (e.g. CXL), IPU/DPU, NoC.
  • Ability to reason across multiple abstraction layers, from architectural details to system-level performance behavior.

Responsibilities

  • Build and maintain system-level (e.g., rack-scale) and chip-level performance models for high-bandwidth data movement between devices operating in the scale-up domain.
  • Model workloads from software memory access patterns through data distribution in the network and all the way down to on-device memory channels.
  • Work day-to-day with silicon architects, system designers, and workload owners to align performance expectations and constraints.
  • Identify performance bottlenecks, scaling limits, and sensitivity points across compute, memory, and interconnects in end-to-end workload settings.
  • Clearly communicate modeling assumptions, limitations, and conclusions to both technical and non-specialist stakeholders.

Skills

Performance modeling
System-level modeling
Multi-layer reasoning
Data movement
NVLink

Education

Bachelor’s or Master’s degree in Electrical Engineering, Computer Engineering, or closely related field

Tools

CUDA
NoC
NVLink

Job description

Netpreme is seeking a Member of Technical Staff, Functional and Performance Modeling to develop functional and performance models for our scale-up memory expansion device for AI accelerators. You’ll work with the silicon architecture team to explore design tradeoffs and validate performance assumptions.

This onsite role is based in Santa Clara, CA or Boston, MA, and involves modeling system- and chip-level data movement, identifying bottlenecks, and communicating results to stakeholders across

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Performance Modeling Engineer, AI Hardware Systems
Staff Performance Modeling Engineer, AI Hardware Systems

Netpreme • Boston (MA)

On-site
USD 150,000 - 230,000
Health coverage
Dental coverage
Vision coverage
+12
Senior Performance Modeling Engineer - AI Hardware
Senior Performance Modeling Engineer - AI Hardware

Socket.dev • Santa Clara (CA)

On-site
USD 170,000 - 250,000
Health, Dental, Vision coverage
401(k) match
Equity grant
+2
Member of Technical Staff, Performance Modeling
Member of Technical Staff, Performance Modeling

Netpreme • Boston (MA)

On-site
USD 150,000 - 230,000
Health coverage
Dental coverage
Vision coverage
+12
Member of Technical Staff, Performance Modeling
Member of Technical Staff, Performance Modeling

Socket.dev • Santa Clara (CA)

On-site
USD 170,000 - 250,000
Health, Dental, Vision coverage
401(k) match
Equity grant
+2
Performance Modeling Architect for AI Hardware
Performance Modeling Architect for AI Hardware

Oho Group • San Francisco (CA)

On-site
USD 180,000 - 260,000
Member of Technical Staff, Performance Modeling
Member of Technical Staff, Performance Modeling

Netpreme • Santa Clara (CA)

On-site
USD 180,000 - 260,000
Health, Dental, and Vision coverage
Early-stage equity grant
401(k) match
+2
Performance Modeling Engineer - AI Systems & Infrastructure
Performance Modeling Engineer - AI Systems & Infrastructure

OpenAI • Seattle (WA)

Hybrid
USD 293,000 - 385,000
Relocation assistance
Hybrid work model
Performance Modeling Engineer
Performance Modeling Engineer

Oho Group • San Francisco (CA)

On-site
USD 180,000 - 260,000
AI Systems Performance Architect
AI Systems Performance Architect

Actalent • Raleigh (NC)

Hybrid
USD 200,000 - 500,000
Performance Modeling Engineer
Performance Modeling Engineer

OpenAI • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Relocation assistance