Performance Analysis Lead for ML Accelerator Hardware

The Consensus

San Jose (CA)

On-site

USD 180,000 - 260,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical coverage
Housing subsidy
Relocation support
Wellness benefits
Lunch provided
Unlimited compute budget

Job summary

Etched, building hardware for frontier intelligence, is seeking a performance analysis engineer to design a tooling suite for its ML accelerator. You will enable ML engineers and customers to understand workload behavior, identify bottlenecks, and unlock full hardware potential across hosts, PCIe, and accelerators.

Collaborate with architects, firmware, drivers, compilers, and ML teams to define requirements and deliver intuitive analyses and visualizations that communicate performance

Qualifications

  • Strong proficiency in C++ or Rust.
  • Python proficiency is a plus.
  • Deep understanding of computer architecture and PCIe interconnects.
  • Experience with low-level performance analysis on complex hardware.
  • Familiar with performance analysis tools such as Nsight, VTune, and perf.

Responsibilities

  • Lead the design and architecture of a performance analysis suite.
  • Develop methods to capture performance data from custom hardware.
  • Implement tracing for host-side API calls and system events.
  • Correlate performance events across host, drivers, PCIe, accelerators, and hosts.
  • Build analysis modules to interpret trace and counter data.
  • Develop visualizations like timelines and graphs.
  • Collaborate with hardware, firmware, driver, compiler, and ML teams.

Skills

C++/Rust
Python
Computer architecture
Performance analysis
Hardware counters

Tools

NVIDIA Nsight
Intel VTune
AMD uProf
perf
Tracy
ETW

Job description

Etched, building hardware for frontier intelligence, is seeking a performance analysis engineer to design a tooling suite for its ML accelerator. You will enable ML engineers and customers to understand workload behavior, identify bottlenecks, and unlock full hardware potential across hosts, PCIe, and accelerators.

Collaborate with architects, firmware, drivers, compilers, and ML teams to define requirements and deliver intuitive analyses and visualizations that communicate performance

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Performance Tools Engineer for ML Accelerator Profiling
Performance Tools Engineer for ML Accelerator Profiling

Etched • San Jose (CA)

On-site
USD 150,000 - 275,000
Performance Tools Engineer, ML Accelerator
Performance Tools Engineer, ML Accelerator

Etched.ai, Inc. • San Jose (CA)

On-site
USD 120,000 - 160,000
Medical, dental, and vision packages
Housing subsidy
Relocation support
+1
Performance Tools Intern — Profiling a Next-Gen ML Accelerator
Performance Tools Intern — Profiling a Next-Gen ML Accelerator

The Consensus • San Jose (CA)

On-site
USD 34,000 - 62,000
ML Accelerator Performance Tools Intern
ML Accelerator Performance Tools Intern

Etched • San Jose (CA)

On-site
USD 20,000 - 31,000
Performance Tools Intern
Performance Tools Intern

The Consensus • San Jose (CA)

On-site
USD 34,000 - 62,000
Performance Tools Intern
Performance Tools Intern

Etched • San Jose (CA)

On-site
USD 20,000 - 31,000
ML Accelerator Performance Engineer — Real ML Benchmarks
ML Accelerator Performance Engineer — Real ML Benchmarks

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 143,000 - 195,000
Sign-on payments
Restricted stock units (RSUs)
Health insurance
+3
Hardware Performance Engineer — AI Accelerators (Equity)
Hardware Performance Engineer — AI Accelerators (Equity)

Artificial Analysis, Inc. • San Francisco (CA)

On-site
USD 150,000 - 190,000
Equity
Competitive compensation
Head of Performance Visibility
Head of Performance Visibility

The Consensus • San Jose (CA)

On-site
USD 180,000 - 260,000
Medical/dental/vision
Housing subsidy
Relocation assistance
+3
Head of Performance Engineering - Kernel & Platform
Head of Performance Engineering - Kernel & Platform

NVIDIA • Santa Clara (CA)

On-site
USD 272,000 - 489,000
Equity
Benefits