Performance Tools Intern — Profiling a Next-Gen ML Accelerator

The Consensus

San Jose (CA)

On-site

USD 34,000 - 62,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Etched in San Jose is seeking a talented intern to design and develop a performance analysis tool for our ML accelerator. You will build tooling to understand workload behavior, identify bottlenecks, and help optimize hardware and software interactions.

Work with hardware, compiler, firmware, and inference engineers in an in-person setting. Strong coding skills in C++ or Rust, plus Python, will help you contribute from day one.

Qualifications

  • Proficient in C++ or Rust with strong systems programming background.
  • Interest or experience in low-level performance analysis and profiling.
  • Familiarity with hardware performance counters, traces, and CPUs/accelerators.

Responsibilities

  • Build components of our performance analysis and profiling infrastructure.
  • Collect and analyze performance data from custom ML accelerators, including hardware counters, execution traces, and memory behavior.
  • Develop tooling to trace host-side runtime activity, system behavior, and accelerator execution.
  • Help correlate performance events across CPUs, accelerators, storage, networking, and distributed workloads.
  • Build analysis and visualization tools that help engineers identify performance bottlenecks and optimize models.
  • Work alongside hardware, compiler, firmware, and inference engineers to understand performance challenges and develop tools that improve developer productivity.

Skills

C++
Rust
Python

Tools

Nsight
VTune
Perfetto
Xprof

Job description

Etched in San Jose is seeking a talented intern to design and develop a performance analysis tool for our ML accelerator. You will build tooling to understand workload behavior, identify bottlenecks, and help optimize hardware and software interactions.

Work with hardware, compiler, firmware, and inference engineers in an in-person setting. Strong coding skills in C++ or Rust, plus Python, will help you contribute from day one.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Accelerator Performance Tools Intern
ML Accelerator Performance Tools Intern

Etched • San Jose (CA)

On-site
USD 20,000 - 31,000
Performance Tools Intern
Performance Tools Intern

The Consensus • San Jose (CA)

On-site
USD 34,000 - 62,000
Performance Tools Intern
Performance Tools Intern

Etched • San Jose (CA)

On-site
USD 20,000 - 31,000
Performance Tools Engineer for ML Accelerator Profiling
Performance Tools Engineer for ML Accelerator Profiling

Etched • San Jose (CA)

On-site
USD 150,000 - 275,000
Performance Analysis Lead for ML Accelerator Hardware
Performance Analysis Lead for ML Accelerator Hardware

The Consensus • San Jose (CA)

On-site
USD 180,000 - 260,000
Medical coverage
Housing subsidy
Relocation support
+3
Performance Tools Engineer, ML Accelerator
Performance Tools Engineer, ML Accelerator

Etched.ai, Inc. • San Jose (CA)

On-site
USD 120,000 - 160,000
Medical, dental, and vision packages
Housing subsidy
Relocation support
+1
Performance Profiling Engineer: Build Chip-Level Profilers
Performance Profiling Engineer: Build Chip-Level Profilers

Infinity Artificial Intelligence Institute • San Francisco (CA)

On-site
USD 150,000 - 210,000
AI Systems Engineer: High-Performance ML on Accelerators
AI Systems Engineer: High-Performance ML on Accelerators

AMD • San Jose (CA)

On-site
USD 150,000 - 190,000
Software Engineer – Performance Profiling
Software Engineer – Performance Profiling

The Consensus • San Jose (CA)

On-site
USD 180,000 - 260,000
Medical coverage
Housing subsidy
Relocation support
+3
ML Accelerator Performance Engineer — Real ML Benchmarks
ML Accelerator Performance Engineer — Real ML Benchmarks

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 143,000 - 195,000
Sign-on payments
Restricted stock units (RSUs)
Health insurance
+3