Performance Profiling Engineer: Build Chip-Level Profilers

Infinity Artificial Intelligence Institute

San Francisco (CA)

On-site

USD 150,000 - 210,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Infinity is an early-stage AI infrastructure research company building software that makes non-NVIDIA chips competitive for AI inference. We’re hiring a performance-analysis engineer to build a profiler across chips, with per-operation attribution and automation that exposes bottlenecks.

Based in San Francisco, this full-time on-site role requires deep systems background, Python and Rust/C/C++, and experience building measurement tools when the docs are thin.

Qualifications

  • Real performance-analysis experience profiling hard problems.
  • Low-level systems background with counters, tracing, and instrumentation.
  • Ability to build measurement tools when documentation is thin.
  • Statistical care about noise and sample size before reporting results.
  • Fluency in Python and at least one systems language (Rust, C, or C++).

Responsibilities

  • Profiler-generation: build a profiler for a supported chip and ensure visibility across the stack.
  • Per-operation attribution: break down an inference pass to show which operation costs time.
  • Reimplement surface-area tooling per chip (nvprof/Nsight) for accelerators without profilers.
  • Counter and telemetry discovery: identify signals and validate counters for performance.
  • Faithful instrumentation: measure without perturbing timings.
  • Bottleneck surfacing: translate measurements into exact blockers for optimization.

Skills

Performance-analysis
Low-level systems
Instrumentation
Python
Rust/C/C++

Tools

perf
VTune
Nsight

Job description

Infinity is an early-stage AI infrastructure research company building software that makes non-NVIDIA chips competitive for AI inference. We’re hiring a performance-analysis engineer to build a profiler across chips, with per-operation attribution and automation that exposes bottlenecks.

Based in San Francisco, this full-time on-site role requires deep systems background, Python and Rust/C/C++, and experience building measurement tools when the docs are thin.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff - Performance Profiling
Member of Technical Staff - Performance Profiling

Infinity Artificial Intelligence Institute • San Francisco (CA)

On-site
USD 150,000 - 210,000
Performance Tools Intern — Profiling a Next-Gen ML Accelerator
Performance Tools Intern — Profiling a Next-Gen ML Accelerator

The Consensus • San Jose (CA)

On-site
USD 34,000 - 62,000
Performance Tools Engineer for ML Accelerator Profiling
Performance Tools Engineer for ML Accelerator Profiling

Etched • San Jose (CA)

On-site
USD 150,000 - 275,000
AI Accelerator Performance Profiling Architect
AI Accelerator Performance Profiling Architect

Etched • San Jose (CA)

On-site
USD 200,000 - 300,000
Embedded Systems Performance Engineer for AI Devices
Embedded Systems Performance Engineer for AI Devices

jobr.pro • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Embedded Systems Performance Engineer (AI Devices)
Embedded Systems Performance Engineer (AI Devices)

OpenAI • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Senior AI Performance Profiling Engineer - Equity & Benefits
Senior AI Performance Profiling Engineer - Equity & Benefits

2100 NVIDIA USA • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Performance Tools Intern
Performance Tools Intern

Etched • San Jose (CA)

On-site
USD 20,000 - 31,000
Performance Tools Intern
Performance Tools Intern

The Consensus • San Jose (CA)

On-site
USD 34,000 - 62,000
Hardware Performance Engineer — AI Accelerators (Equity)
Hardware Performance Engineer — AI Accelerators (Equity)

Artificial Analysis, Inc. • San Francisco (CA)

On-site
USD 150,000 - 190,000
Equity
Competitive compensation