AI Compiler Engineer: Graph Optimizations & HW/SW Co-Design

Ampere Computing

Santa Clara (CA)

Hybrid

USD 195,000 - 292,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
401K retirement plan
Unlimited flextime
Paid holidays

Job summary

Ampere Computing is seeking a Software Principal Engineer- AI Compiler in California to optimize DL graphs for our energy-efficient AI accelerator. You will work across the SW/HW stack—from inference serving and framework integration to compiler, runtime, and compute kernels—to unlock performance.

You’ll enable PyTorch and Llama.cpp models, implement graph-level optimizations like fusion, and collaborate on HW/SW co-design to push efficiency on Ampere's hardware. 5+ years of exp preferred.

Qualifications

  • Strong CS fundamentals: algorithms, data structures, systems.
  • Proficiency in Python and C/C++.
  • Experience optimizing deep learning graphs and compilers is a plus.
  • Competitive programming achievements (IOI/ACM/USACO) are a big plus.

Responsibilities

  • Optimize deep learning computational graphs for performance, throughput, and latency on Ampere's accelerator hardware.
  • Enable popular models and frameworks (PyTorch, Llama.cpp) and serving platforms (vLLM, SGLang).
  • Identify and implement graph-level optimizations: op fusion, pattern recognition, redundancy elimination.
  • Collaborate on HW/SW co-design to push computational efficiency.
  • Work with cross-functional teams to integrate AI solutions into Ampere's AI hardware platforms.

Skills

Python
C/C++
Algorithms
Data structures
Systems
Competitive programming
MLIR

Education

Bachelor's degree in Computer Science or Mathematics
Master's degree

Tools

LLVM

Job description

Ampere Computing is seeking a Software Principal Engineer- AI Compiler in California to optimize DL graphs for our energy-efficient AI accelerator. You will work across the SW/HW stack—from inference serving and framework integration to compiler, runtime, and compute kernels—to unlock performance.

You’ll enable PyTorch and Llama.cpp models, implement graph-level optimizations like fusion, and collaborate on HW/SW co-design to push efficiency on Ampere's hardware. 5+ years of exp preferred.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Compiler Architect for Efficient Deep Learning
Senior AI Compiler Architect for Efficient Deep Learning

Ampere • Santa Clara (CA)

On-site
USD 195,000 - 292,000
Health insurance
401K plan
Unlimited flextime
Software Principal Engineer- AI Compiler
Software Principal Engineer- AI Compiler

Ampere Computing • Santa Clara (CA)

Hybrid
USD 195,000 - 292,000
Health insurance
401K retirement plan
Unlimited flextime
+1
Software Principal Engineer- AI Compiler
Software Principal Engineer- AI Compiler

Ampere • Santa Clara (CA)

On-site
USD 195,000 - 292,000
Health insurance
401K plan
Unlimited flextime
Senior AI Compiler Engineer - XLA & DL Graphs
Senior AI Compiler Engineer - XLA & DL Graphs

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
AI Compiler Engineer, Custom Silicon
AI Compiler Engineer, Custom Silicon

River AI Inc. • Austin (TX), Palo Alto (CA)

On-site
USD 200,000 - 420,000
Visa sponsorship
Relocation assistance
Comprehensive benefits
AI Accelerator Architect - Energy-Efficient LLM Compute
AI Accelerator Architect - Energy-Efficient LLM Compute

Normal Computing Corporation • Palo Alto (CA)

On-site
USD 240,000 - 420,000
Senior AI MLIR Compiler Engineer – GPU Inference (Equity)
Senior AI MLIR Compiler Engineer – GPU Inference (Equity)

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Senior AI Compiler Engineer (MLIR) — Equity & Impact
Senior AI Compiler Engineer (MLIR) — Equity & Impact

NVIDIA • California (MO)

On-site
USD 152,000 - 288,000
Senior Graph Optimization Engineer
Senior Graph Optimization Engineer

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 180,000 - 230,000
Equity
Medical benefits
Retirement plan
+1
Senior AI Compiler Architect — MLIR & GPU Performance
Senior AI Compiler Architect — MLIR & GPU Performance

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 288,000