Get more replies from employers
Send a job-specific resume in minutes.
Ampere Computing is seeking a Software Principal Engineer- AI Compiler in California to optimize DL graphs for our energy-efficient AI accelerator. You will work across the SW/HW stack—from inference serving and framework integration to compiler, runtime, and compute kernels—to unlock performance.
You’ll enable PyTorch and Llama.cpp models, implement graph-level optimizations like fusion, and collaborate on HW/SW co-design to push efficiency on Ampere's hardware. 5+ years of exp preferred.
Ampere Computing is seeking a Software Principal Engineer- AI Compiler in California to optimize DL graphs for our energy-efficient AI accelerator. You will work across the SW/HW stack—from inference serving and framework integration to compiler, runtime, and compute kernels—to unlock performance.
You’ll enable PyTorch and Llama.cpp models, implement graph-level optimizations like fusion, and collaborate on HW/SW co-design to push efficiency on Ampere's hardware. 5+ years of exp preferred.