An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Annapurna Labs (U.S.) Inc. in Seattle seeks an experienced software engineer to advance inference for Generative AI on AWS accelerators. You will work on distributed inference with PyTorch in the Neuron SDK, tuning models for Trainium and Inferentia, and collaborating across compiler, runtime, and hardware teams.
You will design high‑performance kernels, optimize memory and parallel computing, and contribute to a startup‑like culture that values experimentation and mentorship.
Annapurna Labs (U.S.) Inc. in Seattle seeks an experienced software engineer to advance inference for Generative AI on AWS accelerators. You will work on distributed inference with PyTorch in the Neuron SDK, tuning models for Trainium and Inferentia, and collaborating across compiler, runtime, and hardware teams.
You will design high‑performance kernels, optimize memory and parallel computing, and contribute to a startup‑like culture that values experimentation and mentorship.