Get more replies from employers
Send a job-specific resume in minutes.
Amazon in Cupertino, CA, is seeking a Senior Development Manager for LLM Inference Model Enablement to lead a team of AI/ML engineers focused on onboarding and optimizing state-of-the-art open-source and customer LLMs for Trainium accelerators.
You will drive model enablement speed, usability, and quality across the PyTorch inference library, Neuron compiler, runtime, and collectives, while managing priorities and delivering high-performance inference on Trainium hardware.
Amazon in Cupertino, CA, is seeking a Senior Development Manager for LLM Inference Model Enablement to lead a team of AI/ML engineers focused on onboarding and optimizing state-of-the-art open-source and customer LLMs for Trainium accelerators.
You will drive model enablement speed, usability, and quality across the PyTorch inference library, Neuron compiler, runtime, and collectives, while managing priorities and delivering high-performance inference on Trainium hardware.