Get more replies from employers
Send a job-specific resume in minutes.
NVIDIA is building the software stack for fast LLM inference on edge AI hardware in Westford, MA. This role focuses on evaluating open-source inference frameworks and mapping architectures to NVIDIA GPUs to maximize throughput and minimize latency.
You will own validation workflows, develop recipes, and collaborate with the community and partners to solve hardware-specific inference challenges, with equity and benefits on offer.
NVIDIA is building the software stack for fast LLM inference on edge AI hardware in Westford, MA. This role focuses on evaluating open-source inference frameworks and mapping architectures to NVIDIA GPUs to maximize throughput and minimize latency.
You will own validation workflows, develop recipes, and collaborate with the community and partners to solve hardware-specific inference challenges, with equity and benefits on offer.