Stand out for this role — generate a tailored resume and cover letter in about a minute.
Qualcomm Technologies, Inc. seeks a Staff Engineer to lead end-to-end AI model optimization for LLMs, VLMs, and diffusion on Qualcomm accelerators. You will transform PyTorch models, manage KVcache behavior, and drive deployment with PyTorch, ONNX, and torch.compile across multi-core systems.
You will collaborate with compiler and performance teams to craft lowering strategies and scalable tooling, ensuring high throughput with low latency and production-grade reliability.
Qualcomm Technologies, Inc. seeks a Staff Engineer to lead end-to-end AI model optimization for LLMs, VLMs, and diffusion on Qualcomm accelerators. You will transform PyTorch models, manage KVcache behavior, and drive deployment with PyTorch, ONNX, and torch.compile across multi-core systems.
You will collaborate with compiler and performance teams to craft lowering strategies and scalable tooling, ensuring high throughput with low latency and production-grade reliability.