Get more replies from employers
Send a job-specific resume in minutes.
Intel is seeking an AI Infrastructure Engineer to push LLM inference on Intel GPUs to new limits in Santa Clara and across multiple US locations. You will profile bottlenecks, write high-performance GPU kernels, and upstream improvements to open-source projects like vLLM and PyTorch.
You will work end-to-end across the stack, shape hardware roadmaps with roofline analysis, and collaborate with architecture and compiler teams to optimize performance for state-of-the-art generative AI workloads.
Intel is seeking an AI Infrastructure Engineer to push LLM inference on Intel GPUs to new limits in Santa Clara and across multiple US locations. You will profile bottlenecks, write high-performance GPU kernels, and upstream improvements to open-source projects like vLLM and PyTorch.
You will work end-to-end across the stack, shape hardware roadmaps with roofline analysis, and collaborate with architecture and compiler teams to optimize performance for state-of-the-art generative AI workloads.