Get more replies from employers
Send a job-specific resume in minutes.
Intel is seeking a performance-driven AI Infrastructure Engineer to push LLM inference on next-gen Intel GPUs. You will profile, optimize, and write high-performance kernels, upstream improvements into vLLM and PyTorch, and collaborate with architecture teams to shape future GPU roadmaps.
You will contribute across the stack from kernel design to open-source integration, targeting state-of-the-art generative AI workloads and multi-node scale-out setups with a strong emphasis on performance and
Intel is seeking a performance-driven AI Infrastructure Engineer to push LLM inference on next-gen Intel GPUs. You will profile, optimize, and write high-performance kernels, upstream improvements into vLLM and PyTorch, and collaborate with architecture teams to shape future GPU roadmaps.
You will contribute across the stack from kernel design to open-source integration, targeting state-of-the-art generative AI workloads and multi-node scale-out setups with a strong emphasis on performance and