Get more replies from employers
Send a job-specific resume in minutes.
Intel is seeking an AI Infrastructure Engineer in Santa Clara to push LLM inference on Intel GPUs, optimizing the end-to-end stack from profiling to kernel development. You will upstream improvements to open-source platforms like vLLM, SGLang, and PyTorch and collaborate with architecture teams to shape future GPU roadmaps.
The role requires strong C++/Python skills and 4+ years in GPU computing or HPC, with a passion for performance and scale-out inference across multi-node setups.
Intel is seeking an AI Infrastructure Engineer in Santa Clara to push LLM inference on Intel GPUs, optimizing the end-to-end stack from profiling to kernel development. You will upstream improvements to open-source platforms like vLLM, SGLang, and PyTorch and collaborate with architecture teams to shape future GPU roadmaps.
The role requires strong C++/Python skills and 4+ years in GPU computing or HPC, with a passion for performance and scale-out inference across multi-node setups.