Get more replies from employers
Send a job-specific resume in minutes.
Intel is seeking an AI Infrastructure Engineer to push LLM inference on next‑generation GPUs. You will profile stacks, write high‑performance kernels, and upstream optimizations to vLLM, SGLang, and PyTorch, shaping Intel hardware for GenAI workloads.
This hybrid role is based in the United States with primary location Santa Clara, CA, and includes multiple US sites. Strong C++/Python and HPC experience are required, with 4+ years in GPU computing or AI systems.
Intel is seeking an AI Infrastructure Engineer to push LLM inference on next‑generation GPUs. You will profile stacks, write high‑performance kernels, and upstream optimizations to vLLM, SGLang, and PyTorch, shaping Intel hardware for GenAI workloads.
This hybrid role is based in the United States with primary location Santa Clara, CA, and includes multiple US sites. Strong C++/Python and HPC experience are required, with 4+ years in GPU computing or AI systems.