Get more replies from employers
Send a job-specific resume in minutes.
Intel Corporation is seeking a software engineer to make models fast on hardware people own, optimizing inference engines for edge environments. You will work with llama.cpp and vLLM, tuning KV cache, batching, and quantization while reducing CPU overhead and startup costs.
You’ll benchmark across hardware tiers and contribute patches to open‑source engines, aligning with Intel’s mission to improve AI safety, privacy, and efficiency on local devices.
Intel Corporation is seeking a software engineer to make models fast on hardware people own, optimizing inference engines for edge environments. You will work with llama.cpp and vLLM, tuning KV cache, batching, and quantization while reducing CPU overhead and startup costs.
You’ll benchmark across hardware tiers and contribute patches to open‑source engines, aligning with Intel’s mission to improve AI safety, privacy, and efficiency on local devices.