Senior Edge AI Inference Engineer

Intel

Hillsboro (OR)

Hybrid

USD 195,200 - 361,200

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work model
Competitive compensation

Job summary

Intel is hiring for an AI inference engineer focusing on optimizing local inference engines (llama.cpp, vLLM) for edge devices and hybrid environments. You will work on KV cache, batching, quantization, and CPU overhead reduction to enable fast, trustworthy local AI solutions.

The role demands deep C++/Python skills, Linux expertise, and experience with LLM inference at scale. This is a hybrid position with multiple US locations including Oregon, Hillsboro, and surrounding areas.

Qualifications

  • BS/MS in CS, EE, Math or related STEM field
  • 8+ years software development background
  • Strong in C++ and/or Python; comfortable reading systems-level code
  • Experience with LLM inference and attention, KV cache, decoding
  • Experience profiling and optimizing real performance problems (CPU or GPU)
  • Linux, build systems, and low-level debugging expertise

Responsibilities

  • Profile and optimize local inference for latency, throughput, and memory on edge hardware
  • Tune KV cache, batching, and scheduling for interactive workloads
  • Drive quantization strategy and validate quality impact with the Post-Training team
  • Cut CPU overhead and improve engine startup, model load, and lifecycle
  • Benchmark across hardware tiers and publish performance comparisons
  • Upstream fixes and patches to open-source engines where it helps us

Skills

C++
Python
LLM inference
Linux
Performance optimization
Profiling

Education

BS/MS in CS, EE, Math or related STEM field

Tools

llama.cpp
vLLM
ggml
CUDA
Vulkan

Job description

Intel is hiring for an AI inference engineer focusing on optimizing local inference engines (llama.cpp, vLLM) for edge devices and hybrid environments. You will work on KV cache, batching, quantization, and CPU overhead reduction to enable fast, trustworthy local AI solutions.

The role demands deep C++/Python skills, Linux expertise, and experience with LLM inference at scale. This is a hybrid position with multiple US locations including Oregon, Hillsboro, and surrounding areas.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Edge Inference Optimization Engineer
Senior Edge Inference Optimization Engineer

Intel • Phoenix (AZ)

Hybrid
USD 195,000 - 362,000
Stock bonuses
Health insurance
Retirement plan
+1
Senior Edge Inference Optimization Engineer
Senior Edge Inference Optimization Engineer

Intel Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 217,000 - 361,000
Senior AI Inference & Kernel Engineer
Senior AI Inference & Kernel Engineer

Intel • Austin (TX)

Hybrid
USD 189,000 - 315,000
Stock bonuses
Health benefits
Vacation
Sr. Inference Optimization Engineer (local / edge runtime)
Sr. Inference Optimization Engineer (local / edge runtime)

Intel • Phoenix (AZ)

Hybrid
USD 195,000 - 362,000
Stock bonuses
Health insurance
Retirement plan
+1
Sr. Inference Optimization Engineer (local / edge runtime)
Sr. Inference Optimization Engineer (local / edge runtime)

Intel • Hillsboro (OR)

Hybrid
USD 195,200 - 361,200
Hybrid work model
Competitive compensation
Sr. Inference Optimization Engineer (local / edge runtime)
Sr. Inference Optimization Engineer (local / edge runtime)

Intel • Folsom (CA)

Hybrid
USD 195,000 - 362,000
Remote AI Inference Engineer - Edge & On-Device Systems
Remote AI Inference Engineer - Edge & On-Device Systems

Jobgether • Israel Township (OH)

On-site
USD 93,000 - 179,000
Fully remote
Cutting-edge AI
High ownership
+4
Sr. Inference Optimization Engineer (local / edge runtime)
Sr. Inference Optimization Engineer (local / edge runtime)

Intel Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 217,000 - 361,000
Senior Edge AI Inference Engineer
Senior Edge AI Inference Engineer

Quadric • Burlingame (CA)

Hybrid
USD 110,000 - 270,000
Competitive salary and meaningful-equy
Medical dental vision plans from day 1
401(k) retirement plan
+5
High-Performance AI Inference Engineer
High-Performance AI Inference Engineer

F5 Networks, Inc.  • San Jose (CA)

Hybrid
USD 177,000 - 265,000