AI Research Engineer - Model Acceleration & Quantization

intel

Prescott (AZ)

On-site

USD 52,000 - 82,000

Full time

6 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Intel Labs China is seeking a PhD-level researcher to perform cutting-edge AI research and engineering in Bejing. You will collaborate with a world-class team to develop efficient algorithms that compress and accelerate large AI models for deployment across client, edge, cloud, and data center platforms.

The role emphasizes post-training quantization techniques, several AI inference frameworks, and strong software skills in PyTorch, C/C++, Python, and shell scripting.

Qualifications

  • Ph.D. graduate or equivalent in AI, CS, EE, Automation, or related field.
  • Strong expertise in developing AI algorithms for compressing and accelerating large models.
  • Experience with quantization of large models (4-bit, 1.58-bit) desirable.

Responsibilities

  • Conduct cutting-edge AI research and engineering at Intel Labs China.
  • Collaborate with team members to develop efficient algorithmic solutions for accelerating large AI models.
  • Focus on deploying models across Intel computing platforms for client, edge, cloud, and data center applications.

Skills

AI research
Model compression
Post-training quantization
PyTorch
C/C++
Python
Shell scripting
OpenVINO

Education

Ph.D. in AI/CS/EE/Automation

Tools

vLLM
Ollama
OpenVINO

Job description

Intel Labs China is seeking a PhD-level researcher to perform cutting-edge AI research and engineering in Bejing. You will collaborate with a world-class team to develop efficient algorithms that compress and accelerate large AI models for deployment across client, edge, cloud, and data center platforms.

The role emphasizes post-training quantization techniques, several AI inference frameworks, and strong software skills in PyTorch, C/C++, Python, and shell scripting.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Research Engineer/Scientist
AI Research Engineer/Scientist

intel • Prescott (AZ)

On-site
USD 52,000 - 82,000
AI Research Engineer, Inference
AI Research Engineer, Inference

OP Recruiting • Chicago (IL)

On-site
USD 150,000 - 210,000
Health benefits
Research Scientist: Efficient AI Inference
Research Scientist: Efficient AI Inference

Bitdeer (NASDAQ: BTDR) • Austin (TX)

On-site
USD 140,000 - 220,000
AI Software Development Engineer
AI Software Development Engineer

intel • Prescott (AZ)

On-site
USD 18,000 - 24,000
Efficient GenAI Research Engineer: Distillation
Efficient GenAI Research Engineer: Distillation

ByteDance • San Jose (CA)

On-site
USD 254,000 - 588,000
AI Systems Engineer: GenAI & Datacenter Innovation
AI Systems Engineer: GenAI & Datacenter Innovation

QUALCOMM, Inc. • San Diego (CA)

On-site
USD 100,000 - 149,000
Research Scientist, Artificial Intelligence
Research Scientist, Artificial Intelligence

Meta Careers • Menlo Park (CA)

On-site
USD 180,000 - 240,000
Senior ML Engineer: LLM Inference & Quantization
Senior ML Engineer: LLM Inference & Quantization

XPENG • Santa Clara (CA)

On-site
USD 175,000 - 296,000
Competitive compensation
Snacks and meals
Cutting-edge technologies
+1
AI Research Scientist - On-Device Large Language Models
AI Research Scientist - On-Device Large Language Models

Boulder Connect • San Diego (CA)

On-site
USD 130,000 - 180,000
Research Scientist: AI Infrastructure & Large-Scale ML
Research Scientist: AI Infrastructure & Large-Scale ML

Bytedance • San Jose (CA), Northern (KY)

Hybrid
USD 140,000 - 230,000