Distributed LLM Inference Engineer

Intel Corporation

Prescott (AZ)

On-site

USD 52,000 - 82,000

Full time

11 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Intel Corporation in Shanghai is hiring a College Grad Software Engineer to contribute to distributed LLM inference systems within the AI Frameworks team. You will design, implement, and optimize components for high-performance AI software on diverse hardware, collaborating with researchers and engineers to push AI capabilities forward.

The role emphasizes distributed algorithms, model/data parallelism, and efficient communication, with a strong foundation in Python and C++ and knowledge of

Qualifications

  • Master’s degree or higher in Computer Science, AI, or related field.
  • Proficiency in Python and modern C++ programming.
  • Foundational knowledge of deep learning frameworks (e.g., PyTorch).
  • Experience debugging and optimizing software for performance.
  • Strong problem-solving and quick learning of unfamiliar systems.

Responsibilities

  • Design, develop, and optimize distributed LLM inference systems and AI framework components.
  • Implement distributed algorithms and asynchronous communication for deep learning.
  • Develop components like schedulers, workers, and KV cache management.
  • Profile workloads to identify bottlenecks in compute, communication, memory, and scheduling.
  • Collaborate with teams to improve latency, throughput, and scalability.

Skills

Python
C++
PyTorch
Distributed systems
Performance optimization

Education

Master’s degree in CS/AI/Software Engineering

Tools

Git

Job description

Intel Corporation in Shanghai is hiring a College Grad Software Engineer to contribute to distributed LLM inference systems within the AI Frameworks team. You will design, implement, and optimize components for high-performance AI software on diverse hardware, collaborating with researchers and engineers to push AI capabilities forward.

The role emphasizes distributed algorithms, model/data parallelism, and efficient communication, with a strong foundation in Python and C++ and knowledge of

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer — Distributed LLM Inference Systems
Software Engineer — Distributed LLM Inference Systems

Intel Corporation • Prescott (AZ)

On-site
USD 52,000 - 82,000
Distributed Systems Engineer for LLM Inference Platform
Distributed Systems Engineer for LLM Inference Platform

Baseten • San Francisco (CA)

On-site
USD 180,000 - 360,000
Competitive equity
Medical insurance
Flexible PTO
+4
AI Frameworks Engineer: SGLang & Kernels
AI Frameworks Engineer: SGLang & Kernels

Intel Corporation • Prescott (AZ)

On-site
USD 45,000 - 89,000
Senior LLM Inference Architect — Heterogeneous Hardware
Senior LLM Inference Architect — Heterogeneous Hardware

d-Matrix inc. • Santa Clara (CA)

On-site
USD 130,000 - 170,000
Competitive compensation
Equity
Inclusive work environment
Senior LLM Inference Systems Engineer
Senior LLM Inference Systems Engineer

Snowflake • Menlo Park (CA)

On-site
USD 236,000 - 310,000
Bonus and equity plan
Medical, dental, vision insurance
401(k) retirement plan
+1
Senior LLM Inference Systems Engineer
Senior LLM Inference Systems Engineer

Snowflake • Menlo Park (CA)

On-site
USD 236,000 - 310,000
Medical insurance
Bonus & equity plan
401(k) retirement plan
+1
Principal LLM Inference Engineer — Agentic AI Systems
Principal LLM Inference Engineer — Agentic AI Systems

NVIDIA • United States

Remote
USD 272,000 - 431,000
Equity
Benefits
AI Framework Software Engineer - vLLM
AI Framework Software Engineer - vLLM

Intel Corporation • Prescott (AZ)

On-site
USD 45,000 - 81,000
Senior Software Engineer, AI Infrastructure for LLMs
Senior Software Engineer, AI Infrastructure for LLMs

Jobgether SRL • United States

Remote
USD 140,000 - 220,000
Senior AI Systems Research Engineer - LLM Inference
Senior AI Systems Research Engineer - LLM Inference

Snowflake • Bellevue (WA)

On-site
USD 236,000 - 310,000
Bonus and equity plan