AI Framework Software Engineer – SGlang & Kernels

Intel Corporation

Prescott (AZ)

On-site

USD 45,000 - 89,000

Full time

11 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Intel Corporation is seeking a College Grad to design and develop AI software, optimize DL frameworks, and transform neural network graphs. You will work on distributed model efficiency and collaborate with researchers, applying your C++/Python skills and DL theory to real-world problems.

The role emphasizes on-site work in Shanghai with opportunities to contribute to public upstream and advanced AI systems, leveraging PyTorch and related tools.

Qualifications

  • Master's or Ph.D. in Computer Science, AI, Software Engineering, or related field.
  • Proficiency in C++ and Python.
  • Solid foundations in Deep Learning theory with hands-on experience.
  • Fluent English communication (written and spoken).
  • Experience in performance optimization or high-efficiency kernel development is a plus.
  • Experience with PyTorch, SGLang, and vLLM is a plus.
  • Experience with LLMs and understanding model structures is a plus.

Responsibilities

  • Designs and develops AI software and frameworks (e.g., SGLang).
  • Transforms neural network graphs and develops ML primitives.
  • Profiles distributed DL models to identify bottlenecks and proposes solutions.
  • Optimizes code for various hardware backends and collaborates with researchers.

Skills

C++
Python
DL theory
English proficiency
Problem solving
Performance optimization
LLM/AI frameworks
PyTorch/SGLang/vLLM
Model architectures

Education

Master or PhD in CS/AI/Software Eng

Job description

Job Details:
Job Description:

Conducts design and development to build and optimize AI software. Designs, develops, and optimizes for AI frameworks (e.g., SGLang) and contribute to public upstream. Implements various distributed algorithms such as model/data parallel frameworks, parameter servers, dataflow based asynchronous data communication in machine learning, and/or deep learning frameworks. Transforms computational graph representation of neural network model, and develops machine learning and/or deep learning primitives in mathematical libraries. Profiles distributed deep learning models to identify performance bottlenecks and proposes solutions across individual component teams. Optimizes code for various computing hardware backends, and interacts with machine learning and/or deep learning researchers, and utilizing experience with machine learning and/or deep learning frameworks.

Qualifications:
  • Master degree or Ph.D. in in Computer Science, Artificial Intelligence, Software Engineering, or related fields.
  • Good programming skills of modern C++ and Python
  • Know foundations of Deep Learning theory and some hands-on experience
  • Communicate in English fluently (both written and spoken)
  • Are passionate on solving problems and positive thinker
  • Experience of performance optimization or high efficiency kernel development experience is a plus
  • Experience of PyTorch, SGLang, vLLM is a plus
  • Experience of LLM and deep understanding of model structure is a plus
Job Type:

College Grad

Shift:

Shift 1 (China)

Primary Location:

PRC, Shanghai

Posting Statement:

All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance.

Position of Trust

N/A

Work Model for this Role

This role will require an on-site presence. Job posting details (such as work model, location or time type) are subject to change.

ADDITIONAL INFORMATION:

Intel is committed to Responsible Business Alliance (RBA) compliance and ethical hiring practices. We do not charge any fees during our hiring process. Candidates should never be required to pay recruitment fees, medical examination fees, or any other charges as a condition of employment. If you are asked to pay any fees during our hiring process, please report this immediately to your recruiter.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Framework Software Engineer - vLLM
AI Framework Software Engineer - vLLM

Intel Corporation • Prescott (AZ)

On-site
USD 45,000 - 81,000
Software Engineer — Distributed LLM Inference Systems
Software Engineer — Distributed LLM Inference Systems

Intel Corporation • Prescott (AZ)

On-site
USD 52,000 - 82,000
AI Software Solutions Engineer
AI Software Solutions Engineer

Intel Corporation • Prescott (AZ)

On-site
USD 39,000 - 63,000
AI Frameworks Engineer: SGLang & Kernels
AI Frameworks Engineer: SGLang & Kernels

Intel Corporation • Prescott (AZ)

On-site
USD 45,000 - 89,000
AI Research Engineer/Scientist
AI Research Engineer/Scientist

Intel Corporation • Prescott (AZ)

On-site
USD 52,000 - 82,000
Agentic AI Customer Co-Val CG
Agentic AI Customer Co-Val CG

Intel Corporation • Prescott (AZ)

On-site
USD 27,000 - 52,000
AI Systems Performance and Simulation Engineer
AI Systems Performance and Simulation Engineer

Intel Corporation • Prescott (AZ)

On-site
USD 65,000 - 90,000
AI Frameworks Engineer
AI Frameworks Engineer

Intel Corporation • Santa Clara (CA)

Hybrid
USD 150,000 - 276,000
Stock bonuses
Health benefits
Retirement plan
+1
Ecosystem Technical Enabling Engineer
Ecosystem Technical Enabling Engineer

Intel Corporation • Prescott (AZ)

On-site
USD 120,000 - 160,000
Platform Power Thermal Performance Engineer
Platform Power Thermal Performance Engineer

Intel Corporation • Prescott (AZ)

On-site
USD 27,000 - 36,000