Advanced Engineer (High-Efficiency AI Computing)

Beijing Foreign Enterprise Management Consultants Co.,Ltd.

Singapore

On-site

SGD 180,000 - 260,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Huawei seeks an Advanced Engineer (High-Efficiency AI Computing) to shape the next generation of AI hardware. You will bridge algorithmic innovation and silicon design, advancing low-precision and sparse computing on in-house AI accelerators and driving energy efficiency and throughput for large-scale AI models.

Responsibilities include leading quantization research, building scalable kernels (GEMM, FlashAttention), and partnering with IC design to deliver technical specs, benchmarks, and patent

Qualifications

  • Ph.D. in Computer Science, Electronic Engineering, Automation, or a highly related technical field.
  • Deep foundational knowledge in computer architecture, specifically GPU/NPU microarchitecture implementation.
  • Proficiency in algorithm research focusing on low-precision quantization and sparsity.
  • Strong track record in inference optimization for LLMs or multi-modal models with kernel development.
  • Concrete system-level architectural design experience for high-performance AI inference accelerators.
  • Robust publication record in tier-1 architecture and AI conferences (ISCA, MICRO, HPCA, ASPLOS, NeurIPS, CVPR).
  • Exceptional cross-functional collaboration and communication across algorithm, silicon design, software, testing teams.

Responsibilities

  • Spearhead research into optimal quantization methodologies for low-precision data formats.
  • Architect and deploy high-performance kernels for low-precision and sparse computing.
  • Drive microarchitectural evolution of AI chips to improve energy efficiency and throughput.
  • Collaborate with IC design teams to deliver specifications, benchmarks, and patent proposals.

Skills

GPU/NPU microarchitecture
Computer architecture
Low-precision quantization
Sparsity
Inference optimization for LLMs
High-performance kernel development
Cross-functional collaboration

Education

PhD in Computer Science or Electrical/Electronic Engineering

Tools

GEMM kernels
FlashAttention
CUDA
Microarchitecture design tools

Job description

On behalf of Huawei, a world-renowned information and communication technology company, we are seeking passionate and talented individuals to join our team as Advanced Engineer (High-Efficiency AI Computing).

As an Advanced Engineer for High-Efficiency AI Computing, you will be at the forefront of shaping the next generation of AI hardware. Bridging the gap between algorithmic innovation and silicon design, you will play a critical role in pushing the boundaries of low-precision and sparse computing. Your work will directly influence the microarchitectural evolution of our flagship AI accelerators, driving massive improvements in energy efficiency and computational performance for large-scale AI models.

Job Description:

  • Algorithmic Innovation: Spearhead research into optimal quantization methodologies for advanced low-precision data formats, driving the strategic adoption and expansion of our low-precision computing ecosystem to secure a competitive edge in numerical computing.
  • Kernel Architecture: Architect and deploy high-performance, highly scalable kernels for low-precision and sparse computing (e.g., GEMM, FlashAttention), explicitly tailored to maximize the microarchitectural strengths of our in-house AI accelerators.
  • Hardware-Software Co-Design: Identify systemic performance bottlenecks and drive the microarchitectural evolution of next-generation AI chips, significantly enhancing energy efficiency and throughput.
  • Partner closely with IC design teams to ensure to deliver comprehensive technical specifications, rigorous benchmark reports, and high-value patent proposals.

Skills / Qualifications:

  • Education: Ph.D. in Computer Science, Electronic Engineering, Automation, or a highly related technical field.
  • Technical Expertise: Deep foundational knowledge in computer architecture, specifically regarding GPU/NPU microarchitecture implementation.
  • Domain Knowledge: Profound, demonstrable expertise in algorithm research focused on low-precision quantization and sparsity.
  • Proven Impact: A strong track record in inference optimization for Large Language Models (LLMs) or multi-modal models, coupled with hands-on high-performance kernel development.
  • System-Level Design: Concrete experience in system-level architectural design for high-performance AI inference accelerators.
  • Academic Excellence: A robust publication record in tier-1 architecture and AI conferences (e.g., ISCA, MICRO, HPCA, ASPLOS, NeurIPS, CVPR).
  • Professional Attributes: Exceptional cross-functional collaboration and communication skills, with the ability to build consensus across algorithm, silicon design, software, and testing teams.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Hardware Architect - Low-Precision & Sparse Compute
AI Hardware Architect - Low-Precision & Sparse Compute

Beijing Foreign Enterprise Management Consultants Co.,Ltd. • Singapore

On-site
SGD 180,000 - 260,000
AI Computing Architecture Researcher
AI Computing Architecture Researcher

Beijing Foreign Enterprise Management Consultants Co.,Ltd. • Singapore

On-site
SGD 120,000 - 180,000
LPU Chip Architecture Engineer
LPU Chip Architecture Engineer

CANAAN CREATIVE GLOBAL PTE. LTD. • Singapore

On-site
SGD 180,000 - 240,000
AI Engineer (ML Systems & Infrastructure)
AI Engineer (ML Systems & Infrastructure)

SwapeTech • Singapore

On-site
SGD 180,000 - 260,000
Expert R&D Software Engineer – AI / ML
Expert R&D Software Engineer – AI / ML

Keysight Technologies Singapore (Sales) Pte. Ltd....- • Singapore

On-site
SGD 180,000 - 240,000
Research Engineer (AI Accelerator and Energy-Efficient Computing)
Research Engineer (AI Accelerator and Energy-Efficient Computing)

National University of Singapore • Singapore

On-site
SGD 60,000 - 90,000
AI Engineer
AI Engineer

REALTEK SINGAPORE PRIVATE LIMITED • Singapore

On-site
SGD 120,000 - 210,000
AI System Design Engineer (Internship)
AI System Design Engineer (Internship)

OMNIVISION • Singapore

On-site
Hardware Engineer
Hardware Engineer

RUNSUN SERVICE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Expert R&D Software Engineer – AI / ML
Expert R&D Software Engineer – AI / ML

Keysight Technologies • Singapore

On-site
SGD 180,000 - 320,000