Researcher - AI Computing System

Huawei Canada

Vancouver

On-site

CAD 106,238 - 204,304

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Huawei Canada is hiring a Researcher for a 12-month contract to enhance AI systems on the Ascend platform. This role focuses on developing optimization solutions for AI training and inference systems, ensuring scalability and reliability of AI training clusters. Ideal candidates hold a Ph.D or Masters in relevant fields, have a solid programming foundation in Python/C/C++, and must be familiar with AI model structures. The position offers competitive compensation ranging from $78,000 to $150,000 based on experience.

Qualifications

  • Ph.D or Masters in Computer Science, AI, or related fields.
  • Familiarity with model structures for large models.
  • Knowledge of hardware architecture of AI accelerators.

Responsibilities

  • Focus on enhancing AI systems performance on Ascend platform.
  • Design optimization solutions for AI training and inference systems.
  • Build stable and efficient AI training clusters.

Skills

Programming in Python/C/C++
Problem-solving
Communication
Hands-on practice

Education

Ph.D or Master's degree in relevant fields

Tools

AI training frameworks
AI reasoning engines

Job description

Huawei Canada has an immediate 12 month contract opening for a Researcher

About the team:

The Advanced Computing and Storage Lab, currently a part of the Vancouver Research Centre, aims to explore adaptive computing system architectures to address the challenges posed by flexible and variable application loads in the future. It assists in ensuring the stability and quality of training clusters, constructs dynamic cluster configuration strategy solvers, and establishes precision control systems to create stable and efficient computing power clusters. One of the lab's goals is to focus on key industry AI application scenarios such as large model training/inference, based on key technologies like low-precision training, multi-modal training, and reinforcement learning, responsible for bottleneck analysis and the design and development of optimization solutions, thereby improving training and inference performance as well as usability.

About the job:

  • Aiming at key industry AI application scenarios such as large model training and inference, this role focuses on advancing performance, efficiency, and usability of AI systems on the Ascend platform. The work involves low-precision training, multimodal optimization, reinforcement learning, and training resource optimization to address system bottlenecks and deliver next-generation AI capabilities.
  • Responsible for design and development of optimization solutions for AI training and inference systems, with a focus on FP8 optimization, RL-driven training agents, multimodal reinforcement learning or next-generation multi-modal understanding & generation.
  • Combine AI algorithm requirements with system-level architectural optimization in computing, I/O, scheduling, and precision control to improve performance.
  • Build stable, efficient AI training clusters, leveraging dynamic cluster configuration and precision control to ensure scalability and reliability.
  • Develop software frameworks, operator libraries, acceleration libraries, and system-level optimizations for NPU platforms to accelerate large-model AI training.
  • Drive innovation in optimizing large-model training and inference with low-precision training, parallel strategy tuning, and reinforcement learning.
  • Grasp the latest research progress and technological trends in AI computing cluster architecture design, training acceleration, and inference acceleration across academia and industry to strengthen the competitiveness of AI computing cluster systems.

The target annual compensation (based on 2080 hours per year) ranges from $78,000 to $150,000 depending on education, experience and demonstrated expertise.

Job requirements

About the ideal candidate:

  • Ph.D or Masters degree in Computer Science, Computer Engineering majors in artificial intelligence, computer science, software, automation, electronics, communications, robotics, etc.
  • Familiar with the common model structures of large models such as Deepseek and Llama, and have basic technical accumulation in large model training and inference optimization in the fields of LLM, MoE, multimodality, etc.
  • Familiar with the hardware architecture and programming system of AI accelerators such as GPU/NPU, and have experience in optimizing AI systems with coordinated software and hardware cores.
  • Those with any of the following experience is an asset:
  • Solid programming foundation, familiar with Python/C/C++ programming languages, good architecture design and programming habits.
  • Ability to work independently and solve problems, good at communication, willing to cooperate, keen on new technologies, good at summarizing and sharing, and like hands-on practice.
  • Experience in the development of AI training frameworks and AI reasoning engines, or algorithm hardware and related experience.
  • Strong research capabilities in new technologies and new architectures, can quickly track and gain insights into the most cutting-edge AI technologies in the industry, and lead the continuous leadership of system architecture innovation.
Additional Information

Huawei Canada is committed to a fair, inclusive, and accessible recruitment process. If you require accommodation during any stage of the hiring process, please let us know and we will work with you to meet your needs.

All applications for this position are reviewed directly by our hiring team; we do not use artificial intelligence tools to screen or select candidates.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Intern Researcher - AI Computing System
Intern Researcher - AI Computing System

Huawei Canada • Vancouver

On-site
CAD 78,000 - 150,000
Research Engineer - AI Workload & Systems
Research Engineer - AI Workload & Systems

Huawei Technologies Canada Co., Ltd. • Markham

On-site
CAD 178,000 - 316,000
Senior Researcher – Hardware Efficient AI Foundation Model Training
Senior Researcher – Hardware Efficient AI Foundation Model Training

Huawei Canada • Markham

On-site
CAD 127,000 - 225,000
Co-op Researcher - AI Computing System
Co-op Researcher - AI Computing System

Huawei Canada • Vancouver

On-site
CAD 76,000 - 109,000
Researcher – AI/ML
Researcher – AI/ML

Huawei Canada • Markham

On-site
CAD 106,000 - 156,000
Senior Researcher – AI Data Systems
Senior Researcher – AI Data Systems

Huawei Canada • Markham

On-site
CAD 127,000 - 225,000
Intern Researcher - AI/ML Technology Insight and Planning
Intern Researcher - AI/ML Technology Insight and Planning

Huawei Canada • Vancouver

On-site
CAD 58,000 - 104,000
Senior Researcher - AI Data Platforms
Senior Researcher - AI Data Platforms

Huawei Canada • Markham

On-site
CAD 127,000 - 225,000
Researcher - Technical Insight and Planning (AI/ML)
Researcher - Technical Insight and Planning (AI/ML)

Huawei Canada • Markham

On-site
CAD 127,000 - 225,000
Senior Researcher - AI Data Platforms
Senior Researcher - AI Data Platforms

Huawei Technologies Canada Co., Ltd. • Markham

On-site
CAD 127,000 - 225,000