Engineer - ML & RL

Huawei Technologies Canada Co., Ltd.

Edmonton

On-site

CAD 80,000 - 110,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading technology company in Canada is seeking an Engineer for a 12-month contract. The role involves designing and building scalable infrastructure for machine learning and reinforcement learning systems, collaborating with research teams, and developing efficient ML solutions. The ideal candidate has a Master's or PhD in Computer Science or related fields, strong Python programming skills, and proven research excellence. This position offers the opportunity to work with cutting-edge technologies and a commitment to an inclusive recruitment process.

Qualifications

  • Excellent programming skills with strong software engineering practices.
  • Demonstrated experience implementing RL algorithms beyond academic prototypes.
  • Proven research excellence, including at least one publication in top-tier venues.

Responsibilities

  • Design and build scalable infrastructure for ML systems.
  • Develop efficient ML solutions for RL and Recommendation Systems.
  • Conduct systematic benchmarking and validate in real-world environments.
  • Collaborate with research teams to improve training capabilities.

Skills

Python programming
Reinforcement Learning
Deep Learning
Recommender Systems
Transformer-based architectures

Education

Master's or PhD in Computer Science, Machine Learning, or a related field

Tools

PyTorch
DeepSpeed

Job description

Huawei Canada has an immediate 12-month contract opening for an Engineer.

About the team

The Software-Hardware System Optimization Lab continuously improves the power efficiency and performance of smartphone products through software-hardware systems optimization and architecture innovation. We keep tracking the trends of cutting-edge technologies, building the competitive strength of mobile AI, graphics, multimedia, and software architecture for mobile phone products.

About the job
  • Design and build scalable infrastructure to support Reinforcement Learning, Online Search, Recommendation Systems, large model fine-tuning and evaluation/deployment.
  • Develop efficient ML solutions for Recommendation Systems and RL problems, including Multi‑Armed and Contextual Bandit, Tree Search, and Multi‑Agent system orchestration.
  • Implement and optimize deep learning architectures, including custom Transformers for agentic and decision‑making systems.
  • Apply search and optimization techniques to efficiently fine‑tune RL and ML models.
  • Work with large multimodal models (LLMs, VLMs), analyze their components, and fine‑tune them for task‑specific applications.
  • Conduct systematic benchmarking, new papers reading, experimentation, and validation in both simulation and real‑world product environments.
  • Collaborate closely with research teams to scale online RL training capabilities and improve system robustness and accuracy.
  • Explore and integrate emerging AI methodologies and tools into production platforms.
About the ideal candidate
  • Master’s or PhD in Computer Science, Machine Learning, or a related field.
  • Excellent Python programming skills with strong software engineering practices.
  • Strong foundation in Reinforcement Learning, Deep Learning, Recommender Systems, and Transformer‑based architectures.
  • Demonstrated experience implementing RL algorithms beyond academic prototypes.
  • Hands‑on experience with PyTorch and distributed training frameworks such as DeepSpeed.
  • Proven research excellence, including at least one publication in top‑tier venues (e.g., NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV, ICRA, RLC).
  • Familiarity with LLM post‑training techniques such as RLHF, PPO/GRPO, SFT, LoRA, or MoE is an asset.
  • Experience with multi‑agent RL systems or tool‑use agents is an asset.
Additional Information

Huawei Canada is committed to a fair, inclusive, and accessible recruitment process. If you require accommodation during any stage of the hiring process, please let us know and we will work with you to meet your needs.

All applications for this position are reviewed directly by our hiring team; we do not use artificial intelligence tools to screen or select candidates.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Agentic RL Researcher – Distributed Computing
Agentic RL Researcher – Distributed Computing

Huawei Canada • Markham

On-site
CAD 106,000 - 156,000
Researcher - Reinforcement Learning
Researcher - Reinforcement Learning

Huawei Technologies Canada Co., Ltd. • Edmonton

On-site
CAD 80,000 - 120,000
Intern Engineer – RL Post-Training for LLMs
Intern Engineer – RL Post-Training for LLMs

Huawei Canada • Vancouver

On-site
CAD 58,000 - 104,000
Intern Researcher - LLMs Agentic AI and RL
Intern Researcher - LLMs Agentic AI and RL

Huawei Technologies Canada Co., Ltd. • Markham

On-site
CAD 58,000 - 104,000
Researcher – AI/ML
Researcher – AI/ML

Huawei Canada • Markham

On-site
CAD 106,000 - 156,000
Machine Learning Researcher -LLM Agents & Efficient Deep Learning
Machine Learning Researcher -LLM Agents & Efficient Deep Learning

Huawei Canada • Montreal (administrative region)

On-site
CAD 106,000 - 156,000
Intern Researcher - LLMs Agentic AI and RL
Intern Researcher - LLMs Agentic AI and RL

Huawei Canada • Markham

On-site
CAD 79,000 - 143,000
Research Engineer - AI Workload & Systems
Research Engineer - AI Workload & Systems

Huawei Technologies Canada Co., Ltd. • Markham

On-site
CAD 178,000 - 316,000
Intern Research Engineer - AI Agent Systems
Intern Research Engineer - AI Agent Systems

Huawei Canada • Markham

On-site
CAD 58,000 - 104,000
Researcher - AI Computing System
Researcher - AI Computing System

Huawei Canada • Vancouver

On-site
CAD 106,000 - 205,000