Remote Intern Researcher - Reinforcement Learning and LLM

Bilinguallink

Caledon

Remote

CAD 93,000 - 104,000

Part time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Bilinguallink in Ontario, Canada, is seeking an InternResearcher for a 12-month internship to conduct cutting-edge research in large language models and agent-building technologies.

You will leverage reinforcement learning to boost agent performance in business applications, design and maintain research prototypes with thorough docs, publish papers at top conferences, and collaborate with a multidisciplinary team.

Qualifications

  • PhD in CS/AI/ML/Math required or closely related field.
  • Proficient in Python with significant model development experience.
  • Strong analytical, problem-solving and troubleshooting abilities.
  • Excellent written and verbal communication; team collaboration.

Responsibilities

  • Conduct cutting-edge research in large language models and agent-building technologies.
  • Leverage reinforcement learning to enhance agent performance in specialized business applications.
  • Design, develop and maintain research prototypes with thorough documentation and best practices.
  • Contribute to publishing full research papers in top-tier conferences and journals.

Skills

Python
Model development
Optimization
Research skills
Communication

Education

PhD in Computer Science, Artificial Intelligence, Machine Learning, Mathematics, or a closely related technical field

Job description

Our team has an immediate 12-month internship opening for an InternResearcher.
Responsibilities:
  • Conduct cutting-edge research in large language models (LLMs), focusing on agent-building technologies.
  • Leverage reinforcement learning techniques to enhance agent performance in specialized business applications.
  • Design, develop and maintain research prototypes, ensuring comprehensive documentation and adherence to software development best practices.
  • Contribute to publishing full research papers in top-tier conferences and journals.

The target annual compensation (based on 2080 hours per year) ranges from $93,000 to $104,000 depending on education, experience and demonstrated expertise

What you’ll bring to the team:
  • PhD in Computer Science, Artificial Intelligence, Machine Learning, Mathematics, or a closely related technical field.
  • Proficient in Python programming with significant experience in model development and optimization.
  • Excellent analytical, problem-solving, and troubleshooting abilities with a focus on delivering innovative research solutions.
  • Strong written and verbal communication skills, with the ability to collaborate effectively within a team.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Intern Researcher - Embodied AI
Intern Researcher - Embodied AI

Huawei Canada • Markham

On-site
CAD 70,000 - 90,000
Intern Researcher - LLMs Agentic AI and RL
Intern Researcher - LLMs Agentic AI and RL

Huawei Technologies Canada Co., Ltd. • Markham

On-site
CAD 58,000 - 104,000
Intern Researcher - LLMs Agentic AI and RL
Intern Researcher - LLMs Agentic AI and RL

Huawei Canada • Markham

On-site
CAD 79,451 - 142,465
Research Internship: AI/NLP Frontier (Remote, 4–6 mo)
Research Internship: AI/NLP Frontier (Remote, 4–6 mo)

Cohere • Toronto, Montreal (administrative region)

Hybrid
CAD 33,000 - 45,000
Lunch stipend
Health benefits
Retirement plan
+6
Intern Engineer – RL Post-Training for LLMs
Intern Engineer – RL Post-Training for LLMs

Huawei Canada • Vancouver

On-site
CAD 58,000 - 104,000
Intern Researcher - AI Agent Evaluation
Intern Researcher - AI Agent Evaluation

Huawei Canada • Markham

On-site
CAD 58,000 - 104,000
Intern Researcher – AI Foundation Model Training
Intern Researcher – AI Foundation Model Training

Huawei Canada • Markham

On-site
CAD 58,000 - 104,000
Research Internship (Winter 2027)
Research Internship (Winter 2027)

cohere • Toronto

Hybrid
CAD 32,000 - 42,000
Weekly lunch stipend
Health and dental benefits
RRSP matching / Pension
+4
Machine Learning Researcher -LLM Agents & Efficient Deep Learning
Machine Learning Researcher -LLM Agents & Efficient Deep Learning

Huawei Canada • Montreal (administrative region)

On-site
CAD 106,000 - 156,000
Intern Researcher - Embodied AI
Intern Researcher - Embodied AI

Huawei Canada • Markham

On-site
CAD 58,000 - 104,000