Member of Technical Staff, Small Language Models

Purchaseyourfreedom

Montreal (administrative region)

On-site

CAD 100,000 - 130,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

E+E Consulting is hiring a Member of Technical Staff for Small Language Models in Montreal. In this role, you will co-own the SLM training stack and push the frontiers of privacy and model performance.

You will utilize cutting-edge ML tools to ensure efficient model deployment while collaborating closely with other experts. Ideal candidates have substantive experience in small language models and are willing to relocate if necessary.

Qualifications

  • Substantive shipped work in small language models, efficient training, and distillation.
  • Track record of defending technical positions on open research questions.
  • Comfort with deployment challenges alongside modeling.

Responsibilities

  • Co-own the SLM training stack and Expert Fidelity evals.
  • Push for technical advancements in privacy and model performance.
  • Ship to production every sprint with real implications.

Skills

Deep ML technical chops
Expertise in small language models
Experience with on-device inference
Knowledge of federated learning

Education

PhD in relevant field or equivalent experience

Tools

PyTorch
HuggingFace tooling

Job description

Member of Technical Staff, Small Language Models

Full-time

Onix is the Personal Intelligence platform. Each onix is a small language model trained exclusively on a single expert’s private corpus: clinical notes, unpublished research, proprietary methods that have never been on the internet and never will be. Fully isolated, never bleeding across experts.

You will co‑own the SLM training stack and the Expert Fidelity evals that already beat frontier models on the dimensions our experts and users care about.

Privacy is an engineering surface, not a compliance line. Per‑expert isolation, on‑device inference paths, federated update strategies, and grounding guarantees are open research areas you will own.

Every session generates refinement signal. Every expert validates outputs in the loop. This is exclusive, expert‑graded data that Big AI cannot scrape, replicate, or buy.

Our SLMs run at orders of magnitude lower cost and faster latency than frontier models. We own the inference stack. We are not a wrapper. That economics is what makes a profitable consumer subscription business possible, and what makes the category structurally impossible for Big AI to follow into without cannibalizing their core.

20+ founding experts are live in App Store early access, including Dave Rabin, Mark Sisson, Ashley Koff, William Li, and Elissa Epel. NYT bestsellers and category‑defining voices. They pull in their peers unprompted. The data moat compounds with every conversation.

Small senior team in Old Montreal. In person. You report to our CTO and partner closely with engineering.

The next great AI lab will prove that human expertise is a moat, not a training set. That is what we are proving.

What You Will Do
  • Co‑own the SLM training stack: corpus, training, evals, deploy, monitor. Every layer.
  • Push per‑expert grounding, preference learning (RLHF / RLVR), and Fidelity research that makes our privacy‑as‑architecture story technically undeniable.
  • Co‑own Expert Fidelity evals. Make the bar harder. Make us pass it.
  • Ship to production every sprint. Real users. Real experts. Real consequences if we get it wrong.
  • Sit on calls with our experts directly when their voice, fidelity, or persona needs ML‑level attention.
What You Will Work With
  • Training stack: modern PyTorch and HuggingFace tooling for accelerated per‑expert fine‑tuning at scale. Distillation pipelines that generate training data without exposing a private corpus. Expert Fidelity evals as the gating bar.
  • Inference stack: autoscaled per‑expert endpoints. You own quantization, caching, and the per‑expert latency budget.
  • Preference learning at the per‑expert level: RLHF where each expert is the literal human in the loop, and RLVR grounded in our Expert Fidelity reward surface. The data and the experts are exclusive. The technique stack is yours to push.
  • On‑device and edge deployment for iOS. Latency and battery are first‑class constraints, not afterthoughts.
  • Per‑expert isolation infrastructure. No data crosses expert boundaries. Privacy is architecture, you own the surface.
  • Production scale on real users from day one. Real corpora. Real consequences if you get it wrong.
Who You Are

You can read a paper, prototype the model, and ship it to production in the same week. You have substantive work in small language models, efficient training, distillation, on‑device inference, federated learning, retrieval, or grounding. You came from Mila, Vector, Cohere, or a frontier lab (Anthropic, OpenAI, DeepMind, Hugging Face, Mistral). Or you are finishing a PhD and want your next system to ship to real users instead of a benchmark.

You came here for the mission. Technology should amplify human genius, not replace it.

We publish on our timeline, not a journal’s.

You have
  • Deep ML technical chops across training, eval, and deploy.
  • Substantive shipped work in SLMs, distillation, on‑device inference, federated learning, retrieval, or grounding.
  • A track record of defending technical positions on open research questions with evidence, not vibes.
  • Comfort owning infra alongside modeling. The deploy is your problem too.
You are
  • A research engineer first. Engineers here do research. Researchers here do engineering.
  • Opinionated. You can articulate a position on three open research questions in our space within an hour of joining.
  • Direct. You tell teammates they are wrong when they are, and accept the same in return.
  • In Old Montreal in person, or willing to relocate.

We do not care which lab you came from. We care what you have shipped, what you have measured, and what you would publish next.

What Success Looks Like
  • You own SLM and Expert Fidelity end to end. You decide what we train next, what we deprecate, what we publish.
  • Our published technical narrative is undeniable. Experts and serious researchers read it and say “they have a real moat.”
  • The ML team rallies to your technical direction.
  • Big AI watches what we publish and copies us.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Intern, Small Language Models (Mitacs)
Research Intern, Small Language Models (Mitacs)

Onix (doing business as “Onix”; legal entity listed as 16445039 Canada Inc.) • Montreal (administrative region)

Hybrid
CAD 17,000 - 23,000
Senior Product Manager
Senior Product Manager

EviSmart™ • Vancouver

On-site
CAD 120,000 - 180,000
AI Product Manager
AI Product Manager

INTO Inc. • Montreal (administrative region)

On-site
CAD 166,000 - 250,000
Remote work
Bonus program
Internal training
+1
Machine Learning Engineer
Machine Learning Engineer

Urban Ridge Supplies • Toronto

On-site
CAD 210,000 - 224,000
Competitive equity
Generous PTO
Parental leave
+4
AI Engineer
AI Engineer

Valsoft Corporation • Canada

On-site
CAD 100,000 - 140,000
Senior Research Scientist
Senior Research Scientist

adaption • Toronto

On-site
CAD 90,000 - 130,000
Flexible work
Annual travel stipend
Weekly lunch stipend
+1
Technical Product Manager
Technical Product Manager

AltaML • Edmonton

On-site
CAD 110,000 - 130,000
Uncapped Vacation
Make an Impact
Working with PhD and Master Level Coll
+3
Technical Product Manager
Technical Product Manager

AltaML Inc. • Edmonton

On-site
CAD 110,000 - 130,000
Uncapped Vacation
Make an Impact
Work with PhD and Master Level Colleag
+3
Intermediate Full Stack Software Engineer
Intermediate Full Stack Software Engineer

AltaML • Edmonton

On-site
CAD 90,000 - 130,000
Uncapped Vacation
Make an Impact
Office as a Resource
+1
Senior Member of Technical Staff, Safety and Security for Agents
Senior Member of Technical Staff, Safety and Security for Agents

Cohere • Montreal (administrative region)

On-site
CAD 120,000 - 180,000
Open and inclusive culture
Collaborative AI research environment
Weekly lunch stipend
+5