Lead Research Engineer - Post-Training ML Systems (Remote)

Giotto.ai

Lausanne

Vor Ort

CHF 180.000 - 240.000

Vollzeit

Vor 4 Tagen
Sei unter den ersten Bewerbenden

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Zusammenfassung

Giotto.ai, a Switzerland-based AI company, seeks a Lead Research Engineer/Scientist to own and lead the training and optimisation of post-training stacks. From pretrained checkpoints, you will design, scale, and operate methods to produce capable, reliable production models, including SFT, RLHF, and policy distillation, across multi-machine environments.

You will shape technical direction, mentor researchers, and collaborate with data, infra, and evaluation teams while staying hands-on in the

Qualifikationen

  • Experience in large-scale model training/post-training on multi-machine systems.
  • Strong Python and PyTorch fluency; distributed execution expertise.
  • Ability to design end-to-end post-training pipelines and curricula.
  • Experience with RLHF, reinforcement learning for language models, and reward modeling.
  • Proven ability to guide researchers and engineers and make design decisions.

Aufgaben

  • Own end-to-end post-training pipeline from pretrained checkpoints to production candidate.
  • Set technical direction and priorities for post-training and RL work.
  • Design and execute full-parameter and parameter-efficient SFT.
  • Implement preference optimisation, RLHF, RLAIF and related methods.
  • Develop training strategies for reasoning, coding, tool use, multilingual behaviour, and long-horizon agent tasks.
  • Collaborate with data, evaluation, infra, and inference teams; provide mentorship.

Kenntnisse

Python
PyTorch
Distributed training
Reinforcement learning
Research leadership
RLHF
SFT
Policy design

Tools

PyTorch Distributed
FSDP
DeepSpeed
Megatron-Core
CUDA

Jobbeschreibung

Giotto.ai, a Switzerland-based AI company, seeks a Lead Research Engineer/Scientist to own and lead the training and optimisation of post-training stacks. From pretrained checkpoints, you will design, scale, and operate methods to produce capable, reliable production models, including SFT, RLHF, and policy distillation, across multi-machine environments.

You will shape technical direction, mentor researchers, and collaborate with data, infra, and evaluation teams while staying hands-on in the

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Post-Training AI Engineer for Scalable Systems
Senior Post-Training AI Engineer for Scalable Systems

Embodied AI • Lausanne

Hybrid
CHF 150.000 - 210.000
ML Engineer - Scalable AI & LLM Systems (Remote)
ML Engineer - Scalable AI & LLM Systems (Remote)

Embodied AI • Lausanne

Hybrid
CHF 120.000 - 160.000
Senior Research Engineer / Research Scientist - Post-Training, Reinforcement Learning & Training Systems
Senior Research Engineer / Research Scientist - Post-Training, Reinforcement Learning & Training Systems

Embodied AI • Lausanne

Hybrid
CHF 150.000 - 210.000
Remote ML Engineer — LLMs, PyTorch & Distributed Systems
Remote ML Engineer — LLMs, PyTorch & Distributed Systems

Giotto.ai • Lausanne

Hybrid
CHF 120.000 - 170.000
Remote AI Research Scientist - LLMs & Multimodal Reasoning
Remote AI Research Scientist - LLMs & Multimodal Reasoning

Giotto.ai • Lausanne

Hybrid
CHF 140.000 - 210.000
Machine Learning Engineer
Machine Learning Engineer

Embodied AI • Lausanne

Hybrid
CHF 120.000 - 160.000
Machine Learning Engineer
Machine Learning Engineer

Giotto.ai • Lausanne

Hybrid
CHF 120.000 - 170.000
AI Research Scientist
AI Research Scientist

Giotto.ai • Lausanne

Hybrid
CHF 140.000 - 210.000
AI Research Scientist
AI Research Scientist

Embodied AI • Lausanne

Hybrid
CHF 120.000 - 170.000
Remote Senior AI Engineer: Build Production LLMs & AI Apps
Remote Senior AI Engineer: Build Production LLMs & AI Apps

Jobgether • Schweiz

Hybrid
CHF 140.000 - 210.000
Remote-friendly
Home-office budget
Learning budget
+2