Senior Post-Training AI Scientist (RL & Scaling)

Giotto.ai

Lausanne

Hybrid

CHF 150,000 - 210,000

Full time

8 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Giotto.ai in Switzerland seeks a Lead Research Scientist to own and lead the post-training stack, including supervised fine-tuning, RLHF, and verification components. You will design scalable training systems, manage memory and parallelism, and mentor researchers while remaining hands-on.

The role emphasizes end-to-end pipelines, multi-machine training, and collaboration with data, evaluation, and infrastructure teams to deliver reliable, controllable production models.

Qualifications

  • Experience leading large-scale language-model training or post-training runs across multiple machines.
  • Proven ability to design and run end-to-end post-training pipelines in production.
  • Strong Python and PyTorch proficiency with distributed execution expertise.
  • Hands-on experience with DeepSpeed, Megatron-Core, or equivalent frameworks.

Responsibilities

  • Own the end-to-end post-training pipeline from pretrained checkpoint to production candidate.
  • Set technical direction for post-training and reinforcement-learning work.
  • Design and execute full-parameter and parameter-efficient SFT.
  • Implement preference optimisation, RLHF, and related methods.
  • Develop training strategies for reasoning, tool use, multilingual behaviour, and long-horizon tasks.
  • Integrate reward models and verifiers; build scalable rollout-generation systems.
  • Scale training across multiple machines and accelerators with suitable parallelism.

Skills

LLM training
Python
PyTorch
Distributed training
DeepSpeed
Transformers
CUDA/NCCL
Reinforcement learning
Experiment design
Parallelism strategies

Tools

DeepSpeed
Megatron-Core
Hugging Face Transformers
CUDA
NCCL
Docker
GCP
Weights & Biases

Job description

Giotto.ai in Switzerland seeks a Lead Research Scientist to own and lead the post-training stack, including supervised fine-tuning, RLHF, and verification components. You will design scalable training systems, manage memory and parallelism, and mentor researchers while remaining hands-on.

The role emphasizes end-to-end pipelines, multi-machine training, and collaboration with data, evaluation, and infrastructure teams to deliver reliable, controllable production models.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Research Engineer — Post-Training & Scale
Senior AI Research Engineer — Post-Training & Scale

Giotto.ai • Lausanne

Remote
CHF 180,000 - 240,000
Lead Research Engineer - Post-Training ML Systems (Remote)
Lead Research Engineer - Post-Training ML Systems (Remote)

Giotto.ai • Lausanne

On-site
CHF 180,000 - 240,000
Senior Research Engineer / Research Scientist - Post-Training, Reinforcement Learning & Trainin[...]
Senior Research Engineer / Research Scientist - Post-Training, Reinforcement Learning & Trainin[...]

Giotto.ai • Lausanne

Remote
CHF 180,000 - 240,000
Lead Research Scientist - Post-Training
Lead Research Scientist - Post-Training

Giotto.ai • Lausanne

Hybrid
CHF 150,000 - 210,000
Junior AI Researcher
Junior AI Researcher

Startupvalleys • Lausanne

Hybrid
CHF 65,000 - 110,000
Junior ML Engineer: Production AI, Hybrid in Switzerland
Junior ML Engineer: Production AI, Hybrid in Switzerland

Startupvalleys • Lausanne

Hybrid
CHF 90,000 - 120,000
Hybrid work model
LLM Research Scientist — Hybrid (Remote)
LLM Research Scientist — Hybrid (Remote)

Giotto.ai • Lausanne

Hybrid
CHF 120,000 - 190,000
LLM Research Scientist
LLM Research Scientist

Giotto.ai • Lausanne

Hybrid
CHF 120,000 - 190,000
Director, Model R&D — Hybrid AI Leadership in Regulated Work
Director, Model R&D — Hybrid AI Leadership in Regulated Work

Thomson Reuters Foundation • Zug

Hybrid
CHF 180,000 - 240,000
Hybrid Work Model
Career Development
Mental Health Days
+2
Production AI Post-Training Engineer
Production AI Post-Training Engineer

Anthropic • Zürich

Hybrid
CHF 80,000 - 120,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours