Apertus Engineer: Post-training

Eidgenössische Technische Hochschule Zürich

Zürich

Hybrid

CHF 120.000 - 170.000

Vollzeit

vor 22 Stunden
Sei unter den ersten Bewerbenden

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Benefits dieser Stelle

Remote work options
Access to HPC infrastructure

Zusammenfassung

ETH Zürich is seeking a skilled engineer to join the Apertus post-training effort. You will develop, run, and evaluate SFT and reinforcement learning pipelines to turn Apertus base models into capable assistants, in a research-focused HPC setting.

You will work with researchers and engineers across ETH Zürich, EPFL, and CSCS, building containerised environments and running Slurm-based jobs on Alps infrastructure. Flexible remote work options are available.

Qualifikationen

  • MSc or PhD in Computer Science, Data Science, Artificial Intelligence, Machine Learning, or related field
  • Exceptional BSc candidates with strong engineering experience will also be considered
  • Experience in AI and neural network architectures
  • Strong collaboration and communication skills and ability to work across research and engineering teams
  • Prior hands-on experience in the core domains of this role is required
  • This can be project or study based experience; formal work experience is preferred
  • A high degree of flexibility: priorities, tools, and day-to-day tasks shift with training schedules, releases, and a fast-moving field
  • Hands-on experience with LLM post-training, be it alignment (SFT, preference optimisation) or reinforcement learning
  • Experience with frameworks such as veRL, slime, Megatron-LM, DeepSpeed, TRL, vLLM, SGLang, or similar tools

Aufgaben

  • Build and maintain scalable post-training workflows for Apertus
  • Run and monitor distributed training jobs on HPC infrastructure
  • Debug failures in distributed execution, checkpointing, and GPU utilisation
  • Collaborate with researchers and CSCS engineers to improve reliability of large-scale experiments
  • Support SFT, preference optimisation, and reinforcement learning workflows
  • Develop RL environments and verifier-based training

Kenntnisse

LLM post-training
HPC collaboration
Distributed training
Slurm HPC
Containerization
Research collaboration

Ausbildung

MSc or PhD in CS/DS/AI/ML
Strong engineering experience (BSc)

Tools

veRL
slime
Megatron-LM
DeepSpeed
TRL
vLLM
SGLang

Jobbeschreibung

We are seeking a skilled engineer to join the Apertus post-training effort. The ideal candidate will develop, run, and evaluate the SFT and reinforcement learning pipelines used to turn Apertus base models into capable assistants. This role requires a strong background in LLM post-training, solid software engineering skills, and the ability to work collaboratively in a research-focused HPC environment.

Project background

We train open foundation models with hundreds of billions of parameters on thousands of GPUs on one of the largest AI-ready supercomputers in Europe. The team counts more than a dozen full-time engineers working alongside leading researchers from EPFL and ETH Zürich, has released the Apertus 1 and Apertus 1.5 models, and works with over thirty academic collaborators to deliver fully open (open source), responsibly trained, multilingual, multimodal AI models for research and industry.

Apertus is trained and developed on Alps, the Swiss National Supercomputing Centre's supercomputing infrastructure. The role requires someone who is comfortable working in an HPC environment and collaborating with researchers and infrastructure engineers.

The engineer will contribute to the development, execution, and evaluation of scalable post-training workflows for Apertus.

Infrastructure and systems engineering

  • Build and maintain containerised environments for LLM post-training and RL workloads
  • Adapt containers and dependencies for execution on Alps / CSCS infrastructure
  • Run and monitor Slurm-based training and evaluation jobs
  • Debug failures related to distributed execution, checkpointing, filesystem performance, networking, and GPU utilisation
  • Help maintain reproducible training recipes, configuration files, launch scripts, and documentation
  • Work with researchers and CSCS engineers to improve the reliability and performance of large-scale experiments

LLM post-training and reinforcement learning

  • Support SFT, preference optimisation, and reinforcement learning workflows
  • Build and run RL environments for tasks with verifiable outcomes, such as mathematics, code, tool-use, and reasoning
  • Implement and run reward modelling, reward calibration, and verifier-based training
  • Generate and validate synthetic or gym training tasks
  • Run ablation studies comparing algorithms, reward functions, data mixtures, hyperparameters, and infrastructure settings
  • Evaluate model behaviour across reasoning, coding, mathematics, instruction-following, multilingual, tool-use, and safety benchmarks
  • Debug common post-training issues, including optimisation instability, reward hacking, regressions, and evaluation failures
Profile
  • MSc or PhD in Computer Science, Data Science, Artificial Intelligence, Machine Learning, or a related field
  • Exceptional BSc candidates with strong engineering experience will also be considered
  • Experience in AI and neural network architectures
  • Strong collaboration and communication skills and ability to work across research and engineering teams
  • Prior hands‑on experience in the core domains of this role is required
  • This can be project or study based experience; formal work experience is preferred
  • A high degree of flexibility: priorities, tools, and day‑to‑day tasks shift with training schedules, releases, and a fast‑moving field
  • Hands‑on experience with LLM post‑training, be it alignment (SFT, preference optimisation) or reinforcement learning
  • This means experience with frameworks such as veRL, slime, Megatron‑LM, DeepSpeed, TRL, vLLM, SGLang, or similar tools
Strongly preferred
  • Familiarity with distributed training concepts such as data parallelism, tensor parallelism, pipeline parallelism, checkpointing, and GPU communication
  • Experience with Slurm or another HPC workload manager
  • Experience building or adapting containers for HPC or GPU clusters
Nice to have
  • Published research in the domains relevant to this role, or familiarity with recently published research on these topics
  • Experience creating verifiable tasks for mathematics, code, reasoning, or tool use
  • Familiarity with lower‑level GPU/distributed libraries such as NCCL, Transformer Engine, FlashAttention, or communication backends
  • Experience with large‑scale evaluation pipelines
Workplace
We offer
  • A stimulating academic environment at one of the world's leading technical universities
  • The opportunity to work with state-of-the-art supercomputing infrastructure and cutting‑edge AI research
  • Collaboration with top researchers and engineers from EPFL, ETH Zürich, CSCS, and other Swiss institutions
  • Flexible working arrangements, including options for remote work
  • Professional development opportunities, including conference attendance and specialised training
  • The chance to contribute to open‑source projects with global impact
  • Access to the broader Swiss academic ecosystem and industry partnerships
  • Being part of Switzerland's sovereign AI development, working on technology with national significance

In line with our values , ETH Zurich encourages an inclusive culture. We promote equality of opportunity, value diversity and nurture a working and learning environment in which the rights and dignity of all our staff and students are respected. Visit our Equal Opportunities and Diversity website to find out how we ensure a fair and open environment that allows everyone to grow and flourish. Sustainability is a core value for us – we are consistently working towards a climate‑neutral future .

Curious? So are we.

Further information about the ETH AI Center and the Swiss AI Initiative can be found on our website . Questions regarding the position should be directed to Dr. Imanol Schlag, email ischlag@ethz.ch (no applications).

ETH Zurich is one of the world’s leading universities specialising inscience and technology. We are renowned for our excellent education,cutting-edge fundamental research and direct transfer of new knowledgeinto society. Over 30,000 people from more than 120 countries find ouruniversity to be a place that promotes independent thinking and anenvironment that inspires excellence. Located in the heart of Europe,yet forging connections all over the world, we work together todevelop solutions for the global challenges of today and tomorrow.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Apertus Engineer: Post-training
Apertus Engineer: Post-training

Immigration Policy Lab • Zürich

Vor Ort
CHF 120.000 - 180.000
Remote work options
Professional development
Open-source collaboration
+1
Apertus Engineer: Post-training
Apertus Engineer: Post-training

Master in Integrated Building Systems ETH Zürich • Zürich

Hybrid
CHF 120.000 - 170.000
Flexible working arrangements
Professional development opportunities
Access to cutting-edge HPC
Apertus Engineer: Post-training 100%
Apertus Engineer: Post-training 100%

ETH Zürich • Zürich

Hybrid
CHF 110.000 - 150.000
Remote work options
Professional development
Open‑source projects
Apertus Engineer: Post-training
Apertus Engineer: Post-training

ETH Zürich • Zürich

Hybrid
CHF 110.000 - 170.000
Access to HPC infrastructure
Open-source collaboration
Remote Post-Training LLM Engineer: HPC & RL Pipelines
Remote Post-Training LLM Engineer: HPC & RL Pipelines

Immigration Policy Lab • Zürich

Vor Ort
CHF 120.000 - 180.000
Remote work options
Professional development
Open-source collaboration
+1
Remote HPC LLM Post-Training & RL Engineer
Remote HPC LLM Post-Training & RL Engineer

Eidgenössische Technische Hochschule Zürich • Zürich

Hybrid
CHF 120.000 - 170.000
Remote work options
Access to HPC infrastructure
LLM Post-Training Engineer | HPC & RL Pipelines
LLM Post-Training Engineer | HPC & RL Pipelines

ETH Zürich • Zürich

Hybrid
CHF 110.000 - 170.000
Access to HPC infrastructure
Open-source collaboration
Senior Research Engineer / Research Scientist - Post-Training, Reinforcement Learning & Trainin[...]
Senior Research Engineer / Research Scientist - Post-Training, Reinforcement Learning & Trainin[...]

Giotto.ai • Lausanne

Vor Ort
CHF 140.000 - 210.000
Remote work supported
Swiss office meetups
Competitive compensation
Postdoctoral Positions in Robot Learning and Soft, Musculoskeletal, and Biohybrid Robotics
Postdoctoral Positions in Robot Learning and Soft, Musculoskeletal, and Biohybrid Robotics

Master in Integrated Building Systems ETH Zürich • Zürich

Vor Ort
CHF 85.000 - 120.000
Public transport season tickets
Childcare support
Pension benefits
Head of Geospatial AI - Foundation Models and Geo-Intelligence Platform
Head of Geospatial AI - Foundation Models and Geo-Intelligence Platform

Immigration Policy Lab • Zürich

Vor Ort
CHF 150.000 - 210.000
GeoLab leadership role
Impactful AI platforms
Root/Lucerne location
+2