Applied AI Researcher – Foundation Models

Translated Srl

Roma

In loco

EUR 35.000 - 55.000

Tempo pieno

14 giorni+

Ricevi più risposte dai datori di lavoro

Invia un CV specifico per questa offerta in pochi minuti.

Vantaggi offerti da questo lavoro

Flexible remote policy
Periodic presence at Rome HQ

Descrizione del lavoro

Translated is seeking an Applied AI Researcher for its Foundation Models team to advance multilingual Large Language Models, pre-training, and data curation at scale. The role spans model design, experiment execution, and benchmarking across languages, with collaboration across researchers and engineers.

The team works on data selection, scaling laws, and distributed training on large GPU clusters, with periodic presence at the Rome HQ and a flexible remote policy.

Competenze

  • 3+ years of research/industry experience in Deep Learning, ML, NLP, or Large Language Models.
  • Strong understanding of modern deep learning and Transformer-based language models.
  • Excellent programming skills in Python.
  • Experience designing and running machine learning experiments.
  • Familiarity with GPU-based training environments and Unix/Linux systems.
  • Ability to analyze experimental results and make research decisions based on empirical evidence.
  • Interest in large-scale experimental research and language model pre-training.
  • Excellent written and spoken English.
  • Ability to work effectively with both researchers and engineers.

Mansioni

  • Design and conduct research on Large Language Model pre-training.
  • Design experiments, implement them in code, run them at scale, and analyze results.
  • Investigate model architectures, optimization strategies, training dynamics, and scaling behavior.
  • Research and develop multilingual training strategies.
  • Work on data selection, quality, composition, and data mixture experiments for LLM pre-training.
  • Evaluate models across multilingual and general-purpose benchmarks.
  • Monitor and benchmark the state of the art in Large Language Models.
  • Run experiments on large-scale GPU and HPC infrastructure.
  • Translate research hypotheses into measurable experiments and actionable decisions for large-scale training.

Conoscenze

Python
Experiment design
Deep Learning
NLP/LLMs
Unix/Linux
GPU-based training
English communication
Collaboration

Strumenti

Megatron Bridge

Descrizione del lavoro

About Translated

Translated is a leading provider of AI-powered language solutions. Founded in 1999 by a linguist and a computer scientist, we are on a mission to allow everyone to understand and be understood in their own language.

We envision a world where people from different cultures can communicate seamlessly, gaining unprecedented access to knowledge, cultural exchange, and opportunity. To make this possible, we combine proprietary translation AI (Lara Translate) with advanced text, audio, and video translation technologies (Matecat, Matesub, and Matedub) and the world’s largest network of vetted, native-speaking language professionals.

We welcome complex technical challenges from our customers and engineer tailored solutions that often become part of our core products, whether integrated into our enterprise localization platform, the tools designed to support translators, or Lara Translate, our online AI translator for teams and individuals.

Everything we build reflects a simple principle: technology should amplify human potential, not replace it.

We believe in humans.

We are looking for an Applied AI Researcher to join our Foundation Models team, working on the development of large-scale textual language models from scratch.

The ideal candidate has a strong interest in Large Language Models, large-scale deep learning, and experimental research, and is excited about pushing the state of the art in multilingual language modeling.

The Project: Foundation Models

Translated is building its own generation of multilingual Foundation Models, with the ambition of advancing the state of the art in open Large Language Models, particularly in multilingual capabilities.

Our work covers the complete LLM pre-training lifecycle: from data selection and curation to model architecture, scaling experiments, large-scale distributed training, evaluation, and continued training.

A central research question for the team is how to build models that perform strongly not only in English, but across a broad range of languages, including languages that are traditionally underrepresented in large-scale pre-training datasets.

This involves research across areas such as:

  • large-scale language model pre-training
  • multilingual language modeling
  • data curation, filtering, scoring, and mixture design
  • scaling laws and training dynamics
  • model architecture and optimization
  • distributed and multi-GPU training
  • evaluation and benchmarking of Large Language Models
  • continued pre-training and mid-training strategies

You will work in the Foundation Models team, a multidisciplinary group of scientists and engineers responsible for designing, training, and evaluating Translated’s next generation of Large Language Models.

The team works closely with Translated's AI Research and Engineering organizations, combining experimental research with the engineering required to train models at scale on large GPU clusters and HPC infrastructure.

Responsibilities
  • design and conduct research on Large Language Model pre-training
  • design experiments, implement them in code, run them at scale, and analyze their results
  • investigate model architectures, optimization strategies, training dynamics, and scaling behavior
  • research and develop multilingual training strategies
  • work on data selection, quality, composition, and data mixture experiments for LLM pre-training
  • evaluate models across multilingual and general-purpose benchmarks
  • monitor and benchmark the state of the art in Large Language Models
  • run experiments on large-scale GPU and HPC infrastructure
  • translate research hypotheses into measurable experiments and actionable decisions for large-scale training
Requirements
  • 3+ years of research/industry experience in a relevant area of Deep Learning, Machine Learning, Natural Language Processing, or Large Language Models
  • strong understanding of modern deep learning and Transformer-based language models
  • excellent programming skills in Python
  • experience designing and running machine learning experiments
  • familiarity with GPU-based training environments and Unix/Linux systems
  • ability to analyze experimental results and make research decisions based on empirical evidence
  • interest in large-scale experimental research and language model pre-training
  • ability to follow, understand, and reproduce recent scientific literature
  • excellent written and spoken English
  • ability to work effectively with both researchers and engineers
Bonus points if...
  • you have direct experience pre-training or continuing the pre-training of Large Language Models
  • you have experience with distributed and multi-GPU training
  • you have worked with large-scale training frameworks such as Megatron Bridge
  • you have experience optimizing GPU utilization, throughput, memory consumption, or large-scale training stability
  • you have worked on multilingual NLP or multilingual language models
  • you have experience with large-scale dataset curation, filtering, deduplication, or data mixture design
  • you have experience with LLM evaluation and benchmarking
  • you have experience working with HPC environments
Headquarter

Translated is hosted at Pi Campus, a working environment immersed in nature where six luxury villas in Rome, Italy, have been converted into functional offices designed to foster talent growth. Pi Campus is also a venture firm created by Translated to reinvest part of its profits into promising AI startups.

Our Offer

Depending on expertise, the offered salary typically ranges between €35.000,00 and €55.000,00. Compensation generally grows quickly as experience leads to greater contributions. Exceptionally strong candidates may be offered salaries above this range to acknowledge their higher potential impact.

We offer a flexible remote work policy, combined with a requirement for periodic presence at our Rome HQ.

Benefits and Perks

At Translated, we see our people as athletes, and Pi Campus as the place where they can reach their full potential. This vibrant environment fosters talent aggregation and continuous growth of mind, body, and spirit.

We nurture and support the team’s personal and professional development every day through growth-oriented initiatives like personalized learning paths and networking events with industry leaders, health-oriented initiatives like massages, sauna, and workouts in our gym and swimming pool, as well as relaxation rooms, open meeting rooms, a cafeteria, and fully equipped kitchens in every villa. In case of need, every team member has access to psychological, legal, and financial support.

Learn more about our company:

https://translated.com/work-at-translated-onboarding

Diversity Statement

At Translated, we proudly embrace and celebrate each individual's unique qualities to our team, regardless of race, sexual orientation, gender identity, or any other differences. We recognize that these diverse perspectives empower us to overcome challenges, foster innovation, and drive excellence. As an inclusive and equal-opportunity employer, we are committed to cultivating an environment where everyone feels welcome, valued, and supported to achieve their full potential.

Privacy Policy

Ottieni la revisione del curriculum gratis e riservata.
o trascina qui il file.
Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Applied AI Researcher – Foundation Models
Applied AI Researcher – Foundation Models

Translated • Roma

Ibrido
EUR 35.000 - 55.000
Flexible remote work policy
Presence at Rome HQ
Wellness and gym facilities
Senior Deep Learning Scientist (On-site)
Senior Deep Learning Scientist (On-site)

Translated • Roma

In loco
EUR 48.000 - 68.000
Wellness program
Gym and pool access
Massages and sauna
+3
Multilingual Foundation Models Research Scientist
Multilingual Foundation Models Research Scientist

Translated • Roma

Ibrido
EUR 35.000 - 55.000
Flexible remote work policy
Presence at Rome HQ
Wellness and gym facilities
Remote Multilingual Foundation Models Researcher
Remote Multilingual Foundation Models Researcher

Translated Srl • Roma

Ibrido
EUR 35.000 - 55.000
Flexible remote policy
Periodic presence at Rome HQ
Full Stack Engineer (On-site)
Full Stack Engineer (On-site)

Translated • Roma

Ibrido
EUR 35.000 - 55.000
Remote-friendly policy
Gym and wellness facilities
On-site HQ at Pi Campus (Rome)
+2
Senior CloudOps Engineer (On-site) at Translated
Senior CloudOps Engineer (On-site) at Translated

Translated • Roma

In loco
EUR 38.000 - 58.000
Gym & swimming pool
Massages & wellness program
Cafeteria & kitchens
Senior Backend Engineer - LARA API (On-site)
Senior Backend Engineer - LARA API (On-site)

Translated • Roma

In loco
EUR 35.000 - 55.000
Senior Backend Engineer - Lara API Team (On-site)
Senior Backend Engineer - Lara API Team (On-site)

Translated Srl • Roma

In loco
EUR 35.000 - 55.000
Senior CloudOps Engineer (On-site)
Senior CloudOps Engineer (On-site)

Translated • Roma

Ibrido
EUR 38.000 - 58.000
Flexible remote policy
Periodic presence at Rome HQ
Wellness facilities
Senior Deep Learning Scientist (On-site)
Senior Deep Learning Scientist (On-site)

Translated Srl • Roma

In loco
EUR 70.000 - 100.000
Personalized learning paths
Networking events
Health support initiatives
+1