Doctoral Researcher in Speech and Language Technology

Aalto University Executive Education Oy

Espoo, Helsinki

Hybrid

EUR 32,000 - 38,000

Full time

5 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Aalto University invites applications for a Doctoral Researcher in Speech and Language Technology in Espoo, Finland. The role is hybrid, starting January 2027, under the COWAMA project led by Assistant Professor Lauri Juvela.

You will develop methods to detect generated audio content, build datasets, and publish in leading venues while contributing open-source software. Your profile includes a Master’s degree (or near completion), strong Python skills, and interest in deep learning,

Qualifications

  • Master’s degree in a relevant field or near completion.
  • Strong interest in doctoral research on audio deepfake detection and watermarking.
  • Good programming skills, preferably Python.
  • Basic knowledge of deep learning and signal processing.
  • Interest in working with speech/audio datasets and detector models.
  • Motivation to publish scientific articles and contribute to open-source.
  • Ability to work independently and in international teams.

Responsibilities

  • Detect generated audio content from multiple sources.
  • Create diverse datasets of real and generated speech/audio.
  • Adapt detector baselines for multi-objective detection.
  • Develop explainable and localized detection methods for watermarking and deepfake attribution.
  • Study robustness against signal-processing and adversarial attacks.
  • Publish research and contribute to open-source software and models.

Skills

Python programming
Deep learning basics
Signal processing basics
Research independence
International collaboration
Academic writing/publishing

Education

Master’s degree (or near completion)

Tools

Python
Git

Job description

Doctoral Researcher in Speech and Language Technology

18.9.2026

Application closes on

31.10.2026

Unit

School of Electrical Engineering

Job category

Doctoral Researchers

Aalto University is where science and art meet technology and business. We shape a sustainable future by making research breakthroughs in and across our disciplines, sparking the game changers of tomorrow and creating novel solutions to major global challenges. Our community is made up of 16 000 students and 5 200 employees, including 446 professors. Our campus is in Espoo, Greater Helsinki, Finland. Diversity is part of who we are, and we actively work to ensure our community’s diversity and inclusiveness. This is why we warmly encourage qualified candidates from all backgrounds to join our community.

We are now looking for a Doctoral Researcher in Speech and Language Technology

Are you excited about speech and audio AI, deepfake detection, and the question of how we can identify the origin of AI-generated content? We are looking for aDoctoral Researcherto join the funded projectCOWAMA: Content-based watermarking for AI-generated audio, led by Assistant Professor Lauri Juvela.

The rapid development of generative AI has made it possible to create realistic speech and music, but it has also created risks related to misinformation, malicious deepfakes, copyright, ownership, attribution, and accountability. The project addresses these challenges by developing methods for detecting the origin and authenticity of generated and watermarked audio data, improving content-based audio watermarking, and evaluating robustness against removal, spoofing, adversarial attacks, and realistic downstream processing.

Your role and goals

As the Doctoral Researcher, you will take primary responsibility for detecting generated audio content from multiple sources, and work jointly with the Postdoctoral Researcher onrobustness, evaluation, and generalization.

Your work will include
  • Developing methods for detecting the authenticity and origin of speech and audio content generated by multiple models and containing different watermarking methods.
  • Creating datasets of diverse real and generated speech and audio, including content from different generative models and multiple watermarking methods.
  • Adapting and improving detector baselines for multi-objective detection involving deepfake detection, content origin detection, and watermark identification.
  • Developing explainable and localized detection methods for watermark and deepfake attribution, especially when generated content is mixed with other audio sources.
  • Studying robustness against signal-processing attacks, generative resynthesis attacks, adversarial attacks, and realistic downstream processing such as mixing.
  • Publishing research in leading journals and conferences in speech, audio, and machine learning, and contributing to open-source releases of software, trained models, and reproducible research outputs.
Your network and team

You will be supervised byAssistant Professor Lauri Juvela, who leads the Speech Synthesis research group at Aalto University. The group works on deep generative models for speech and audio, speech synthesis, differentiable signal processing, deepfake detection, and watermarking.

In addition to the Aalto speech groups and the Postdoctoral Researcher in the project, you will work with an international collaboration network including Prof. Junichi Yamagishi at the National Institute of Informatics in Japan, Prof. Xavier Serra at Universitat Pompeu Fabra in Spain, Prof. Gustav Eje Henter at KTH Royal Institute of Technology in Sweden.

Your experience and ambitions

We are looking for a curious and motivated early-career researcher who wants to develop expertise in speech, audio, machine learning, and trustworthy generative AI. To succeed in this role, you should have:

  • A master’s degree, or be close to completing a master’s degree, in speech or audio processing, machine learning, signal processing, computer science, electrical engineering, or a related field.
  • Strong interest in doctoral research on audio deepfake detection, watermarking, source tracing, and generative AI.
  • Good programming skills, preferably including Python and modern machine learning tools.
  • Basic knowledge of deep learning, signal processing, speech/audio processing, or statistical machine learning.
  • Interest in working with speech and audio datasets, detector models, generative audio models, and experimental evaluation.
  • Motivation to publish scientific articles and contribute to open-source and reproducible research.
  • Ability to work both independently and collaboratively in an international research environment.

Fluency in English is required. Finnish language is not required.

If you are chosen for this position, you will apply for the study right in doctoral studies at Aalto University School of Electrical Engineering. Thus, please see the student information and admission criteria at https://www.aalto.fi/en/study-options/aalto-doctoral-programme-in-electrical-engineering.

What we offer
  • A doctoral research position in a timely and socially meaningful field: improving transparency, traceability, and accountability for AI-generated speech and audio.
  • The opportunity to work on technical methods that support safer use of generative AI, including deepfake detection, watermark attribution, and robustness evaluation.
  • Excellent computing and audio research infrastructure, including CPU/GPU clusters, access to CSC and LUMI, FIN-CLARIN resources, and the Aalto Acoustics Lab.
  • A supportive research team and supervision by Assistant Professor Lauri Juvela.
  • International collaboration opportunities with leading researchers in speech synthesis, music technology, deepfake detection, and generative audio.
  • A strong open-science environment: the project aims to publish in leading venues and release software source code and trained models to support reproducibility and FAIR data management.
  • Great possibilities for competence development and learning, including professional development opportunities, staff training, and development projects based on your interests and needs.
  • A culture guided by responsibility, courage, and collaboration, where equality and inclusion support curiosity, innovation, collaboration, and wellbeing.
  • Our vast array of professional development opportunities means you will grow and learn, having the chance to participate actively in staff training and development projects based on your interests and needs.

The starting salary for this position is 3143 €/month and increases after a mid-term evaluation. The position is fixed term and follows school’s standard 2+2 model. It will be made initially for two years, with a six-month probationary period, and extended by two further years after a successful mid-term review, giving a total duration of four years. The position starts in January 2027 or as mutually agreed.

We value work-life balance and well-being in all aspects of life. We work in a hybrid model, with the primary workplace located at the Otaniemi Campus in Espoo, Finland. Life on the revitalized campus is vibrant, featuring stunning architecture, tranquil nature, and a variety of cafes, restaurants, and services, all complemented by excellent public transportation connections.

For more information about the role, please contact Professor Lauri Juvela (lauri.juvela@aalto.fi; +358 50 464 6653). For questions about the application process, please contact HR Advisor Johanna Haapalainen at hr-elec@aalto.fi.

We will go through applications, and we may invite suitable candidates to interview already during the application period. You will hear from us the latest in the second week of November. We aim to have a transparent and equal recruitment process, so feel free to ask us for feedback.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Postdoctoral Researcher in Speech and Language Technology
Postdoctoral Researcher in Speech and Language Technology

Aalto University Executive Education Oy • Espoo

Hybrid
EUR 43,000 - 51,000
Open science & open-source releases
Excellent computing infrastructure
Hybrid work model
Doctoral Researcher in Speech and Language Technology
Doctoral Researcher in Speech and Language Technology

Aalto University • Espoo

Hybrid
EUR 32,000 - 38,000
Hybrid work model
International collaboration
Open science environment
+1
Doctoral Researcher in Speech and Language Technology
Doctoral Researcher in Speech and Language Technology

Aalto University • Finland

Hybrid
EUR 32,000 - 38,000
Doctoral Researcher in Speech and Language Technology
Doctoral Researcher in Speech and Language Technology

Aalto • Espoo

Hybrid
EUR 30,000 - 40,000
Postdoctoral Researcher in Speech and Language Technology
Postdoctoral Researcher in Speech and Language Technology

Aalto University • Finland

Hybrid
EUR 43,000 - 51,000
PhD Researcher in Speech & Audio Watermarking & Detection
PhD Researcher in Speech & Audio Watermarking & Detection

Aalto University • Finland

Hybrid
EUR 32,000 - 38,000
PhD Researcher: Speech & Audio AI & Deepfake Detection
PhD Researcher: Speech & Audio AI & Deepfake Detection

Aalto University Executive Education Oy • Espoo, Helsinki

Hybrid
EUR 32,000 - 38,000
Tutkijatohtori puhe- ja kieliteknologian alalle
Tutkijatohtori puhe- ja kieliteknologian alalle

Aalto University Executive Education Oy • Espoo

Hybrid
EUR 40,000 - 54,000
PhD Candidate in Speech & Audio Deepfake Detection
PhD Candidate in Speech & Audio Deepfake Detection

Aalto University • Espoo

Hybrid
EUR 32,000 - 38,000
Hybrid work model
International collaboration
Open science environment
+1
PhD Researcher in Speech & Audio AI: Watermarking & Deepfakes
PhD Researcher in Speech & Audio AI: Watermarking & Deepfakes

Aalto • Espoo

Hybrid
EUR 30,000 - 40,000