ML & Data Engineer - Create Audio Augmentation Models

Embodied AI

Zürich

Hybrid

CHF 20.000 - 36.000

Teilzeit

Vor 11 Tagen
Bewerbungsgenerator

Verschicke keinen generischen Lebenslauf — erstelle einen Lebenslauf und ein Anschreiben, die genau auf diese Rolle zugeschnitten sind.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Paid Master thesis or internship
One problem that is genuinely yours
Your work goes into the training stack
Flexible hybrid setup

Zusammenfassung

Aurigin.ai in Zurich is building an augmentator to make recordings sound as if captured in different environments. You will design a generative model that takes source audio and a target acoustic scenario and outputs the transformed speech.

The role blends research and implementation in Python with PyTorch, data pipelines, and Acoustic Atlas contributions. It offers a paid Master thesis or internship and a flexible hybrid setup with two days per week in the Zurich office.

Qualifikationen

  • Pursuing or completed MSc in CS, EE, ML, or related field.
  • Strong Python and hands-on ML, ideally PyTorch.
  • Interest in audio, speech, or signal processing.
  • Fluent in English; available for around 6 months; able to be in the Zurich office at least two days a week.

Aufgaben

  • Build the augmentator: a generative model that takes source audio plus a target acoustic scenario and returns that voice as it would have sounded in that setup.
  • Make the scenario space open-ended: represent acoustic conditions so the model can produce combinations it has never seen.
  • Improve how Acoustic Atlas collects: the capture flow, the metadata, and the quality control that decides whether a contribution is usable for training.
  • Get data in at volume, by crowdsourcing, paid recording sessions, open datasets, or a public competition.
  • Prove it out: measure whether detectors trained on generated audio hold up better on replay and real calls; report results clearly.

Kenntnisse

Python
PyTorch
Data pipelines
English communication

Ausbildung

MSc in CS/EE/ML

Tools

Python
PyTorch

Jobbeschreibung

Aurigin.ai in Zurich is building an augmentator to make recordings sound as if captured in different environments. You will design a generative model that takes source audio and a target acoustic scenario and outputs the transformed speech.

The role blends research and implementation in Python with PyTorch, data pipelines, and Acoustic Atlas contributions. It offers a paid Master thesis or internship and a flexible hybrid setup with two days per week in the Zurich office.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

ML / Data Engineer (Master Thesis / Internship)
ML / Data Engineer (Master Thesis / Internship)

Embodied AI • Zürich

Hybrid
CHF 20.000 - 36.000
Paid Master thesis or internship
One problem that is genuinely yours
Your work goes into the training stack
+1
Audio ML Thesis: Real-World Data, Real-Time Detection
Audio ML Thesis: Real-World Data, Real-Time Detection

Embodied AI • Zürich

Hybrid
CHF 20.000 - 27.000
Paid Master thesis
Own Topic
Real Attack Data
+2
Open Audio ML Research (Master Thesis / Internship)
Open Audio ML Research (Master Thesis / Internship)

Embodied AI • Zürich

Hybrid
CHF 20.000 - 27.000
Paid Master thesis
Own Topic
Real Attack Data
+2
Paid AI Research Intern - Flexible Hours & Perks
Paid AI Research Intern - Flexible Hours & Perks

AGIGO • Zürich

Vor Ort
CHF 28.000 - 39.000
Paid internship
Market-level salary
Flexible hours
+1
Staff ML Engineer - Multimodal Gen AI & Synthesis
Staff ML Engineer - Multimodal Gen AI & Synthesis

Apple • Zürich

Vor Ort
CHF 180.000 - 240.000
Research Internship: Audio Toolbox and (Micro) Speech-LLMs
Research Internship: Audio Toolbox and (Micro) Speech-LLMs

AGIGO • Zürich

Vor Ort
CHF 28.000 - 39.000
Paid internship
Market-level salary
Flexible hours
+1
Remote Audio ML Data Engineer - Data Pipelines & Model Dev
Remote Audio ML Data Engineer - Data Pipelines & Model Dev

Logitech • Lausanne

Hybrid
CHF 90.000 - 130.000
Research Internship: Efficient Inference for Text-to-Speech Models and LLMs
Research Internship: Efficient Inference for Text-to-Speech Models and LLMs

AGIGO • Zürich

Vor Ort
CHF 20.000 - 29.000
Paid internship
Flexible hours
Access to GPU clusters (Hopper, Blackw
Bilingual Music Producer for AI Evaluation (German/English)
Bilingual Music Producer for AI Evaluation (German/English)

Mercor • Zürich

Vor Ort
CHF 55.000 - 83.000
Multimodal Data Research Engineer & Dataset Architect
Multimodal Data Research Engineer & Dataset Architect

Microsoft Schweiz GmbH • Zürich

Vor Ort
CHF 110.000 - 140.000