Audio Deepfake Research Intern (Remote)

Reality Defender

United States

Remote

USD 55,800 - 89,280

Part time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

Reality Defender seeks an Applied Scientist Intern (Audio) to advance research in multimodal deepfake detection and audio understanding. This remote role partners with our research scientists and engineers to develop, fine-tune, and evaluate models that operate across speech, text, and other modalities.

You will contribute to building robust, scalable detection pipelines, publish findings, and help shape a privacy-first platform used by enterprises and governments.

Qualifications

  • Currently enrolled in a PhD program with specialization in machine learning/deep learning, natural language processing, and/or speech processing.
  • 2+ years of experience with training/fine-tuning large models, esp. audio language models, speech foundation models, multi-modal foundation models.
  • Experience with end-to-end model building pipeline for ML tasks: dataset curation/cleaning, model implementation, benchmarking, and result analysis
  • Familiarity with distributed multi-GPU training
  • Prior publications in ML/Audio/NLP venues (e.g., NeurIPS, Interspeech, ICASSP, ACL, EMNLP).

Responsibilities

  • Explore and conceptualize novel methods to leverage different modalities (e.g., speech, text) for deepfake detection and relevant audio understanding tasks.
  • Perform fundamental and applied research to advance the current state-of-the-art on audio deepfake detection.
  • Build models with generalizability to unseen generative methods - Collaborate with scientists and engineers across the organization
  • Summarize, publish, and present research findings

Skills

PhD in ML/NLP/Speech
Fine-tuning large models (audio)
End-to-end ML pipeline
Distributed multi-GPU training

Education

PhD program in ML/deep learning/NLP/speech processing

Job description

Reality Defender seeks an Applied Scientist Intern (Audio) to advance research in multimodal deepfake detection and audio understanding. This remote role partners with our research scientists and engineers to develop, fine-tune, and evaluate models that operate across speech, text, and other modalities.

You will contribute to building robust, scalable detection pipelines, publish findings, and help shape a privacy-first platform used by enterprises and governments.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Applied Scientist Intern (Audio) at Reality Defender Remote
Applied Scientist Intern (Audio) at Reality Defender Remote

Reality Defender • United States

Remote
Equity
Benefits
Audio ML Research Intern: Turn Research into Real Voice AI
Audio ML Research Intern: Turn Research into Real Voice AI

Bland AI • San Francisco (CA)

On-site
USD 40,000 - 65,000
Competitive intern compensation
Mentorship from researchers on front‑f
All the tools you need
+2
Senior Audio-to-Audio AI Research Scientist
Senior Audio-to-Audio AI Research Scientist

Google Inc. • Mountain View (CA), New York (NY)

On-site
USD 207,000 - 300,000
Remote Video AI Research Scientist II - Deepfake Detection
Remote Video AI Research Scientist II - Deepfake Detection

Remote Jobs • United States

Remote
USD 160,000 - 185,000
RSUs (Restricted Stock Units)
Unlimited Paid Time Off (PTO)
Generous health and welfare plans
+1
Remote Applied Scientist - Deep Learning (Audio AI)
Remote Applied Scientist - Deep Learning (Audio AI)

Incept • United States

On-site
USD 100,000 - 130,000
Remote Research Engineer — AI Audio & Models
Remote Research Engineer — AI Audio & Models

ElevenLabs • Town of Poland (NY)

On-site
USD 140,000 - 190,000
Lead Audio AI Research for Multimodal Systems
Lead Audio AI Research for Multimodal Systems

Thinking Machines Lab Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 350,000 - 475,000
Health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
Research, Audio Expertise
Research, Audio Expertise

Mosaic.tech • San Francisco (CA)

On-site
USD 350,000 - 475,000
Health benefits
Dental benefits
Vision benefits
+3
Research Scientist (Fixed-Term) — Speech & Audio AI
Research Scientist (Fixed-Term) — Speech & Audio AI

Spotify • United States

On-site
USD 48,000 - 62,000
Remote ML Engineer - DeepFake Media & AI Security
Remote ML Engineer - DeepFake Media & AI Security

dHired.com • Agency (MO)

On-site
USD 90,000 - 130,000
Flexible working hours
Significant professional growth
Tea, coffee, water, snacks