Consigue una respuesta de este empleador — un currículum y una carta de presentación adaptados exactamente a lo que busca para contratar.
Mercor seeks Latin America-based voice actors with native Latin American Spanish fluency and a neutral, internationally intelligible delivery. The engagement is a single recording session of about four hours, with potential a second session if needed.
Prior voice-acting experience is a plus but not required, and cloning consent is noted. Recordings are for training and evaluating state-of-the-art TTS models, used exclusively by the client’s CX AI Agent.
Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking Latin America-based voice actors with native Latin American Spanish fluency and neutral, internationally intelligible delivery to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is open to both experienced voice professionals and newcomers with a naturally clear, expressive voice and a good recording setup — prior voice-acting experience is a plus but not required.
This is an ongoing opening — we onboard new voices on a rolling basis, so applications are always welcome. The engagement itself is a single recording session of approximately 4 hours. A second session of similar length may follow depending on the client's needs, though this isn't guaranteed.
Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose.
Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.)
Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation
Maintain consistency in voice, accent, and delivery across recording sessions
Follow detailed recording guidelines (environment, microphone setup, file formatting)
Perform multiple takes with variation in emotion, emphasis, and style when required
Native Latin American Spanish speaker currently based in Latin America, with the ability to speak neutral, internationally intelligible Spanish — without a strong regional accent
A clear, natural, and expressive voice (prior voice acting, narration, or broadcast experience is a plus, not required)
Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.)
Strong command of intonation, diction, and emotional range
Ability to follow scripts precisely while maintaining natural delivery
Availability for a single ~4-hour recording session, with the possibility of a further session of similar length later on
[IMP]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case.
Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets
Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper)
Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm)