Voice Actor for AI Text-to-Speech Development at Mercor | Torre

Voice Actor for AI Text-to-Speech Development

Emma highlights
This highlight was written by Emma’s AI. Ask Emma to edit it.
Freelance
A project
Compensation
USD50 - 150/hour
Non-negotiable
location_on
Remote (for United States residents)
flightsmode
Visa sponsorship: No
Match
skeleton-gauges
You have opted out of job matches in .
Posted 2 months ago

Responsibilities and deliverables


We are partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking USA-based voice actors with native American English accents to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. Key Responsibilities Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation Maintain consistency in voice, accent, and delivery across recording sessions Follow detailed recording guidelines (environment, microphone setup, file formatting) Perform multiple takes with variation in emotion, emphasis, and style when required Requirements Native American English speaker currently based in the United States Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) Strong command of intonation, diction, and emotional range Ability to follow scripts precisely while maintaining natural delivery Reliable availability for 5-10 hours per week over the project duration [IMP]: Your voice may be cloned for the clients CX AI Agent so please only apply if you are okay with voice cloning Preferred Qualifications Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets