1

Ai Voice Trainer Jobs (NOW HIRING)

Your voice may be cloned for the clients CX AI Agent so please only apply if you are okay with voice cloning . Preferred * Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets

We are seeking talented voice professionals based in the United States with native American English ... Ensure recordings meet quality standards suitable for AI training and evaluation. Required ...

AI Engineer

Pittsburgh, PA · Remote

$70 - $76/hr

In this role, you will work on developing, training, and refining AI models for voice synthesis, voice cloning, speech recognition, and/or voice transformation. Your work will contribute to cutting ...

... voice samples for training AI systems, or reviewing AI-generated speech to ensure it sounds authentic. Qualifications : Required : • Familiarity with AI workflows • Natural language processing ...

... training, fine-tuning, evaluation PyTorch/TensorFlow) Good to have: Cloud (Azure/AWS/Google Cloud Platform), Docker/Kubernetes Real-time systems, document AI, voice AI Profile: Strong backend ...

next page

Showing results 1-20

Ai Voice Trainer information

See salary details

$13

$24

$36

How much do ai voice trainer jobs pay per hour?

As of Jul 27, 2026, the average hourly pay for ai voice trainer in the United States is $24.74, according to ZipRecruiter salary data. Most workers in this role earn between $18.75 and $26.44 per hour, depending on experience, location, and employer.

What is the difference between Ai Voice Trainer vs Speech Language Pathologist?

AspectAi Voice TrainerSpeech Language Pathologist
CredentialsTypically requires training in AI, machine learning, and voice technology; certifications varyRequires a master's degree in speech-language pathology and state licensure
Work EnvironmentTech companies, AI development labs, remote or office settingsHospitals, clinics, schools, private practices
Industry UsagePrimarily in AI and tech industries focusing on voice recognition and synthesisHealthcare and rehabilitation sectors focusing on speech and language disorders

While both roles involve working with voice, an Ai Voice Trainer focuses on developing and refining voice AI systems using technology and machine learning. In contrast, a Speech Language Pathologist works directly with individuals to diagnose and treat speech and language disorders. The roles differ in credentials, work environment, and industry focus, but both aim to improve voice communication.

What are some common challenges faced by AI Voice Trainers when improving speech recognition accuracy?

AI Voice Trainers often encounter challenges such as handling diverse accents, dialects, and background noises that can affect speech recognition models. Ensuring that training data is inclusive and representative of real-world usage is crucial, as is working closely with linguists and engineers to fine-tune models. Additionally, maintaining user privacy and addressing potential biases in voice datasets are ongoing priorities, requiring collaborative problem-solving and continuous learning.

What are the key skills and qualifications needed to thrive as an AI Voice Trainer, and why are they important?

To thrive as an AI Voice Trainer, you need a strong background in linguistics, phonetics, or computational linguistics, often supported by a relevant degree or experience in speech technology. Familiarity with annotation tools, speech recognition systems, and machine learning platforms is typically required. Excellent attention to detail, cross-cultural communication, and analytical thinking are crucial soft skills for refining voice data and training conversational AI. These skills ensure accurate, inclusive, and natural-sounding AI voice models, directly impacting product usability and user satisfaction.

What is an AI Voice Trainer?

An AI Voice Trainer is a professional who helps develop, refine, and improve the performance of voice recognition and generation systems powered by artificial intelligence. They work by collecting, annotating, and analyzing voice data, as well as training AI models to better understand human speech, accents, and emotions. Their role is crucial in ensuring that voice assistants, speech-to-text applications, and other AI-powered voice technologies are accurate, inclusive, and user-friendly.
More about Ai Voice Trainer jobs
What cities are hiring for Ai Voice Trainer jobs? Cities with the most Ai Voice Trainer job openings:
What states have the most Ai Voice Trainer jobs? States with the most job openings for Ai Voice Trainer jobs include:
Infographic showing various Ai Voice Trainer job openings in the United States as of July 2026, with employment types broken down into 72% Full Time, 21% Part Time, and 7% Contract. Highlights an 61% In-person, and 39% Remote job distribution, with an average salary of $51,453 per year, or $24.7 per hour.
Voice Narrator - AI Trainer

Voice Narrator - AI Trainer

Mercor

San Diego, CA • Remote

$50 - $150/hr

Full-time

Posted 6 days ago


Job description

About the job

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.

Position: Voice Actor: CX Agent Voice Cloning (USA)
Type: Contract
Compensation: $50–$150/hour
Location: Remote
Commitment: 5–10 hours/week

Role Responsibilities

  • Record high-quality voice samples across diverse scripts such as conversational, narrative, and instructional.
  • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation.
  • Maintain consistency in voice, accent, and delivery across recording sessions.
  • Follow detailed recording guidelines including environment, microphone setup, and file formatting.
  • Perform multiple takes with variation in emotion, emphasis, and style when required.

Qualifications

Must-Have

  • Native American English speaker currently based in the United States.
  • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting.
  • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.).
  • Strong command of intonation, diction, and emotional range.
  • Ability to follow scripts precisely while maintaining natural delivery.
  • Reliable availability for 5-10 hours per week over the project duration.
  • [IMP]: Your voice may be cloned for the clients CX AI Agent so please only apply if you are okay with voice cloning.

Preferred

  • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets.
  • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper).
  • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm).

Application Process (Takes 20–30 mins to complete)

  • Upload resume
  • AI interview based on your resume
  • Submit form

Resources & Support

  • For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome
  • For any help or support, reach out to: support@mercor.com

PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.


#hiringmercor