1

Voice Model Jobs in Illinois (NOW HIRING)

Seeking a Superstar Voice Teacher Do you love working with kids? Do you want to make a difference ... We believe in a family first business model where we are raising tomorrows leaders through the ...

IL

$107K - $147K/yr

Optimize for Voice: Drive selective Small Language Model (SLM) fine-tuning and inference optimization for voice latency and cost. Qualifications * You have shipped production AI agents serving real ...

Experience with deliverability, routing, or pricing models. * Knowledge of CLI rules, emergency routing, or other Voice compliance topics. * Experience with WebRTC, Voice bots, OTT Voice, or AI ...

IL ยท On-site

$107K - $147K/yr

Optimize for Voice: Drive selective Small Language Model (SLM) fine-tuning and inference optimization for voice latency and cost. Qualifications * You have shipped production AI agents serving real ...

next page

Showing results 1-20

Voice Model information

What is a voice model?

A Voice Model job involves providing high-quality voice recordings that are used to create or enhance AI-powered speech systems. These recordings help train text-to-speech (TTS) models, virtual assistants, and other voice-enabled applications. Voice Models may work on projects requiring natural speech, emotional expression, or specific accents and tones. The role can involve reading scripts, responding to prompts, or improvising speech patterns to capture a variety of vocal nuances.

What are the key skills and qualifications needed to thrive in a voice model position, and why are they important?

To thrive as a Voice Model, you need an excellent command of vocal techniques, clear diction, and the ability to adjust your voice for various styles or characters, often supported by formal vocal or acting training. Familiarity with professional recording equipment, audio editing software, and sometimes home studio setups is essential. Adaptability, reliability, and taking direction well are important soft skills for succeeding in client-driven environments. These skills and qualities ensure Voice Models can deliver high-quality vocal performances that meet diverse client needs across advertising, entertainment, and media industries.

What are the typical work arrangements and environments for a voice model?

Voice Models often work on a freelance basis or as part of talent agencies, providing voice recordings for commercials, animations, video games, and audiobooks. Many professionals operate from home studios using specialized equipment, while some projects require sessions at recording studios with directors and sound engineers present. Flexibility is important, as schedules can include last-minute bookings and varied project durations. Collaboration with creative teams, such as producers and scriptwriters, is common to ensure that the final product matches the intended vision. This dynamic environment offers both autonomy and opportunities for skill development across different media industries.

What are popular job titles related to Voice Model jobs in Illinois?

For Voice Model jobs in Illinois, the most frequently searched job titles are:

What job categories do people searching Voice Model jobs in Illinois look for?

The top searched job categories for Voice Model jobs in Illinois are:

Infographic showing various Voice Model job openings in Illinois as of August 2026, with employment types broken down into 1% As Needed, 79% Full Time, 18% Part Time, and 2% Contract. Highlights an 87% Physical, 3% Hybrid, and 10% Remote job distribution.

Principal Research Scientist Speech Voice Foundation Models

Intelix.AI

Mundelein, IL โ€ข On-site

Other

Posted 13 days ago


Job description

Principal Research Scientist Speech & Audio Foundation Models

Text-to-speech, voice cloning, speech synthesis, realtime conversational voice.


$270,000โ€“$500,000 base plus bonus, equity and benefits (US).

Relocation assistance available. Visa transfer supported

San Francisco on-site preferred | Remote considered in the US, UK and parts of Europe.

Permanent, full-time.



A top end research lab building realtime voice models text-to-speech, speech-to-text and speech-to-speech delivered as an API. The models run in production behind consumer applications used at very large scale, across health, learning, therapy, companionship, media and gaming.


Text-to-speech, voice cloning, speech synthesis, realtime conversational voice.



The role


Build the models that are the product! This is full-stack research ownership: you frame the question, run the experiments, and ship the result. Research is only finished when it is in production and measurable.



Responsibilities



  • Train foundation models: pre-training, reinforcement learning, reward modelling, post-training, new architectures, scaling.
  • Build and improve voice and speech models across TTS, STT and speech-to-speech.
  • Design the evaluation that proves the work: benchmarks, eval loops, LLM-as-judge, failure analysis. Evaluation is treated as a research product in its own right, not as a pre-launch checkbox.
  • Work on frontier problems adjacent to the roadmap: multimodal, agents and tool use, test-time compute.
  • Take models into production alongside the serving engineering team, inside a sub-200ms latency budget and across 100+ languages.



Essential



  • Hands-on foundation-model training. Pre-training, RL, reward modelling, post-training, scaling. Fine-tuning or building on top of someone else's model is a different discipline and is not what this role is.
  • Real voice or speech research: TTS, STT or speech-to-speech. Speech-to-speech is the strongest signal; TTS and ASR both count. Text-only research does not transfer.
  • Evidence you can point at: papers, shipped models, open-source contributions, or systems in production.



Desirable



  • Evaluation depth: benchmarks, eval loops, quality measurement, failure analysis.
  • Publications at ICML, ICLR, NeurIPS, EMNLP, ACL, AAAI, Interspeech or ICASSP.
  • PhD in ML or NLP, or equivalent practical experience you can point to.
  • Frontier exposure: multimodal, agents, tool use, test-time compute.
  • Public work: side projects, open-source, technical write-ups.