1

Speech Research Engineer Jobs (NOW HIRING)

$200K - $400K/yr

About the Role As a Research Engineer focused on Audio and Speech, you'll be responsible for building the models and agent harnesses that power Decagon's real-time voice agents and taking them all ...

$200K - $400K/yr

About the Role As a Research Engineer focused on Audio and Speech, you'll be responsible for building the models and agent harnesses that power Decagon's real-time voice agents and taking them all ...

About the Role As a Research Engineer focused on Audio and Speech, you'll be responsible for building the models and agent harnesses that power Decagon's real-time voice agents and taking them all ...

About The Role As a Research Engineer at Phonic, you'll sit at the intersection of cutting-edge ML ... You'll design, train, and iterate on models across the voice stack (speech, audio, language, and ...

Engineering Group, Engineering Group > Multimedia Systems General Summary: Qualcomm's Multimedia R&D Group is seeking candidates to join its world-class Speech team developing novel technologies ...

Engineering Group, Engineering Group > Multimedia Systems General Summary: Qualcomm's Multimedia R&D Group is seeking candidates to join its world-class Speech team developing novel technologies ...

About The Role As a Research Engineer at Phonic, you'll sit at the intersection of cutting-edge ML ... You'll design, train, and iterate on models across the voice stack (speech, audio, language, and ...

Speech & Audio Research Engineer

San Diego, CA ยท On-site

$148K - $222K/yr

Research, development, and standardization of new conversational speech codecs for wireless ... Programming skills in C/C++ and Python, and experience in Windows or Unix/Linux development ...

AI Research Engineer- Speech 1

Redmond, WA ยท On-site

$229K/yr

About Job AI Engineer: Speech/Audio Centific AI Research About Centific AI Research Centific AI Research is at the forefront of developing cutting-edge AI solutions that bridge the gap between ...

AI Research Engineer- Speech 1

Redmond, WA ยท On-site

$229K/yr

About Job AI Engineer: Speech/Audio Centific AI ResearchAbout Centific AI Research Centific AI Research is at the forefront of developing cutting-edge AI solutions that bridge the gap between ...

Research Engineer About Us Lotus Health is a groundbreaking primary care app that integrates your ... Experience building speech or multimodal pipelines for medical settings * Contributions to open ...

next page

Showing results 1-20

Speech Research Engineer information

See salary details

$37K

$106K

$142.5K

How much do speech research engineer jobs pay per year?

As of Sep 13, 2026, the average yearly pay for speech research engineer in the United States is $106,012.00, according to ZipRecruiter salary data. Most workers in this role earn between $104,000.00 and $104,000.00 per year, depending on experience, location, and employer.

What is a speech research engineer?

Speech Research Engineers are professionals who design, develop, and optimize technologies that enable computers to understand, interpret, and generate human speech. They work at the intersection of signal processing, machine learning, and linguistics to improve applications such as speech recognition, voice synthesis, and language translation systems. Their responsibilities include creating algorithms, training models on large datasets, and evaluating the performance of speech systems. These engineers often collaborate with researchers and product teams to bring cutting-edge speech technology into real-world applications.

What are the key skills and qualifications needed to thrive as a speech research engineer, and why are they important?

To thrive as a Speech Research Engineer, you need a strong background in signal processing, machine learning, and proficiency in programming languages like Python or C++, often backed by an advanced degree in computer science, electrical engineering, or a related field. Experience with frameworks such as TensorFlow or PyTorch and familiarity with speech recognition toolkits like Kaldi or HTK are typically required. Strong analytical thinking, creativity, and effective communication skills help you design innovative algorithms and collaborate in multidisciplinary teams. These skills are crucial for developing robust speech technologies and driving advancements in voice-driven applications.

What are common challenges faced by speech research engineers when working with real-world speech data?

Speech Research Engineers often encounter diverse challenges when handling real-world speech data, such as dealing with background noise, accents, dialectal variations, and inconsistent audio quality. These factors can significantly affect the accuracy and robustness of speech recognition models. Engineers must frequently preprocess and augment data, implement noise-robust algorithms, and collaborate closely with linguists and data scientists to ensure models perform well across various environments and populations.

What are popular job titles related to Speech Research Engineer jobs?

For Speech Research Engineer jobs, the most frequently searched job titles are:

Infographic showing various Speech Research Engineer job openings in the United States as of September 2026, with employment types broken down into 1% Internship, 1% As Needed, 88% Full Time, 9% Part Time, and 1% Contract. Highlights an 79% Physical, 3% Hybrid, and 18% Remote job distribution, with an average salary of $106,012 per year, or $51 per hour.

Research Engineer, Audio and Speech

San Francisco, CA โ€ข On-site

$200K - $400K/yr

Other

Medical, Dental, Vision, Life, Retirement, PTO

Posted 6 days ago


Job description

About Decagon

Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experiences.

Our technology enables industry-defining enterprises like Avis Budget Group, Blockโ€™s Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents that power personalized, deeply satisfying interactions across voice, chat, email, SMS, and every other channel.

Weโ€™re building a future where customer experiences are being redefined from support tickets and hold music to faster resolutions, richer conversations, and deeper relationships. Weโ€™re proud to be backed by world-class investors who share that vision, including a16z, Accel, Bain Capital Ventures, Coatue, and Index Ventures, along with many others.

Weโ€™re an in-office company, driven by a shared commitment to excellence and velocity. Our values โ€” Just Get It Done, Invent What Customers Want, Winnerโ€™s Mindset, and The Polymath Principle โ€” shape how we work and grow as a team.

About the Team

Read more about the Speech Research Team's work:

  • https://decagon.ai/blog/audio-native-semantic-speaker-change-detection
  • https://decagon.ai/blog/scaling-real-time-tts-inference
  • https://decagon.ai/blog/teaching-flow-matching-tts-with-rl

The Research team develops the model and decision-making stack that powers Decagonโ€™s conversational agents for enterprise support. We research, adapt, and implement state-of-the-art techniques in model training, prompting, orchestration, and evaluation in order to make our agents more accurate, robust, and efficient in real-world deployments.

Our goal is to push the frontier of applied conversational AI: agents that reliably understand nuanced intent, track long context, and take the right actions under uncertainty. We measure success the way customers feel it: higher resolution rates, better user satisfaction, and consistent behavior at scale.

About the Role

As a Research Engineer focused on Audio and Speech, youโ€™ll be responsible for building the models and agent harnesses that power Decagonโ€™s real-time voice agents and taking them all the way from idea to production. Your work will advance multimodal and full-duplex systems that can listen, reason, speak, and respond naturally in real time.

Weโ€™re looking for strong engineers who want to build the next generation of AI voice agents. People here own their work end-to-end, ship real improvements, and are trusted to make high-impact technical decisions.

In this role, you will
  • Design and build next-generation agent harnesses optimized for streaming speech, turn-taking, interruptions, overlapping speech, and continuous interaction
  • Research and train multimodal and full-duplex models that jointly understand audio, reason, and generate speech
  • Improve speech recognition, voice activity detection, endpointing, and speech generation across diverse speakers, environments, domains, and languages
  • Build evaluations and use production calls to ship measurable improvements in accuracy, latency, naturalness, and task outcomes
  • Optimize end-to-end inference for responsiveness, throughput, stability, and cost, partnering with Voice Platform and Infrastructure teams to deploy at scale
Your background looks something like this
  • 2+ years of experience in speech, audio ML, multimodal ML, or production machine learning
  • Experience developing or adapting autoregressive, diffusion, flow-matching, or codec-based speech models
  • Hands-on experience with streaming agent systems, low-latency inference, production model serving, and evaluation on real-world audio
  • Fluency in Python and a modern deep-learning framework such as PyTorch, with strong foundations in machine learning and signal processing
  • A track record of taking research ideas from prototype to reliable, measurable production impact
Even better if you have
  • Familiarity with speech-to-speech or full-duplex models
  • Experience with telephony, multilingual speech, noisy-channel robustness, speaker adaptation, or expressive speech generation
Compensation

$200K โ€“ $400K + Offers Equity

Benefits

We proudly offer the following benefits for our full-time employees:

  • Medical, Dental, and Vision benefits for you and your family
  • Life Insurance and Disability Benefits
  • Retirement Plan (e.g., 401K, pension)
  • Parental Leave
  • Fertility and family building benefits through Carrot
  • Monthly stipend to support your wellness, lifestyle, and work-life balance
  • Daily lunches and snacks in the office to keep you at your best
  • Take what you need vacation policy (subject to local requirements; UK employees receive 25 days of statutory leave)

These benefits are described in more detail in Decagonโ€™s policies, may vary by location, and can change at any time according to applicable compensation and benefits plans.

#J-18808-Ljbffr