1

Speech Research Jobs (NOW HIRING)

$150K - $225K/yr

In this role, you'll have the opportunity to work on innovative foundational research in speech technology and machine learning. As a member of the team, you will be inspired by a diversity of ...

AI Research Engineer- Speech 1

Redmond, WA Β· On-site

$229K/yr

Speech/Audio Centific AI Research About Centific AI Research Centific AI Research is at the forefront of developing cutting-edge AI solutions that bridge the gap between research innovation and real ...

Qualcomm's Multimedia R&D Group is seeking candidates to join its world-class Speech team developing novel technologies enabling the future of real-time communication and collaboration. We conduct ...

Qualcomm's Multimedia R&D Group is seeking candidates to join its world-class Speech team developing novel technologies enabling the future of real-time communication and collaboration. We conduct ...

AI Research Engineer- Speech 1

Redmond, WA Β· On-site

$229K/yr

Speech/Audio Centific AI ResearchAbout Centific AI Research Centific AI Research is at the forefront of developing cutting-edge AI solutions that bridge the gap between research innovation and real ...

General Summary Qualcomm's Multimedia R&D Group is seeking candidates to join its world‑class Speech team developing novel technologies enabling the future of real‑time communication and ...

We are looking for researchers with deep expertise in one of two areas: * Speech & Audio: Build state-of-the-art medical speech-to-text systems using our large proprietary dataset of real-world ...

We are looking for researchers with deep expertise in one of two areas: * Speech & Audio: Build state-of-the-art medical speech-to-text systems using our large proprietary dataset of real-world ...

Own the TTS research and model roadmap - decide which technical directions can materially move speech-generation quality, including the expensive and non-obvious ones, and recognize when an approach ...

Speech Therapist

Fair Lawn, NJ Β· On-site

$65 - $73/hr

... research and best practices in speech-language pathology Job Type: Part-time Pay: $68.00 - $73.00 per hour Expected hours: 6 per week Benefits: * 401(k) * 401(k) matching * Continuing education ...

Showing results 21-40

speech research information

See salary details

$12

$45

$67

How much do speech research jobs pay per hour?

As of Sep 13, 2026, the average hourly pay for speech research in the United States is $45.78, according to ZipRecruiter salary data. Most workers in this role earn between $33.65 and $56.49 per hour, depending on experience, location, and employer.

What is speech research?

A Speech Research job involves studying and developing technologies related to speech processing, recognition, synthesis, and understanding. Researchers in this field work on improving voice assistants, automatic transcription, language modeling, and speech-based AI systems. The role often involves machine learning, linguistics, signal processing, and artificial intelligence to enhance human-computer interaction. Professionals in speech research may work in academia, tech companies, or research institutions to advance the capabilities of speech-driven applications.

What are the typical responsibilities of a speech researcher within a team?

Speech Researchers often work as part of interdisciplinary teams, collaborating with engineers, data scientists, and product managers to design and conduct experiments involving speech recognition, synthesis, or processing. Daily tasks may include collecting and annotating audio data, developing algorithms, analyzing linguistic features, and publishing findings in peer-reviewed forums. Progress is usually shared in regular team meetings, and researchers are encouraged to stay current with advancements in the field. This collaborative environment fosters innovation and the practical application of research in real-world speech technology products.

What are the key skills and qualifications needed to thrive in speech research, and why are they important?

To thrive in Speech Research, you typically need a solid background in linguistics, computational linguistics, or a related field, with strong analytical and experimental design skills. Familiarity with speech analysis software, programming languages such as Python, and tools like Praat or MATLAB is often essential. Exceptional problem-solving abilities, attention to detail, and collaborative communication skills help you excel in multidisciplinary research teams. These competencies enable researchers to conduct high-quality experiments and contribute innovative advances in speech technology.

More about speech research jobs

What cities are hiring for Speech Research jobs?

Cities with the most Speech Research job openings:

What are the most commonly searched types of Speech Research jobs?

The most popular types of Speech Research jobs are:

What states have the most Speech Research jobs?

States with the most job openings for Speech Research jobs include:

What other helpful pages are available for Speech Research?

Other pages related to Speech Research:

Infographic showing various Speech Research job openings in the United States as of September 2026, with employment types broken down into 1% Internship, 1% As Needed, 88% Full Time, 9% Part Time, and 1% Contract. Highlights an 79% Physical, 3% Hybrid, and 18% Remote job distribution, with an average salary of $95,216 per year, or $45.8 per hour.

Senior Research Engineer, Audio and Speech

New York, NY β€’ On-site

$200K - $400K/yr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Re-posted 26 days ago


Job description

About Decagon
Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experiences.
Our technology enables industry-defining enterprises like Avis Budget Group, Block's Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents that power personalized, deeply satisfying interactions across voice, chat, email, SMS, and every other channel.
We're building a future where customer experiences are being redefined from support tickets and hold music to faster resolutions, richer conversations, and deeper relationships. We're proud to be backed by world-class investors who share that vision, including a16z, Accel, Bain Capital Ventures, Coatue, and Index Ventures, along with many others.
We're an in-office company, driven by a shared commitment to excellence and velocity. Our values - Just Get It Done, Invent What Customers Want, Winner's Mindset, and The Polymath Principle - shape how we work and grow as a team.
About the Team
Read more about the Speech Research Team's work:
  • https://decagon.ai/blog/audio-native-semantic-speaker-change-detection
  • https://decagon.ai/blog/scaling-real-time-tts-inference
  • https://decagon.ai/blog/teaching-flow-matching-tts-with-rl

The Research team develops the model and decision-making stack that powers Decagon's conversational agents for enterprise support. We research, adapt, and implement state-of-the-art techniques in model training, prompting, orchestration, and evaluation in order to make our agents more accurate, robust, and efficient in real-world deployments.
Our goal is to push the frontier of applied conversational AI: agents that reliably understand nuanced intent, track long context, and take the right actions under uncertainty. We measure success the way customers feel it: higher resolution rates, better user satisfaction, and consistent behavior at scale.
About the Role
As a Research Engineer focused on Audio and Speech, you'll be responsible for building the models and agent harnesses that power Decagon's real-time voice agents and taking them all the way from idea to production. Your work will advance multimodal and full-duplex systems that can listen, reason, speak, and respond naturally in real time.
We're looking for strong engineers who want to build the next generation of AI voice agents. People here own their work end-to-end, ship real improvements, and are trusted to make high-impact technical decisions.
In this role, you will
  • Design and build next-generation agent harnesses optimized for streaming speech, turn-taking, interruptions, overlapping speech, and continuous interaction
  • Research and train multimodal and full-duplex models that jointly understand audio, reason, and generate speech
  • Improve speech recognition, voice activity detection, endpointing, and speech generation across diverse speakers, environments, domains, and languages
  • Build evaluations and use production calls to ship measurable improvements in accuracy, latency, naturalness, and task outcomes
  • Optimize end-to-end inference for responsiveness, throughput, stability, and cost, partnering with Voice Platform and Infrastructure teams to deploy at scale

Your background looks something like this
  • 4+ years of experience in speech, audio ML, multimodal ML, or production machine learning
  • Experience developing or adapting autoregressive, diffusion, flow-matching, or codec-based speech models
  • Hands-on experience with streaming agent systems, low-latency inference, production model serving, and evaluation on real-world audio
  • Fluency in Python and a modern deep-learning framework such as PyTorch, with strong foundations in machine learning and signal processing
  • A track record of taking research ideas from prototype to reliable, measurable production impact

Even better if you have
  • Familiarity with speech-to-speech or full-duplex models
  • Experience with telephony, multilingual speech, noisy-channel robustness, speaker adaptation, or expressive speech generation

Compensation
$200K - $400K + Offers Equity
Benefits
We proudly offer the following benefits for our full-time employees:
  • Medical, Dental, and Vision benefits for you and your family
  • Life Insurance and Disability Benefits
  • Retirement Plan (e.g., 401K, pension)
  • Parental Leave
  • Fertility and family building benefits through Carrot
  • Monthly stipend to support your wellness, lifestyle, and work-life balance
  • Daily lunches and snacks in the office to keep you at your best
  • Take what you need vacation policy (subject to local requirements; UK employees receive 25 days of statutory leave)

These benefits are described in more detail in Decagon's policies, may vary by location, and can change at any time according to applicable compensation and benefits plans.