We are searching for a Speech Processing ML Algorithm Engineer to advance our work in the domain of speech enhancement, with an emphasis on low-latency processing applications. Description As a ...
We are searching for a Speech Processing ML Algorithm Engineer to advance our work in the domain of speech enhancement, with an emphasis on low-latency processing applications. Description As a ...
Speech Processing ML Algorithm Engineer
$150K - $277K/yr
We are searching for a Speech Processing ML Algorithm Engineer to advance our work in the domain of speech enhancement, with an emphasis on low-latency processing applications. Description As a ...
Speech Processing ML Algorithm Engineer
$150K - $277K/yr
We are searching for a Speech Processing ML Algorithm Engineer to advance our work in the domain of speech enhancement, with an emphasis on low-latency processing applications. Description As a ...
Speech Recognition Engineer
Manhattan, NY · On-site
The ideal candidate will have expertise in speech processing, deep learning, natural language processing (NLP), and machine learning. This role involves building, training, fine-tuning, and deploying ...
Speech Recognition Engineer
Manhattan, NY · On-site
The ideal candidate will have expertise in speech processing, deep learning, natural language processing (NLP), and machine learning. This role involves building, training, fine-tuning, and deploying ...
Sales Engineer | North America
Atlanta, GA · On-site +1
At Speech Processing Solutions, we build speech recognition and voice productivity solutions that help professionals in the healthcare, legal, and other professional sectors to work smarter and more ...
Sales Engineer | North America
Atlanta, GA · On-site +1
At Speech Processing Solutions, we build speech recognition and voice productivity solutions that help professionals in the healthcare, legal, and other professional sectors to work smarter and more ...
At Speech Processing Solutions, we build speech recognition and voice productivity solutions that help professionals in the healthcare, legal, and other professional sectors to work smarter and more ...
Quick apply
At Speech Processing Solutions, we build speech recognition and voice productivity solutions that help professionals in the healthcare, legal, and other professional sectors to work smarter and more ...
Senior Software Engineer - Audio (675080)
Cambridge, MA · On-site
$135K - $177K/yr
Responsibilities : • Develop speech processing algorithms, e.g. voice activity, vocal affect, laughter detection • Research, evaluate, and integrate off-the-shelf solutions for required audio ...
Senior Software Engineer - Audio (675080)
Cambridge, MA · On-site
$135K - $177K/yr
Responsibilities : • Develop speech processing algorithms, e.g. voice activity, vocal affect, laughter detection • Research, evaluate, and integrate off-the-shelf solutions for required audio ...
Citizenship) Preferred : • Knowledge of deep learning frameworks • Familiarity with transformer models • Experience with speech processing • Proficiency in sentiment analysis • ...
Citizenship) Preferred : • Knowledge of deep learning frameworks • Familiarity with transformer models • Experience with speech processing • Proficiency in sentiment analysis • ...
S. Citizenship Preferred : • Knowledge of deep learning frameworks • Familiarity with transformer models • Experience with speech processing • Proficiency in sentiment analysis • ...
S. Citizenship Preferred : • Knowledge of deep learning frameworks • Familiarity with transformer models • Experience with speech processing • Proficiency in sentiment analysis • ...
We are seeking individuals passionate in areas such as Natural Language Processing, Audio and Speech processing, Computer Vision, Machine Learning, Deep Learning, and Reinforcement Learning. Our ...
We are seeking individuals passionate in areas such as Natural Language Processing, Audio and Speech processing, Computer Vision, Machine Learning, Deep Learning, and Reinforcement Learning. Our ...
Senior Machine Learning Engineer (Remote)
New York, NY · On-site +1
$114K - $157K/yr
Experience with modern speech processing frameworks * Good dev ops experience * Advanced DSP experience * MongoDB or SQL experience * Experience with unit, integration, and load testing * Experience ...
Senior Machine Learning Engineer (Remote)
New York, NY · On-site +1
$114K - $157K/yr
Experience with modern speech processing frameworks * Good dev ops experience * Advanced DSP experience * MongoDB or SQL experience * Experience with unit, integration, and load testing * Experience ...
We are seeking individuals passionate in areas such as Natural Language Processing, Audio and Speech processing, Computer Vision, Machine Learning, Deep Learning, and Reinforcement Learning. Our ...
We are seeking individuals passionate in areas such as Natural Language Processing, Audio and Speech processing, Computer Vision, Machine Learning, Deep Learning, and Reinforcement Learning. Our ...
You will contribute to a broad set of technologies including audio signal processing, speech and voice processing, noise reduction, and microphone array systems , and gain experience adapting ...
You will contribute to a broad set of technologies including audio signal processing, speech and voice processing, noise reduction, and microphone array systems , and gain experience adapting ...
You will contribute to a broad set of technologies including audio signal processing, speech and voice processing, noise reduction, and microphone array systems , and gain experience adapting ...
You will contribute to a broad set of technologies including audio signal processing, speech and voice processing, noise reduction, and microphone array systems , and gain experience adapting ...
This hands-on role owns audio processing, dataset curation, annotation and QA workflows, model training, and evaluation for our multilingual speech, translation, and conversational AI systems. Key ...
This hands-on role owns audio processing, dataset curation, annotation and QA workflows, model training, and evaluation for our multilingual speech, translation, and conversational AI systems. Key ...
Speech Science Technology Manager
Richardson, TX · On-site
$130K - $178K/yr
Architect and lead development of real-time speech processing pipeline: Voice Command detection, VAD (voice activity detection), ASR, TTS, noise suppression, echo cancellation, and keyword spotting.
Speech Science Technology Manager
Richardson, TX · On-site
$130K - $178K/yr
Architect and lead development of real-time speech processing pipeline: Voice Command detection, VAD (voice activity detection), ASR, TTS, noise suppression, echo cancellation, and keyword spotting.
Speech Science Technology Manager
Richardson, TX · On-site
$130K - $178K/yr
Architect and lead development of real-time speech processing pipeline: Voice Command detection, VAD (voice activity detection), ASR, TTS, noise suppression, echo cancellation, and keyword spotting.
Speech Science Technology Manager
Richardson, TX · On-site
$130K - $178K/yr
Architect and lead development of real-time speech processing pipeline: Voice Command detection, VAD (voice activity detection), ASR, TTS, noise suppression, echo cancellation, and keyword spotting.
$180K - $270K/yr
This team's responsibilities cover low-level media processing pipelines, integration with internal and external models for speech recognition and synthesis, and globally distributed, WebRTC-based ...
$180K - $270K/yr
This team's responsibilities cover low-level media processing pipelines, integration with internal and external models for speech recognition and synthesis, and globally distributed, WebRTC-based ...
Media Software Engineer, Speech (Senior-Staff Levels)
Sunnyvale, CA · On-site
$180K - $270K/yr
This team's responsibilities cover low-level media processing pipelines, integration with internal and external models for speech recognition and synthesis, and globally distributed, WebRTC-based ...
Media Software Engineer, Speech (Senior-Staff Levels)
Sunnyvale, CA · On-site
$180K - $270K/yr
This team's responsibilities cover low-level media processing pipelines, integration with internal and external models for speech recognition and synthesis, and globally distributed, WebRTC-based ...
Sr. Backend Engineer - Engine Team: API
$150K - $220K/yr
You will design and implement secure, robust, and scalable services for speech processing; efficient, distributed compute orchestration; optimized scheduling, and more. Your skill at building highly ...
Sr. Backend Engineer - Engine Team: API
$150K - $220K/yr
You will design and implement secure, robust, and scalable services for speech processing; efficient, distributed compute orchestration; optimized scheduling, and more. Your skill at building highly ...
Senior Machine Learning Engineer, Speech & LLM Training Data
Overland Park, KS · On-site
$111K - $133K/yr
This hands-on role owns audio processing, dataset curation, annotation and QA workflows, model training, and evaluation for our multilingual speech, translation, and conversational AI systems. Key ...
Senior Machine Learning Engineer, Speech & LLM Training Data
Overland Park, KS · On-site
$111K - $133K/yr
This hands-on role owns audio processing, dataset curation, annotation and QA workflows, model training, and evaluation for our multilingual speech, translation, and conversational AI systems. Key ...
Speech Processing information
See salary details
$19.95 - $23.36
2% of jobs
$23.36 - $26.77
2% of jobs
$26.77 - $30.18
5% of jobs
$30.18 - $33.59
8% of jobs
$35.23 is the 25th percentile. Wages below this are outliers.
$33.59 - $37
15% of jobs
$37 - $40.41
16% of jobs
The median wage is $40.69 / hr.
$40.41 - $43.82
19% of jobs
$45.58 is the 75th percentile. Wages above this are outliers.
$43.82 - $47.22
15% of jobs
$47.22 - $50.63
9% of jobs
$50.63 - $54.04
6% of jobs
$54.04 - $57.45
2% of jobs
$19
$41
$57
How much do speech processing jobs pay per hour?
What is speech processing?
What are the key skills and qualifications needed to thrive as a speech processing engineer, and why are they important?
What are some common challenges faced by professionals in speech processing roles, and how can they be addressed?
What is the difference between Speech Processing vs Speech Recognition?
| Aspect | Speech Processing | Speech Recognition |
|---|---|---|
| Definition | Broad field involving analysis, modification, and synthesis of speech signals | Subfield focused on converting spoken language into text |
| Skills & Certifications | Signal processing, audio engineering, programming; certifications like DSP or audio engineering | Machine learning, NLP, programming; certifications in AI or speech technology |
| Work Environment | Research labs, tech companies, audio hardware firms | Software development, AI companies, voice assistant firms |
| Industry Usage | Telecommunications, audio processing, speech synthesis | Virtual assistants, transcription services, voice command systems |
Speech Processing is a broad field encompassing various aspects of speech signal analysis and synthesis, while Speech Recognition specifically focuses on converting spoken words into written text. Both roles often require similar technical skills and certifications, but their applications differ across industries and job functions.
What cities are hiring for Speech Processing jobs?
Cities with the most Speech Processing job openings:
What states have the most Speech Processing jobs?
States with the most job openings for Speech Processing jobs include:
What job categories do people searching Speech Processing jobs look for?
The top searched job categories for Speech Processing jobs are:
What other helpful pages are available for Speech Processing?
Other pages related to Speech Processing:

Speech Processing ML Algorithm Engineer
Cupertino, CA • On-site
Full-time
Posted 15 days ago
Apple rating
8.1
Based on 684 frontline employees who took The Breakroom Quiz
Job description
The Acoustic ML Algorithm Development team sits in Apple's Hardware organization and develops the new architectures and training paradigms that define the audio experience of the next generation of Apple hardware along with our counterparts in the Software organization. We also work closely with various teams behind microphones, loudspeakers, wireless communication, silicon, etc and we focus on features that improve the customer experience across speech, music, and general audio. We are searching for a Speech Processing ML Algorithm Engineer to advance our work in the domain of speech enhancement, with an emphasis on low-latency processing applications.
Description
As a Speech Processing ML Algorithm Engineer on the Acoustics ML Algorithm Development team, you will design, train, and evaluate models for speech enhancement. Examples include noise suppression, de-reverberation, echo suppression, and general multi-microphone processing under the low-latency, real-time constraints of shipping hardware. You will work with cross functional teams to bring early prototypes through to production.
Minimum Qualifications
MS or PhD in a computational science field or 3+ years of experience in the field of ML for speech processing.
A passion for audio ML research for applied product applications, with a deep understanding of transformers, recurrent and convolutional neural networks, etc.
Experience designing, training, and evaluating machine learning models for speech enhancement, such as noise suppression, dereverberation, source separation, or echo residual suppression.
Experience building models that meet low-latency, real-time requirements, including streaming and causal processing, and a clear understanding of the quality, complexity, and latency trade-offs involved.
Strong grounding in audio and speech signal processing fundamentals, for example STFT analysis/synthesis.
Proficiency with PyTorch and Bash, including version control, code review, testing, and reproducible experiments.
A habit of following the state of the art literature closely, with the ability to reproduce, critique, and build on published results.
Working knowledge of speech quality evaluation, spanning objective metrics (e.g. PESQ, STOI, SI-SDR, DNSMOS) and subjective listening tests, along with the data simulation and augmentation needed to support them.
Preferred Qualifications
Experience applying machine learning to adaptive filter prediction and control, such as echo cancellation, active noise control, or adaptive beamforming, including hybrid classical and learned systems.
Familiarity with Lightning and Hydra.
Familiarity with model efficiency techniques such as quantization-aware training or distillation.
Experience in high-performance cloud computing for model training.
Open source contributions to a repository using in the ML audio community.
About Apple
Sourced by ZipRecruiter
Imagine what you could do here! At Apple, new ideas have a way of becoming extraordinary products, services, and customer experiences very quickly. Bring passion and dedication to your job and there's no telling what you could accomplish. Dynamic, intelligent people and inspiring, innovative technologies are the norm here. The people who work here have reinvented entire industries with all Apple Hardware products. The same real passion for innovation that goes into our products also applies to our practices strengthening our dedication to leave the world better than we found it.
Industry
Computer and electronic product manufacturing
Company size
10,000+ Employees
Headquarters location
Cupertino, CA, US
Year founded
1976