1

Speech Technology Jobs (NOW HIRING)

Lead development efforts and work effectively with a small team of developers to create, improve and maintain applications that support state-of-the-art speech technology. * Full life cycle ownership ...

Showing results 41-60

Speech Technology information

See salary details

$15

$43

$69

How much do speech technology jobs pay per hour?

As of Aug 10, 2026, the average hourly pay for speech technology in the United States is $43.92, according to ZipRecruiter salary data. Most workers in this role earn between $36.06 and $51.68 per hour, depending on experience, location, and employer.

What does a speech technology professional do?

In speech technology roles, professionals commonly work on projects such as developing and improving speech recognition systems, designing voice user interfaces, and implementing natural language processing solutions for products like virtual assistants or automated transcription tools. Day-to-day responsibilities often involve collaborating with software engineers, data scientists, and UX designers, as well as testing and optimizing algorithms using large voice datasets. You may also participate in troubleshooting, model evaluation, and refining features to meet client or product needs. This collaborative, innovative work environment offers an excellent opportunity to contribute to cutting-edge advancements in human-computer interaction.

What is a speech technology?

A Speech Technology job involves developing and improving systems that enable machines to process, recognize, and synthesize human speech. Professionals in this field work with technologies like automatic speech recognition (ASR), text-to-speech (TTS), and natural language processing (NLP) to enhance voice-based applications. They collaborate with engineers, linguists, and data scientists to create efficient and accurate speech-enabled systems for industries such as AI assistants, customer service automation, and accessibility solutions.

What are the key skills and qualifications needed to thrive in speech technology?

To excel in Speech Technology, a strong background in computer science, linguistics, and machine learning—often supported by a relevant degree—is essential. Familiarity with speech recognition frameworks, natural language processing (NLP) libraries, and programming languages like Python or C++ is typically required. Strong analytical thinking, teamwork, and clear communication skills help differentiate top professionals in this field. These abilities enable the development and refinement of advanced speech systems that work accurately and efficiently for end-users.

What cities are hiring for Speech Technology jobs? Cities with the most Speech Technology job openings:
What are the most commonly searched types of Speech Technology jobs? The most popular types of Speech Technology jobs are:
What states have the most Speech Technology jobs? States with the most job openings for Speech Technology jobs include:
Infographic showing various Speech Technology job openings in the United States as of August 2026, with employment types broken down into 72% Full Time, and 28% Part Time. Highlights an 89% In-person, and 11% Remote job distribution, with an average salary of $91,346 per year, or $43.9 per hour.

Machine Learning Architect - Conversational Speech

Apple

Cupertino, CA

$262K - $394K/yr

Full-time

Medical, Dental, Retirement

Re-posted 14 days ago


Apple rating

8.0

Company rating: 8.0 out of 10

Based on 677 frontline employees who took The Breakroom Quiz

7th of 30 rated technology retailers


Job description

Join the team redefining what a deeply personal and integrated assistant can be.
As part of the Siri organization, you will help shape one of the world's most widely used AI assistants, powered by our next-generation of Apple Intelligence, with capabilities like personal context understanding and on-screen awareness, built with privacy from the ground up. Your work will have direct, meaningful impact for users across iOS, iPadOS, macOS, watchOS, and visionOS.
This is a rare opportunity to build at the intersection of cutting-edge AI and human-centered design, shipping technology that is centered around users and their needs.
Description
The Speech organization within Siri drives major speech recognition, synthesis, and speech-to-speech model advances for features deeply embedded throughout Apple's ecosystem. Our mission is to build cutting-edge infrastructure, datasets, and models that empower Siri conversational AI, dictation, and speech-enabled Apple Intelligence features across natural language understanding, dialog generation, speech recognition, and multimodal interaction. We apply these technologies to create engaging, intelligent, and personalized conversational experiences for millions of Apple users.
We are seeking a Machine Learning Architect to serve as a senior technical leader spanning the full Speech organization. You will set the future modeling direction for all of conversational speech-charting the architectural and algorithmic course for how Apple's speech technologies evolve. You will operate as a hands-on expert who not only defines strategy but also digs into the hardest technical problems, working shoulder-to-shoulder with teams to overcome critical obstacles. Reporting directly to the Speech organization leadership, you will have broad visibility and influence across speech recognition, synthesis, dialog, multimodal foundation models, and speech-to-speech systems, ensuring coherent technical vision and cross-team alignment.
","responsibilities":"As the Machine Learning Architect for Conversational Speech, you will define modeling strategy and technical direction across the Speech organization, establishing a unified architectural vision for speech recognition, speech synthesis, dialog systems, multimodal foundation models, and speech-to-speech technologies.
You will serve as the organization's foremost modeling expert, providing deep technical guidance to multiple teams working on interconnected speech capabilities.
You will evaluate emerging research and industry trends-including advances in large language models, multimodal architectures, and full-duplex natural conversational systems-and translate them into actionable roadmaps.
You will champion production-readiness, ensuring architectural decisions account for on-device constraints, latency, scalability, and robustness.
You will collaborate broadly with partner teams across Siri, Apple Intelligence, hardware, and platform engineering to ensure speech modeling investments are well-integrated into Apple's broader AI strategy.
Preferred Qualifications
Ph.D. in Computer Science, Electrical Engineering, Machine Learning, or similar technical field.
Experience architecting or leading development of full-duplex natural conversational systems, speech-to-speech models, or multimodal foundation models that have shipped to large-scale user populations.
Deep familiarity with the full stack of speech technologies-ASR, TTS, spoken dialog, speaker modeling, audio understanding-and an ability to reason about their interactions and dependencies.
Experience with large-scale distributed training and the infrastructure considerations that shape model design at scale.
A data-centric perspective on foundation model development, including experience guiding data collection, curation, annotation, and quality strategies.
Experience with on-device ML deployment, including model compression, quantization, and latency-aware architecture design.
Minimum Qualifications
10+ years of experience in machine learning applied to speech or multimodal systems, with progressively increasing technical scope and leadership.
Demonstrated expertise as a technical leader or architect who has defined modeling direction across multiple teams or product areas.
Deep, hands-on proficiency in modern deep learning, including large language models and end-to-end speech systems.
Significant experience with multimodal LLMs, including architecture design, training, adaptation, and deployment of models that integrate speech, audio, and text modalities.
Direct experience building speech-to-speech conversational systems, with a strong understanding of full-duplex natural conversational interaction and end-to-end speech pipelines.
A track record of translating research into production-quality systems at scale.
Expert programming skills in Python and deep learning frameworks such as PyTorch, JAX, or TensorFlow.
Pay & Benefits
At Apple, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $262,500 and $394,000, and your base pay will depend on your skills, qualifications, experience, and location.
Apple employees also have the opportunity to become an Apple shareholder through participation in Apple's discretionary employee stock programs. Apple employees are eligible for discretionary restricted stock unit awards, and can purchase Apple stock at a discount if voluntarily participating in Apple's Employee Stock Purchase Plan. You'll also receive benefits including: Comprehensive medical and dental coverage, retirement benefits, a range of discounted products and free services, and for formal education related to advancing your career at Apple, reimbursement for certain educational expenses - including tuition. Additionally, this role might be eligible for discretionary bonuses or commission payments as well as relocation. Learn more about Apple Benefits
Note: Apple benefit, compensation and employee stock programs are subject to eligibility requirements and other terms of the applicable plan or program.

What Apple employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Apple logo

About Apple

Sourced by ZipRecruiter

Imagine what you could do here! At Apple, new ideas have a way of becoming extraordinary products, services, and customer experiences very quickly. Bring passion and dedication to your job and there's no telling what you could accomplish. Dynamic, intelligent people and inspiring, innovative technologies are the norm here. The people who work here have reinvented entire industries with all Apple Hardware products. The same real passion for innovation that goes into our products also applies to our practices strengthening our dedication to leave the world better than we found it.

Industry

Computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

Cupertino, CA, US

Year founded

1976