1

Assistant Voice Research Task Jobs in California

Research Engineer (Voice / Speech AI) 📍 San Francisco, CA (Onsite) I am seeking a Research ... voice systems. Rather than building traditional AI assistants, this team is creating full-duplex AI ...

Voice Designer

Fremont, CA · On-site

$45 - $50/hr

... voice research scientists, software engineers, and product managers Responsible for defining ... daily tasks and troubleshooting Strong understanding of DSP, audio synthesis, and digital audio ...

... research and outreach. Building 0→1 has never been easier, but selling at the speed of the market ... The ones that do go through a maze: voicemails, phone trees, virtual assistants. Voice AI is the ...

Position: Human Baseliner for Open-Ended ML Research Tasks Type: Contract Compensation: $75-$90 ... Use preferred tools, including IDEs and AI coding assistants like Cursor , Claude Code , and ...

The Research Assistant Intern will support the ESSC Research Division at Easterseals by performing ... Interns may conduct behavior intervention tasks for research purposes after completing the initial ...

next page

Showing results 1-20

Assistant Voice Research Task information

What are Assistant Voice Research Tasks?

Assistant Voice Research Tasks involve collecting, analyzing, and improving voice data to enhance the performance of voice assistants like Siri, Alexa, or Google Assistant. People in these roles may transcribe voice recordings, annotate speech data, evaluate the quality of voice responses, or test new voice recognition features. The goal is to help make voice assistants more accurate, natural, and helpful for users by identifying errors or areas for improvement. These tasks are typically carried out by researchers, linguists, or specialized annotators working for technology companies.

What are the key skills and qualifications needed to thrive as a Voice Assistant Researcher, and why are they important?

To thrive as a Voice Assistant Researcher, you need expertise in computational linguistics, machine learning, natural language processing (NLP), and usually a degree in computer science, linguistics, or a related field. Familiarity with programming languages like Python, tools such as TensorFlow or PyTorch, and experience with speech recognition systems are typically required. Strong analytical thinking, creativity, and effective collaboration skills help you innovate and work well within multidisciplinary research teams. These competencies are crucial for developing accurate, user-friendly voice technologies that meet evolving user needs.

What are some common challenges faced by professionals working in Assistant Voice Research tasks, and how can they be addressed?

Professionals in Assistant Voice Research often encounter challenges such as ensuring high-quality voice recognition across diverse accents, dialects, and noisy environments. They may also need to handle large datasets and continually update models to improve accuracy and user experience. Collaborating closely with linguists, data scientists, and software engineers is crucial, as is staying current with advancements in AI and natural language processing. Addressing these challenges involves ongoing testing, user feedback collection, and leveraging cutting-edge research to refine voice assistant capabilities.

What is the difference between Assistant Voice Research Task vs Voice Data Annotator?

AspectAssistant Voice Research TaskVoice Data Annotator
Required CredentialsBasic understanding of linguistics, speech technologyNone or minimal; training provided
Work EnvironmentResearch labs, tech companies, remoteData labeling centers, remote or onsite
Employer & IndustryTech companies, AI development, speech recognitionData annotation firms, AI companies
Common Search & ComparisonYesYes

The Assistant Voice Research Task involves supporting speech research with a focus on linguistic and technical understanding, often in a research or development setting. Voice Data Annotators primarily focus on labeling and preparing audio data for machine learning models. While both roles support speech technology, the Assistant Voice Research Task typically requires some foundational knowledge, whereas Voice Data Annotators usually need minimal prior experience.

What are the most commonly searched types of Voice Research Task jobs in California? The most popular types of Voice Research Task jobs in California are:
What cities in California are hiring for Assistant Voice Research Task jobs? Cities in California with the most Assistant Voice Research Task job openings:

Research Engineer

Acceler8 Talent

San Francisco, CA • On-site

Other

This job post has expired today. Applications are no longer accepted.


Job description

Research Engineer (Voice / Speech AI)


📍 San Francisco, CA (Onsite)


I am seeking a Research Engineer to join a frontier AI research lab building the next generation of real-time conversational voice systems. Rather than building traditional AI assistants, this team is creating full-duplex AI companions that can naturally participate in live multiplayer experiences: listening, responding, interrupting appropriately, understanding shared context, and interacting alongside human players in real time.


Alongside building production voice agents, the team also operates as a voice research lab, producing research-grade speech datasets, evaluations, and tooling that support frontier AI research across the wider ecosystem.


What We're Looking For:


  • Experience researching speech foundation models or voice AI systems
  • Hands-on experience training, fine-tuning, evaluating, and benchmarking speech models
  • Experience working on streaming or low-latency speech systems
  • Exposure to full-duplex conversational models such as Moshi, PersonaPlex, or similar research is highly desirable
  • Publications at leading speech, audio, or machine learning conferences (ICASSP, Interspeech, NeurIPS, ICML, ICLR, etc.) are a significant advantage
  • Strong machine learning research background with the ability to translate research into production systems


What You'll Do:


  • Research and prototype ultra-low latency streaming and full-duplex speech systems
  • Develop speech foundation models, including cascaded (ASR → LLM → TTS) and speech-to-speech architectures
  • Improve conversational behaviours including turn-taking, interruption handling, timing, prosody, and emotion
  • Train, fine-tune, benchmark, and evaluate state-of-the-art speech models
  • Design evaluation frameworks that measure conversational quality and social presence—not just transcription accuracy
  • Work across model research, inference optimisation, data generation, and production deployment
  • Collaborate closely with researchers building frontier multimodal and voice foundation models


This is a genuinely research-focused opportunity for someone who wants to push the boundaries of speech AI. You'll be tackling challenges around conversational timing, streaming inference, interruption handling, latency optimisation, and voice quality—building systems that feel socially present rather than simply responsive.


Based in downtown San Francisco, the team works fully onsite and offers highly competitive compensation alongside meaningful equity as they continue to build one of the most ambitious voice AI research groups in the industry.


If you'd like to find out more, apply today!


Ethan Lewis

elewis@acceler8talent.com