1

Audio Annotation Jobs in Seattle, WA (NOW HIRING)

Execute Data labelling and annotation tasks across speech and voice datasets. * Work with audio and language data, including transcription, categorization, and tagging. YOU ARE A FIT IF YOU'RE... * A ...

next page

Showing results 1-20

Audio Annotation information

See Seattle, WA salary details

$33.6K

$96.1K

$195.2K

How much do audio annotation jobs pay per year?

As of Sep 9, 2026, the average yearly pay for audio annotation in Seattle, WA is $96,113.00, according to ZipRecruiter salary data. Most workers in this role earn between $56,900.00 and $128,600.00 per year, depending on experience, location, and employer.

What is audio annotation?

Audio annotation is the process of labeling or tagging audio data with relevant information, such as identifying sounds, speech, speakers, or background noises. This process helps train machine learning models to recognize and understand audio content. Audio annotation can involve tasks like transcribing speech, marking segments with specific sounds, or categorizing audio clips by genre or emotion. It is widely used in developing applications for speech recognition, virtual assistants, and audio analysis.

What are the key skills and qualifications needed to thrive as an audio annotator, and why are they important?

To thrive as an Audio Annotator, you need strong attention to detail, excellent listening skills, and familiarity with linguistic concepts, often supported by relevant coursework or experience in linguistics or audio processing. Proficiency in annotation tools such as ELAN, Audacity, or Praat, as well as experience with data labeling platforms, is typically required. Strong organizational skills, patience, and the ability to work independently make someone stand out in this role. These skills ensure accurate and consistent audio data labeling, which is essential for training reliable AI and speech recognition systems.

What are some common challenges faced by audio annotators, and how can they be managed effectively?

Audio annotators often encounter challenges such as distinguishing overlapping voices, dealing with low-quality recordings, and maintaining consistency in labeling. To manage these, it's important to use high-quality headphones, familiarize yourself with annotation guidelines, and communicate regularly with your team to resolve ambiguities. Many organizations also provide regular feedback sessions and quality checks to ensure accuracy and support continuous improvement.

What are popular job titles related to Audio Annotation jobs in Seattle, WA?

For Audio Annotation jobs in Seattle, WA, the most frequently searched job titles are:

What job categories do people searching Audio Annotation jobs in Seattle, WA look for?

The top searched job categories for Audio Annotation jobs in Seattle, WA are:

What cities near Seattle, WA are hiring for Audio Annotation jobs?

Cities near Seattle, WA with the most Audio Annotation job openings:

Infographic showing various Audio Annotation job openings in Seattle, WA as of September 2026, with employment types broken down into 100% Full Time. Highlights an 100% In-person job distribution, with an average salary of $96,113 per year, or $46.2 per hour.

Research Scientist - Speech and Audio Understanding (Large Models & Multimodal Systems)

Bellevue, WA • On-site

$150 - $200/hr

Other

Medical, Dental, Vision, Life, Retirement, PTO

Posted 7 days ago


Job description

What the Role Entails

We are building large‑scale, native multimodal model systems that jointly support vision, audio, and text to enable comprehensive perception and understanding of the physical world.

Job Responsibilities
  • Develop general‑purpose, end‑to‑end large speech models covering multilingual automatic speech recognition (ASR), speech translation, speech synthesis, paralinguistic understanding, and general audio understanding.
  • Advance research on speech representation learning and encoder/decoder architectures to build unified acoustic representations for multi‑task and multimodal applications.
  • Explore representation alignment and fusion mechanisms between audio/speech and other modalities in large multimodal models, enabling joint modeling with image and text.
  • Build and maintain high‑quality multimodal speech datasets, including automatic annotation and data synthesis technologies.
Who We Look For
  • Ph.D. in Computer Science, Electrical Engineering, Artificial Intelligence, Linguistics, or a related field; or Master’s degree with several years of relevant experience.
  • Solid understanding of speech and audio signal processing, acoustic modeling, language modeling, and large model architectures.
  • Proficient in one or more core speech system development pipelines such as ASR, TTS, or speech translation; experience with multilingual, multitask, or end‑to‑end systems is a plus.
  • Experience with large‑scale training and distributed systems is a plus.
  • Familiar with Transformer‑based architectures and their applications in speech and multimodal training/inference.
Preferred Expertise
  • Speech representation pretraining (e.g., HuBERT, Wav2Vec, Whisper).
  • Multimodal alignment and cross‑modal modeling (e.g., audio‑visual‑text).
  • Experience driving state‑of‑the‑art (SOTA) performance on audio understanding tasks with large models.
  • Proficient in deep learning frameworks such as PyTorch or TensorFlow.
Location

State(s): US-Washington-Bellevue

Compensation

The expected base pay range for this position is $122,500.00 – $229,700.00 per year. Actual pay may vary depending on job‑related knowledge, skills, and experience.

Benefits

Employees may be eligible for a sign‑on payment, relocation package, restricted stock units, medical, dental, vision, life and disability benefits, and participation in the company’s 401(k) plan.

Vacation: up to 15 to 25 days per year (depending on tenure). Holidays: up to 13 days per year. Paid sick leave: up to 10 days per year.

Equal Employment Opportunity

We are an equal‑opportunity employer. We firmly believe that diverse voices fuel our innovation and allow us to better serve our users and the community. Every employee of Tencent feels supported and inspired to achieve individual and common goals.

#J-18808-Ljbffr