Audio AI Engineer
$80K - $160K/yr
Audio AI Engineer, #1085 Multilingual Speech-to-Text Engineer - On-Device Model Optimization, #1085 A Role with Purpose and Impact This role builds the speech recognition core of a mobile translation ...
$80K - $160K/yr
Audio AI Engineer, #1085 Multilingual Speech-to-Text Engineer - On-Device Model Optimization, #1085 A Role with Purpose and Impact This role builds the speech recognition core of a mobile translation ...
$80K - $160K/yr
Audio AI Engineer, #1085 Multilingual Speech-to-Text Engineer - On-Device Model Optimization, #1085 A Role with Purpose and Impact This role builds the speech recognition core of a mobile translation ...
$80K - $160K/yr
Audio AI Engineer, #1085 Multilingual Speech-to-Text Engineer -- On-Device Model Optimization, #1085 A Role with Purpose and Impact This role builds the speech recognition core of a mobile ...
Quick apply
$80K - $160K/yr
Audio AI Engineer, #1085 Multilingual Speech-to-Text Engineer -- On-Device Model Optimization, #1085 A Role with Purpose and Impact This role builds the speech recognition core of a mobile ...
Reston, VA · On-site
$80K - $160K/yr
... Audio AI Engineer, #1085 Multilingual Speech-to-Text Engineer - On-Device Model Optimization, #1085 A Role with Purpose and Impact This role builds the speech recognition core of a mobile translation ...
Reston, VA · On-site
$80K - $160K/yr
... Audio AI Engineer, #1085 Multilingual Speech-to-Text Engineer - On-Device Model Optimization, #1085 A Role with Purpose and Impact This role builds the speech recognition core of a mobile translation ...
Arlington, VA · On-site
Use AI-assisted audio tools where appropriate, including noise reduction, stem separation, automated leveling, voice cleanup, transcription, and related production workflows. * Stay current on audio ...
Arlington, VA · On-site
Use AI-assisted audio tools where appropriate, including noise reduction, stem separation, automated leveling, voice cleanup, transcription, and related production workflows. * Stay current on audio ...
Integrates AI-assisted audio tools such as noise reduction, stem separation, automated leveling, transcription, translation workflow - into the production pipeline Host and Contributor Management
Integrates AI-assisted audio tools such as noise reduction, stem separation, automated leveling, transcription, translation workflow - into the production pipeline Host and Contributor Management
Develop, script, and produce high-quality multimedia content (audio/video/still photography) for ... Generative AI Integration * Use Generative AI platforms for video and animation production for full ...
Develop, script, and produce high-quality multimedia content (audio/video/still photography) for ... Generative AI Integration * Use Generative AI platforms for video and animation production for full ...
Develop, script, and produce high-quality multimedia content (audio/video/still photography) for ... Utilize AI to enhance workflows and train on latest production innovations to develop editing ...
Develop, script, and produce high-quality multimedia content (audio/video/still photography) for ... Utilize AI to enhance workflows and train on latest production innovations to develop editing ...
Develop, script, and produce high-quality multimedia content (audio/video/still photography) for ... Generative AI Integration * Use Generative AI platforms for video and animation production for full ...
Develop, script, and produce high-quality multimedia content (audio/video/still photography) for ... Generative AI Integration * Use Generative AI platforms for video and animation production for full ...
Richmond, VA · On-site
$150 - $230/hr
Applied AI (Voice Agents & ML Systems) The pitch We build and operate production AI voice agents ... Audio handling and the quirks of real human conversation (interruptions, timing, noise)
Richmond, VA · On-site
$150 - $230/hr
Applied AI (Voice Agents & ML Systems) The pitch We build and operate production AI voice agents ... Audio handling and the quirks of real human conversation (interruptions, timing, noise)
Overview Empower AI is AI for government. Empower AI gives federal agency leaders the tools to ... You will oversee audio/visual (A/V) infrastructure, conference room sustainment, video ...
Overview Empower AI is AI for government. Empower AI gives federal agency leaders the tools to ... You will oversee audio/visual (A/V) infrastructure, conference room sustainment, video ...
Richmond, VA · Remote
$121K - $160K/yr
Applied AI (Voice Agents & ML Systems) AMC Health · Remote (US) · Full-time The pitch We build ... Audio handling and the quirks of real human conversation (interruptions, timing, noise)
Richmond, VA · Remote
$121K - $160K/yr
Applied AI (Voice Agents & ML Systems) AMC Health · Remote (US) · Full-time The pitch We build ... Audio handling and the quirks of real human conversation (interruptions, timing, noise)
Overview Empower AI is AI for government. Empower AI gives federal agency leaders the tools to ... You will oversee audio/visual (A/V) infrastructure, conference room sustainment, video ...
Overview Empower AI is AI for government. Empower AI gives federal agency leaders the tools to ... You will oversee audio/visual (A/V) infrastructure, conference room sustainment, video ...
Overview Empower AI is AI for government. Empower AI gives federal agency leaders the tools to ... You will oversee audio/visual (A/V) infrastructure, conference room sustainment, video ...
Overview Empower AI is AI for government. Empower AI gives federal agency leaders the tools to ... You will oversee audio/visual (A/V) infrastructure, conference room sustainment, video ...
Overview Empower AI is AI for government. Empower AI gives federal agency leaders the tools to ... You will oversee audio/visual (A/V) infrastructure, conference room sustainment, video ...
Overview Empower AI is AI for government. Empower AI gives federal agency leaders the tools to ... You will oversee audio/visual (A/V) infrastructure, conference room sustainment, video ...
Reston, VA · On-site
$175K - $220K/yr
Our model enables our customers to perform critical data labeling tasks over multi-media sources such as video, images, audio, and documents. Our AI/ML Engineer is the core mission specialist who ...
Reston, VA · On-site
$175K - $220K/yr
Our model enables our customers to perform critical data labeling tasks over multi-media sources such as video, images, audio, and documents. Our AI/ML Engineer is the core mission specialist who ...
Carnegie Mellon University's Software Engineering Institute is seeking an AI Security Software ... autonomy, audio analysis) • Experience and knowledge in cybersecurity best practices • ...
Carnegie Mellon University's Software Engineering Institute is seeking an AI Security Software ... autonomy, audio analysis) • Experience and knowledge in cybersecurity best practices • ...
Carnegie Mellon University is seeking an AI Security Software Engineer within the CERT Division of ... autonomy, audio analysis) • Experience and knowledge in cybersecurity best practices • ...
Carnegie Mellon University is seeking an AI Security Software Engineer within the CERT Division of ... autonomy, audio analysis) • Experience and knowledge in cybersecurity best practices • ...
Arlington, VA · On-site
$131K - $180K/yr
Carnegie Mellon University is seeking a Senior AI Security Software Engineer to join the CERT ... autonomy, audio analysis) • Experience and knowledge in cybersecurity best practices • ...
Arlington, VA · On-site
$131K - $180K/yr
Carnegie Mellon University is seeking a Senior AI Security Software Engineer to join the CERT ... autonomy, audio analysis) • Experience and knowledge in cybersecurity best practices • ...
Arlington, VA · On-site +1
As AI becomes central to critical infrastructure, advancing its security and resilience offers a ... audio analysis) * Experience and knowledge in cybersecurity best practices * Demonstrated ability ...
Arlington, VA · On-site +1
As AI becomes central to critical infrastructure, advancing its security and resilience offers a ... audio analysis) * Experience and knowledge in cybersecurity best practices * Demonstrated ability ...
Falls Church, VA · On-site
$20 - $24/hr
Geospatial AI Data Annotator About Enabled Intelligence, Inc. Enabled Intelligence, Inc. provides ... video, audio files, or unstructured text (EO, IR, SAR, FMV) * Prior experience with business ...
Falls Church, VA · On-site
$20 - $24/hr
Geospatial AI Data Annotator About Enabled Intelligence, Inc. Enabled Intelligence, Inc. provides ... video, audio files, or unstructured text (EO, IR, SAR, FMV) * Prior experience with business ...
| Aspect | Ai Audio | Voiceover Artist |
|---|---|---|
| Required Credentials | Technical skills, AI and audio editing knowledge | Voice training, acting skills, demo reels |
| Work Environment | Digital, remote, tech-focused | Recording studios, remote, live performances |
| Industry Usage | Media production, AI development, tech companies | Advertising, entertainment, media |
| Search & Comparison Intent | Technical, AI-driven audio solutions | Creative voice work, acting |
Ai Audio involves creating audio content using artificial intelligence technology, focusing on automation and digital tools. Voiceover Artists provide human voice recordings for various media, emphasizing performance and acting skills. While Ai Audio is tech-based and automated, Voiceover Artists rely on vocal talent and creativity. Both roles are essential in media production but serve different purposes and skill sets.

$80K - $160K/yr
Full-time
Posted 5 days ago
Multilingual Speech-to-Text Engineer - On-Device Model Optimization, #1085
A Role with Purpose and Impact
This role builds the speech recognition core of a mobile translation capability supporting a government agency's national security mission. The engineer will take large, high-quality speech-to-text models spanning many language families and adapt, compress, and optimize them so they run performantly on an iPhone - including handling the reality that speakers frequently mix in borrowed English terms mid-utterance, and the model needs to make a sound call on whether to transcribe those terms in English or in the source language's own transliteration.
This is an applied ML role, not a research-only position. The strongest candidate can move fluidly from raw audio data, to model adaptation and compression experiments, to a rigorous evaluation framework - and can clearly explain what they're building, why it's better than the status quo, and how they'll know it worked.
What This Role Is (and Isn't)
This position owns the speech-to-text model - its data, its training/adaptation, its size and latency on-device, and its accuracy across languages. It does not own iOS application development, translation (source-language-to-target-language), or the Swift/AVFoundation integration layer; those are handled by a separate mobile engineering function this role will collaborate closely with.
Key Responsibilities
Required Qualifications
Preferred (Not Required)
The estimated salary range for this position is $80,000 - $160,000. This salary range is not a guarantee of compensation. The offered salary will be based on factors including relevant experience, geographic location, internal equity, and applicable contractual requirements. *Compensation may fall outside this range when appropriate.