1

Text To Speech Jobs in Oregon (NOW HIRING)

Senior Software Engineer, Voice AI

OR · On-site +1

$122K - $161K/yr

You'll work across the full voice AI stack: telephony, speech-to-text, LLM orchestration, text-to-speech, and analytics - building agentic conversational systems that directly improve patient access ...

Over 50 million people use Speechify's text-to-speech products to turn whatever they're reading - PDFs, books, Google Docs, news articles, websites - into audio, so they can read faster, read more ...

Go-to-Market - Bend, OR, USA

Bend, OR · On-site

$150K - $200K/yr

Over 50 million people use Speechify's text-to-speech products to turn whatever they're reading - PDFs, books, Google Docs, news articles, websites - into audio, so they can read faster, read more ...

AI Engineer

OR · On-site +1

Familiarity with Text-to-Speech and Speech-to-Text models/solutions like Deepgram, ElevenLabs, Cartesia etc. * Experience in computer vision, time-series modelling (ARIMA, Prophet) or multimodal AI.

Over 50 million people use Speechify's text-to-speech products to turn whatever they're reading - PDFs, books, Google Docs, news articles, websites - into audio, so they can read faster, read more ...

Over 50 million people use Speechify's text-to-speech products to turn whatever they're reading - PDFs, books, Google Docs, news articles, websites - into audio, so they can read faster, read more ...

Go-to-Market - Eugene, OR, USA

Eugene, OR · On-site

$150K - $200K/yr

Over 50 million people use Speechify's text-to-speech products to turn whatever they're reading - PDFs, books, Google Docs, news articles, websites - into audio, so they can read faster, read more ...

Go-to-Market - Eugene, OR, USA

Eugene, OR · On-site

$150K - $200K/yr

Over 50 million people use Speechify's text-to-speech products to turn whatever they're reading - PDFs, books, Google Docs, news articles, websites - into audio, so they can read faster, read more ...

Travel Speech Language Pathologist

Roseburg, OR · On-site

$1.6K - $2.0K/wk

... text, or email--you consent to receive communications regarding job opportunities. Message ... Speech Language Pathologist:SLP,07:00:00-19:00:00 About Zack Group Zack Group has been active in ...

Travel Speech Language Pathologist

Roseburg, OR · On-site

$1.6K - $2.0K/wk

... text, or email--you consent to receive communications regarding job opportunities. Message ... Speech Language Pathologist:SLP,07:00:00-19:00:00 About Zack Group Zack Group has been active in ...

Travel Speech Language Pathologist

Roseburg, OR · On-site

$1.6K - $2.0K/wk

... text, or email--you consent to receive communications regarding job opportunities. Message ... Speech Language Pathologist:SLP,07:00:00-19:00:00 About Zack Group Zack Group has been active in ...

Showing results 21-40

Text To Speech information

What is a text to speech job?

A Text To Speech (TTS) job typically involves converting written text into spoken audio using specialized software or AI technology. Professionals in this field may work on developing, fine-tuning, or implementing TTS systems for various applications, such as virtual assistants, accessibility tools, or audiobooks. The role can also include tasks like voice data collection, script editing, and quality assurance of generated speech. TTS jobs are important for making digital content more accessible to people with visual impairments or reading difficulties. The field combines elements of linguistics, software engineering, and artificial intelligence.

What are the key skills and qualifications needed to thrive as a text to speech engineer?

To thrive as a Text to Speech Engineer, you need a strong background in computer science, linguistics, and digital signal processing, often supported by a relevant degree. Experience with machine learning frameworks, speech synthesis toolkits (like Tacotron or WaveNet), and programming languages such as Python or C++ is typically required. Creativity, analytical thinking, and cross-functional communication skills help you collaborate with diverse teams and innovate in voice technology. These skills ensure the development of accurate, natural-sounding speech systems that meet user and client needs.

What are some common challenges faced by professionals working in text to speech development roles?

Professionals in Text to Speech development often encounter challenges such as fine-tuning synthetic voices to sound natural and expressive, handling diverse accents or languages, and optimizing algorithms for various platforms. Collaboration with linguists, UX designers, and software engineers is frequent, as ensuring accessibility and seamless integration across applications is a top priority. Staying updated on advances in AI and deep learning is essential, as the field evolves rapidly and demands continuous improvement in both technical and creative aspects.

What is the difference between Text To Speech vs Voice Actor?

AspectText To SpeechVoice Actor
Required CredentialsNone or basic audio editing skillsVoice training, acting skills, often professional demos
Work EnvironmentSoftware, digital platforms, remoteRecording studios, on-location, remote
Industry UsageAutomation, AI, tech companiesMedia, entertainment, advertising
Search & Comparison IntentAutomated voice solutions, TTS technologyVoice acting, narration, character voices

Text To Speech involves using software to convert written text into spoken words, primarily for automation and digital applications. Voice Actors, on the other hand, provide human voice recordings for media, entertainment, and advertising. While TTS is tech-driven and often used in AI and accessibility tools, Voice Actors bring emotional nuance and personality to their performances. Both roles are essential in their respective industries, but they differ significantly in skills, environment, and purpose.

What cities in Oregon are hiring for Text To Speech jobs?

Cities in Oregon with the most Text To Speech job openings:

Infographic showing various Text To Speech job openings in Oregon as of August 2026, with employment types broken down into 63% Full Time, 19% Part Time, 5% Temporary, and 13% Contract. Highlights an 87% In-person, and 13% Remote job distribution.

Senior Software Engineer, Voice AI

Natera

OR • On-site, Remote

$122K - $161K/yr

Full-time

This job post has expired 1 day ago. Applications are no longer accepted.


Natera rating

7.8

Company rating: 7.8 out of 10

Based on 39 frontline employees who took The Breakroom Quiz

53rd of 120 rated laboratories


Job description

Role Description

This is a high-autonomy, high-agency position for a voice AI engineer who thrives at the intersection of real-time systems, conversational AI, and healthcare. You'll own the architecture and delivery of Natera's Voice AI platform - a production system handling thousands of patient calls daily that provides automated test status, identity verification, billing support, and intelligent routing to human agents.

You'll work across the full voice AI stack: telephony, speech-to-text, LLM orchestration, text-to-speech, and analytics - building agentic conversational systems that directly improve patient access to their genetic testing results. This role requires deep understanding of the intricacies unique to voice AI: real-time audio streaming, turn-taking, interruption handling, latency optimization, and the orchestration challenges that distinguish voice from text-based AI systems.

Your work will span two critical domains:

1. Voice AI Platform Engineering

Design, build, and operate Natera's production voice AI system. This includes multi-agent orchestration, real-time WebSocket audio pipelines, telephony integration, and the voice-specific challenges of latency management, VAD tuning, barge-in handling, and ASR accuracy for medical terminology.

2. Agentic Conversational Architecture

Architect and implement autonomous agent workflows that handle complex patient interactions end-to-end - identity verification, OTP validation, personalized test status delivery, billing inquiries, and intelligent escalation. You'll design tool-calling patterns, agent handoff logic, state management across conversation turns, and the analytics infrastructure needed to measure and improve call efficacy.

What You'll Do
  • Own the end-to-end voice AI architecture - from Twilio media streams through LLM orchestration to TTS output and call disposition

  • Design and implement multi-agent systems using tool calling, agent handoffs, and shared conversation state for complex patient workflows

  • Build and optimize real-time audio pipelines - WebSocket streaming, codec handling (mulaw/PCM), VAD configuration, and interruption management

  • Architect analytics and observability infrastructure for voice-specific metrics: per-segment latency (STT/LLM/TTS), call efficacy, disposition accuracy, and ASR error rates

  • Solve voice-specific challenges: turn-taking timing, silence detection thresholds, barge-in recovery, medical term recognition, and end-to-end latency optimization

  • Integrate voice agents with internal services via secure authenticated APIs

  • Drive platform reliability - eliminate single points of failure, implement multi-provider LLM failover, and design graceful degradation paths

  • Collaborate with product and clinical operations to improve self-serve efficacy rates and reduce call escalations

  • Mentor team members on voice AI best practices and contribute to architectural decisions

What We're Looking For
  • 5+ years of software engineering experience, with at least 2 years building production voice AI or conversational AI systems

  • Deep experience with voice AI pipelines - you understand the end-to-end flow from telephony through STT, LLM processing, TTS, and back to the caller, and you've solved real problems at each stage

  • Production experience with agentic architectures - multi-agent orchestration, tool calling, agent handoffs, memory/state management, and LLM-driven decision making in real-time conversation contexts

  • Strong understanding of voice-specific challenges: VAD tuning, turn-taking, interruption/barge-in handling, latency budgets, audio codec management, and the differences between voice and text-based AI UX

  • Hands-on experience with telephony systems - Twilio (media streams, SIP, IVR), or equivalent platforms with WebSocket-based audio streaming

  • Proficiency in TypeScript/Node.js with strong async programming patterns; experience with NestJS or similar frameworks

  • Experience with STT/TTS providers (Deepgram, OpenAI, ElevenLabs, Azure Speech) and understanding of ASR accuracy challenges (domain-specific vocabulary, noise handling)

  • Production experience with LLM APIs - OpenAI (especially Realtime API), Anthropic Claude, or equivalent; prompt engineering for conversational agents

  • High agency and autonomy - you don't wait for permission, detailed specs, or hand-holding. You unblock yourself, seek out the highest-impact work, and drive it to completion

  • Excellent communication - you can translate complex voice AI architecture decisions for product and clinical stakeholders

Preferred
  • Experience in healthcare, biotech, or regulated environments (HIPAA, PHI handling, zero-retention architectures, BAA compliance)

  • AWS infrastructure experience - ECS Fargate, Lambda, DynamoDB, Bedrock, Kafka/MSK, API Gateway, CDK

  • Background in real-time systems: WebSocket lifecycle management, connection resilience, streaming protocols

  • Experience building analytics pipelines for voice/conversational metrics (call efficacy, disposition tracking, latency observability)

  • Familiarity with RAG architectures (vector stores, embedding models, chunking strategies) for knowledge-grounded voice agents

  • Track record of migrating or evaluating vendor platforms while maintaining production uptime

  • Experience with Datadog APM, LLM Observability, or equivalent monitoring for AI systems

  • Prior experience in a high-growth startup or zero-to-one product environment


What Natera employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom