1

Audio Speech Machine Learning Jobs in Texas (NOW HIRING)

Staff Machine Learning Engineer

Austin, TX · On-site +1

$208K - $255K/yr

Jeppesen ForeFlight is seeking a Senior Machine Learning Engineer to help build and scale domain-specialized automatic speech recognition (ASR) systems for aviation and operational audio workflows.

next page

Showing results 1-20

Audio Speech Machine Learning information

What are some common challenges faced when developing machine learning models for audio speech applications?

A key challenge in audio speech machine learning roles is dealing with diverse and noisy audio data, which can significantly affect model accuracy. Additionally, models must be robust to different accents, languages, and speaking styles, requiring large and varied datasets for training and validation. Collaboration with data engineers, linguists, and software developers is often necessary to ensure high-quality data pipelines and model integration into production systems. Staying updated with the latest research and optimizing models for real-time performance are also ongoing aspects of the role.

What is an audio speech machine learning engineer?

An Audio Speech Machine Learning Engineer is a specialized professional who designs, develops, and implements machine learning models that process and analyze audio and speech data. Their work involves tasks like speech recognition, speaker identification, and audio event detection by leveraging algorithms and large datasets. These engineers collaborate with data scientists, software developers, and linguists to create applications such as voice assistants, transcription tools, and automated customer service systems. Expertise in signal processing, deep learning frameworks, and programming languages like Python is crucial for this role.

What is the difference between Audio Speech Machine Learning vs Speech Data Analyst?

AspectAudio Speech Machine LearningSpeech Data Analyst
Required CredentialsDegree in Computer Science, Data Science, or related fields; knowledge of ML frameworksDegree in Data Analysis, Statistics, or related fields; experience with data tools
Work EnvironmentResearch labs, tech companies, AI startupsData analysis teams, research institutions, tech firms
Industry UsageDeveloping speech recognition, voice assistants, NLP applicationsAnalyzing speech datasets, improving speech models, reporting insights

Audio Speech Machine Learning focuses on developing algorithms for speech recognition and processing, often involving model training and AI development. Speech Data Analysts interpret speech data, generate insights, and support model improvements. Both roles require strong analytical skills, but their core tasks differ: one builds models, the other analyzes data.

What are the key skills and qualifications needed to thrive as an audio speech machine learning engineer, and why are they important?

To thrive as an Audio Speech Machine Learning Engineer, you need a solid background in machine learning, signal processing, and programming (typically Python), along with a relevant degree in computer science or a related field. Familiarity with tools like TensorFlow or PyTorch, audio processing libraries (such as Librosa), and experience with speech datasets and ASR systems are commonly required. Critical soft skills include problem-solving, innovation, and effective communication for collaborating with cross-functional teams. These skills are essential to develop accurate, scalable speech recognition systems that advance voice-driven technology.
What are popular job titles related to Audio Speech Machine Learning jobs in Texas? For Audio Speech Machine Learning jobs in Texas, the most frequently searched job titles are:
What job categories do people searching Audio Speech Machine Learning jobs in Texas look for? The top searched job categories for Audio Speech Machine Learning jobs in Texas are:
What cities in Texas are hiring for Audio Speech Machine Learning jobs? Cities in Texas with the most Audio Speech Machine Learning job openings:

Staff Machine Learning Engineer

ForeFlight

Austin, TX • On-site, Remote

$208K - $255K/yr

Full-time

Medical, Dental, Vision, Retirement, PTO

Re-posted 13 days ago


Job description

Jeppesen ForeFlight builds industry-leading aviation software used by pilots, aircraft operators, and major airlines worldwide. As a high-growth, private equity-backed company, we are focused on scaling our operations, strengthening our financial infrastructure, and driving operational excellence across the business. Our team combines deep domain expertise with a collaborative, high-performance culture to solve complex challenges and support continued growth.
Jeppesen ForeFlight is seeking a Senior Machine Learning Engineer to help build and scale domain-specialized automatic speech recognition (ASR) systems for aviation and operational audio workflows. This role focuses on developing vertical ASR models optimized for high-accuracy transcription in noisy, safety-critical environments, including aviation communications, cockpit interactions, operational dispatch, maintenance coordination, and related specialized audio domains.
You will work across the full ML lifecycle - data engineering, model training, evaluation, deployment, and optimization - to deliver production-grade speech intelligence capabilities integrated into ForeFlight and broader aviation platforms.
This position is ideal for someone with deep expertise in speech AI, acoustic modeling, large-scale transcription pipelines, and domain adaptation techniques for specialized vocabularies and constrained communication environments.
Key Responsibilities
  • Design, train, and optimize domain-specific ASR models for aviation and operational communications.
  • Develop verticalized speech models tuned for specialized terminology, accents, abbreviations, call signs, and noisy radio/audio conditions.
  • Build and maintain large-scale transcription and labeling pipelines for supervised and semi-supervised learning workflows.
  • Fine-tune foundation speech models (e.g., Whisper, wav2vec, Conformer, RNN-T, Citrinet, NeMo-based architectures) for aviation-specific use cases.
  • Improve transcription quality through language model adaptation, pronunciation lexicons, contextual biasing, and decoding optimization.
  • Develop evaluation frameworks and benchmarking methodologies using WER, CER, domain entity accuracy, latency, and robustness metrics.
  • Collaborate with product, avionics, data engineering, and platform teams to deploy scalable real-time and batch transcription systems.
  • Optimize inference pipelines for edge, cloud, and low-latency streaming environments.
  • Research emerging techniques in speech enhancement, diarization, speaker adaptation, multilingual ASR, and audio foundation models.
  • Ensure compliance with security, privacy, and operational reliability standards required in aviation environments.

Required Qualifications
  • Bachelor's or Master's degree in Computer Science, Machine Learning, Electrical Engineering, Linguistics, or related field.
  • 3+ years of experience in speech recognition, audio ML, or applied machine learning.
  • Strong experience training and fine-tuning ASR models using frameworks such as PyTorch or TensorFlow.
  • Experience with modern ASR architectures including:

    • Transformer-based ASR
    • Conformer
    • RNN-T
    • CTC-based systems
    • Encoder-decoder speech models

  • Experience working with:

    • Speech/audio preprocessing
    • Forced alignment
    • Language model adaptation
    • Beam search decoding
    • Noise robustness techniques

  • Familiarity with NVIDIA NeMo, Kaldi, ESPnet, Hugging Face, Whisper, DeepSpeed, or equivalent ecosystems.
  • Strong Python engineering skills and experience building production ML systems.
  • Experience with cloud infrastructure and ML deployment workflows (AWS, Kubernetes, Docker, CI/CD).
  • Ability to work with large audio datasets and distributed training environments.

Preferred Qualifications
  • Experience building ASR systems for aviation, air traffic control, public safety, defense, or other mission-critical domains.
  • Familiarity with VHF/UHF radio communications and noisy-channel audio processing.
  • Experience with multilingual or code-switching ASR systems.
  • Background in speech enhancement, keyword spotting, diarization, or speaker verification.
  • Knowledge of LLM-assisted transcription correction and retrieval-augmented speech systems.
  • Experience optimizing models for real-time streaming inference.
  • Active pilot experience or familiarity with aviation operations is a plus.

Why Join Us
At Jeppesen ForeFlight, we know you want a rewarding career. To do that, you need challenging projects, a good work environment, and awesome coworkers. We believe in our employees, and we empower them to make a direct impact on our products and services. We strive to provide our employees with a world-class benefits experience, focused on supporting their physical, financial, and emotional wellbeing. Our benefits package includes but is not limited to the following:
  • Medical, dental, vision insurance with Employer paid health premiums
  • Open PTO Policy
  • 401(k) with up to 10% company matching and immediate vesting
  • 12 Weeks Paid Maternity Leave
  • 4 Weeks Paid Paternity Leave
  • Flight Training Rewards

Pay is based upon candidate experience and qualifications, as well market and business considerations: Summary Pay Range: $208,000-$255,000
Jeppesen ForeFlight - EOE including Disability/Vets | Pay Transparency | E-Verify Participant | Equal Opportunity Employer
Equal Opportunity Employer
This employer is required to notify all applicants of their rights pursuant to federal employment laws.
For further information, please review the Know Your Rights notice from the Department of Labor.