1

Audio Speech Machine Learning Jobs in Boston, MA

Machine Learning Engineer

Somerville, MA ยท On-site

$170 - $200/hr

We're looking for a Senior Machine Learning Engineer to help advance the state of voice ... Experience with audio models or speech systems (ASR, TTS, speaker modeling, etc.) * Experience with ...

New

DSP Engineer - Audio Tech

Framingham, MA ยท On-site

$147K - $171K/yr

Exposure to applied machine learning techniques for audio systems , such as ML-assisted noise reduction, speech enhancement, or audio classification. * Familiarity with Bluetooth audio systems and ...

DSP Engineer - Audio Tech

Framingham, MA

$147K - $171K/yr

Exposure to applied machine learning techniques for audio systems , such as ML-assisted noise reduction, speech enhancement, or audio classification. * Familiarity with Bluetooth audio systems and ...

next page

Showing results 1-20

Audio Speech Machine Learning information

What are some common challenges faced when developing machine learning models for audio speech applications?

A key challenge in audio speech machine learning roles is dealing with diverse and noisy audio data, which can significantly affect model accuracy. Additionally, models must be robust to different accents, languages, and speaking styles, requiring large and varied datasets for training and validation. Collaboration with data engineers, linguists, and software developers is often necessary to ensure high-quality data pipelines and model integration into production systems. Staying updated with the latest research and optimizing models for real-time performance are also ongoing aspects of the role.

What is an audio speech machine learning engineer?

An Audio Speech Machine Learning Engineer is a specialized professional who designs, develops, and implements machine learning models that process and analyze audio and speech data. Their work involves tasks like speech recognition, speaker identification, and audio event detection by leveraging algorithms and large datasets. These engineers collaborate with data scientists, software developers, and linguists to create applications such as voice assistants, transcription tools, and automated customer service systems. Expertise in signal processing, deep learning frameworks, and programming languages like Python is crucial for this role.

What is the difference between Audio Speech Machine Learning vs Speech Data Analyst?

AspectAudio Speech Machine LearningSpeech Data Analyst
Required CredentialsDegree in Computer Science, Data Science, or related fields; knowledge of ML frameworksDegree in Data Analysis, Statistics, or related fields; experience with data tools
Work EnvironmentResearch labs, tech companies, AI startupsData analysis teams, research institutions, tech firms
Industry UsageDeveloping speech recognition, voice assistants, NLP applicationsAnalyzing speech datasets, improving speech models, reporting insights

Audio Speech Machine Learning focuses on developing algorithms for speech recognition and processing, often involving model training and AI development. Speech Data Analysts interpret speech data, generate insights, and support model improvements. Both roles require strong analytical skills, but their core tasks differ: one builds models, the other analyzes data.

What are the key skills and qualifications needed to thrive as an audio speech machine learning engineer, and why are they important?

To thrive as an Audio Speech Machine Learning Engineer, you need a solid background in machine learning, signal processing, and programming (typically Python), along with a relevant degree in computer science or a related field. Familiarity with tools like TensorFlow or PyTorch, audio processing libraries (such as Librosa), and experience with speech datasets and ASR systems are commonly required. Critical soft skills include problem-solving, innovation, and effective communication for collaborating with cross-functional teams. These skills are essential to develop accurate, scalable speech recognition systems that advance voice-driven technology.
What are popular job titles related to Audio Speech Machine Learning jobs in Boston, MA? For Audio Speech Machine Learning jobs in Boston, MA, the most frequently searched job titles are:
What job categories do people searching Audio Speech Machine Learning jobs in Boston, MA look for? The top searched job categories for Audio Speech Machine Learning jobs in Boston, MA are:
What cities near Boston, MA are hiring for Audio Speech Machine Learning jobs? Cities near Boston, MA with the most Audio Speech Machine Learning job openings:

Machine Learning Engineer

Sierra Ventures

Somerville, MA โ€ข On-site

$170 - $200/hr

Other

Medical, Dental, Vision, PTO

Posted 2 days ago

New


Job description

Modulate is the leader in conversational voice intelligence. We enable enterprises to deeply understand how people communicate and take timely action based on those insights. Our products help detect harm, prevent fraud, and build safer, more trusted online and real-world voice environments. We are building a Conversation Intelligence Platform โ€” APIs, workflows, and applications that bring voice understanding to customers at enterprise scale.

We're looking for a Senior Machine Learning Engineer to help advance the state of voice understanding at Modulate. In this role, you'll design, train, evaluate, and deploy cuttingโ€‘edge machine learning models that power our products. You'll work closely with researchers, engineers, product leaders, and executives to bring innovative ML solutions from concept to production.

Your Impact
  • Develop and deploy high-quality machine learning models that power Modulate's products
  • Advance our capabilities in conversational voice intelligence through applied research and engineering
  • Help translate business needs into scalable ML solutions
  • Improve model performance, reliability, and efficiency across our platform
  • Contribute to a collaborative, highโ€‘performing engineering culture
What You Will Do
  • Design, train, evaluate, and deploy machine learning models for production applications
  • Collaborate with engineers, researchers, product managers, and company leadership to define and execute on ML initiatives
  • Conduct experiments to improve model quality, accuracy, robustness, and scalability
  • Partner with platform and infrastructure teams to operationalize and monitor ML systems in production
  • Analyze model performance and identify opportunities for improvement
  • Communicate technical findings, tradeoffs, and recommendations to both technical and nonโ€‘technical stakeholders
  • Contribute to technical strategy and help shape the future direction of Modulate's ML systems
  • Review code, share knowledge, and mentor teammates as needed
  • Stay current on advances in machine learning and identify opportunities to apply new techniques to our products
What We Are Looking For
  • Experience conducting machine learning research and shipping models to production
  • Strong experience building and deploying productionโ€‘grade machine learning systems
  • Strong experience with Python and PyTorch
  • Experience designing experiments and evaluating model performance
  • Ability to work across research and engineering disciplines to deliver business impact
  • Strong communication skills and the ability to explain complex technical concepts clearly
  • Experience working collaboratively in crossโ€‘functional environments
Nice to Have
  • Experience communicating research externally through papers, conferences, or openโ€‘source contributions
  • Experience with audio models or speech systems (ASR, TTS, speaker modeling, etc.)
  • Experience with cloud infrastructure, especially AWS
  • Experience building and maintaining largeโ€‘scale ML infrastructure or MLOps systems
  • Experience in fastโ€‘paced startup environments
Benefits
  • Competitive salary + equity
  • Full health, dental, and vision coverage
  • Flexible PTO with a strong culture of taking it
  • Weekly team lunches with dietary accommodations
  • Hybrid work with core inโ€‘office days and flexible remote options
  • Leadership and technical learning sessions
  • Career development and continued growth support
  • Up to 8 weeks workโ€‘fromโ€‘anywhere policy
  • A deeply inclusive, humanโ€‘centered culture
Pay Transparency

Salary: $170,000โ€“$200,000
Equity: Offered

Additional benefits include HSA, FSA, 15 company holidays, and professional development resources.

Candidates may be located anywhere in the U.S., with preference for proximity to major tech hubs such as Boston, San Francisco, Seattle, New York, or Austin.

About Modulate

Modulate is on a mission to make voice a force for good online. Our tools help communities thrive by proactively detecting toxic behavior, protecting user identity, and empowering safety teams. We're trusted by leaders in gaming and beyondโ€”and we're growing fast.

We believe that great cultures don't just happen. That's why we've built a foundation of intentional systems: from biasโ€‘reducing hiring practices to transparent pay to tools that help teams collaborate across communication styles. At Modulate, we treat people like peopleโ€”and we're building technology that does the same.

Ready to join us? Apply here or reach out directlyโ€”we're excited to meet you.

#J-18808-Ljbffr