1

Audio Speech Machine Learning Jobs in Michigan (NOW HIRING)

Hear speech, speak loudly / softly; Distinguish sounds, colors, hot / cold surfaces, temperatures ... audio/visual equipment and electro physiological recording equipment. Sets up computer and ...

Hear speech, speak loudly / softly; Distinguish sounds, colors, hot / cold surfaces, temperatures ... audio/visual equipment and electro physiological recording equipment. Sets up computer and ...

Hear speech, speak loudly / softly; Distinguish sounds, colors, hot / cold surfaces, temperatures ... audio/visual equipment and electro physiological recording equipment. Sets up computer and ...

... learning, intranet and computer navigation. Ability to use other software required to perform ... Hear speech, speak loudly / softly; Distinguish sounds, colors, hot / cold surfaces, temperatures ...

Senior Engineer, AI

Novi, MI · On-site

$90.75 - $133.10/hr

Engineer audio systems and integrated technology platforms that augment the driving experience ... machine learning systems and/or generative AI applications. * Strong hands‑on experience with ...

About the Job The Varsity Tutors Live Learning Platform has thousands of students looking for ... Adapts instruction using immersive role-play scenarios, authentic audio and video materials, and ...

About the Job The Varsity Tutors Live Learning Platform has thousands of students looking for ... Adapts instruction using role-play scenarios, audio immersion exercises, and practical thematic ...

Showing results 41-60

Audio Speech Machine Learning information

What are some common challenges faced when developing machine learning models for audio speech applications?

A key challenge in audio speech machine learning roles is dealing with diverse and noisy audio data, which can significantly affect model accuracy. Additionally, models must be robust to different accents, languages, and speaking styles, requiring large and varied datasets for training and validation. Collaboration with data engineers, linguists, and software developers is often necessary to ensure high-quality data pipelines and model integration into production systems. Staying updated with the latest research and optimizing models for real-time performance are also ongoing aspects of the role.

What is an audio speech machine learning engineer?

An Audio Speech Machine Learning Engineer is a specialized professional who designs, develops, and implements machine learning models that process and analyze audio and speech data. Their work involves tasks like speech recognition, speaker identification, and audio event detection by leveraging algorithms and large datasets. These engineers collaborate with data scientists, software developers, and linguists to create applications such as voice assistants, transcription tools, and automated customer service systems. Expertise in signal processing, deep learning frameworks, and programming languages like Python is crucial for this role.

What is the difference between Audio Speech Machine Learning vs Speech Data Analyst?

AspectAudio Speech Machine LearningSpeech Data Analyst
Required CredentialsDegree in Computer Science, Data Science, or related fields; knowledge of ML frameworksDegree in Data Analysis, Statistics, or related fields; experience with data tools
Work EnvironmentResearch labs, tech companies, AI startupsData analysis teams, research institutions, tech firms
Industry UsageDeveloping speech recognition, voice assistants, NLP applicationsAnalyzing speech datasets, improving speech models, reporting insights

Audio Speech Machine Learning focuses on developing algorithms for speech recognition and processing, often involving model training and AI development. Speech Data Analysts interpret speech data, generate insights, and support model improvements. Both roles require strong analytical skills, but their core tasks differ: one builds models, the other analyzes data.

What are the key skills and qualifications needed to thrive as an audio speech machine learning engineer, and why are they important?

To thrive as an Audio Speech Machine Learning Engineer, you need a solid background in machine learning, signal processing, and programming (typically Python), along with a relevant degree in computer science or a related field. Familiarity with tools like TensorFlow or PyTorch, audio processing libraries (such as Librosa), and experience with speech datasets and ASR systems are commonly required. Critical soft skills include problem-solving, innovation, and effective communication for collaborating with cross-functional teams. These skills are essential to develop accurate, scalable speech recognition systems that advance voice-driven technology.
What are popular job titles related to Audio Speech Machine Learning jobs in Michigan? For Audio Speech Machine Learning jobs in Michigan, the most frequently searched job titles are:
What job categories do people searching Audio Speech Machine Learning jobs in Michigan look for? The top searched job categories for Audio Speech Machine Learning jobs in Michigan are:
What cities in Michigan are hiring for Audio Speech Machine Learning jobs? Cities in Michigan with the most Audio Speech Machine Learning job openings:
Infographic showing various Audio Speech Machine Learning job openings in Michigan as of August 2026, with employment types broken down into 100% Full Time. Highlights an 100% In-person job distribution.

Research Intern on Generative and Protective AI for Content Creation

Sony

Detroit, MI • On-site, Remote

$50/hr

Full-time

This job post has expired today. Applications are no longer accepted.


Job description

Sony AI America, a branch of Sony AI, is a remotely distributed organization spread across the U.S. and Canada. Sony AI is Sony's new research organization pursuing the mission to use AI to unleash human creativity. Sony AI works closely with Sony's other business units, including Sony Interactive Entertainment LLC., Sony Pictures Entertainment Inc., and Sony Music Entertainment. With some 900 million Sony devices in hands and homes worldwide today, a vast array of Sony movies, television shows and music, and the PlayStation Network, Sony creates and delivers more entertainment experiences to more people than anyone else on earth. To learn more: https://ai.sony/

Position Summary

Sony AI is seeking research interns who are passionate about ML-based technologies that are useful in movie , game , and music creation, as well as technologies for the ethical operations of generative AIs. Our mission is to research and develop technologies for various Sony Group products and for scientific publication. The technologies developed during the internship will have the potential to be applied in film, game , and music production across entertainment studios worldwide.

Responsibilities

As a research intern, you will investigate and apply novel algorithms related to:

  • The generation and editing of video, sound, and 3D visual geometry (including 3D human motion).
  • The analysis and alleviation of ethical flaws in generative models, including techniques for memorization detection and mitigation, concept erasure, and data attribution.

Your goal will be to publish your findings in a top-tier conference. You are expected to be self-motivated and to implement innovative ideas using your research, coding, and problem-solving skills. You will receive support from internal scientists and engineers in your efforts.

Required Qualifications

  • Master's degree in Computer Science , Electrical Engineering, Applied Mathematics, or related fields.
  • Proven knowledge and expertise in generative AI applications, including deep generative modeling, computer vision, and audio signal processing.
  • Strong analytical and programming skills in deep learning using frameworks and tools for machine learning (e.g., PyTorch ) and visualization (e.g., TensorBoard ).
  • Experience in research communities, including published papers at top-tier conferences such as CVPR, ICCV, ECCV, NeurIPS , ICLR, and ICML.
  • Excellent communication and presentation skills.

Preferred Qualifications

  • Currently enrolled in a relevant Ph.D. program in fields such as Computer Science, Electrical Engineering, Applied Mathematics, or related disciplines.

Working Location

Location flexible (Tokyo, NYC, remote)

The target hourly rate for this internship is $50.00 per hour. The individual will be paid hourly and eligible for overtime .

LI-AS1

All qualified applicants will receive consideration for employment without regard to any basis protected by applicable federal, state, or local law, ordinance, or regulation.

Disability Accommodation for Applicants to Sony Corporation of America Sony Corporation of America provides reasonable accommodation for qualified individuals with disabilities and disabled veterans in job application procedures. For reasonable accommodation requests, please contact us by email at careers@sonyusa.com or by mail to: Sony Corporation of America, People Experience Department, 25 Madison Avenue, New York, NY 10010. Please indicate the position you are applying for.

We are aware that unauthorized individuals or organizations may attempt to solicit personal information or payments from job applicants by impersonating our company through fraudulent job postings. We take these matters seriously but cannot control third-party websites. To protect your personal information, please verify that any job posting you respond to also appears on our official Careers page: www.sonyjobs.com . Please also be advised that we never request personal identifying information (such as Social Security numbers, bank details, or copies of identification documents) during the initial stages of our application process. If you have any doubts about the authenticity of a job posting or communication, please contact careers@sonyusa.com before submitting any information.

Right to Work (English/Spanish)

E-Verify Participation (English/Spanish)