1

Machine Learning Data Linguist Jobs in Missouri (NOW HIRING)

$80K - $110K/yr

Our partner is looking for a Senior Geospatial Machine Learning Engineer based in Netherlands. Join ... You will collaborate with teams across Europe and the Americas, influencing data pipelines ...

A successful candidate will have an established hands-on data science, AI, ML experience in driving ... Work on Python and mainstream machine learning frameworks, e.g. TensorFlow or PyTorch and Agentic ...

Staff, Machine Learning Engineer

Noel, MO · On-site

$130K - $260K/yr

A successful candidate will have an established hands-on data science, AI, ML experience in driving ... Work on Python and mainstream machine learning frameworks, e.g. TensorFlow or PyTorch and Agentic ...

A successful candidate will have an established hands-on data science, AI, ML experience in driving ... Work on Python and mainstream machine learning frameworks, e.g. TensorFlow or PyTorch and Agentic ...

Data Science Tutor

Saint Louis, MO · Remote

$18 - $40/hr

Deep knowledge of statistical analysis, data wrangling, exploratory data analysis, machine learning, data visualization, SQL, Python or R programming, hypothesis testing, and communication of data ...

Data Science Tutor

Columbia, MO · Remote

$18 - $40/hr

Deep knowledge of statistical analysis, data wrangling, exploratory data analysis, machine learning, data visualization, SQL, Python or R programming, hypothesis testing, and communication of data ...

Data Science Tutor

Kansas City, MO · Remote

$18 - $40/hr

Deep knowledge of statistical analysis, data wrangling, exploratory data analysis, machine learning, data visualization, SQL, Python or R programming, hypothesis testing, and communication of data ...

Showing results 41-60

Machine Learning Data Linguist information

What is a machine learning data linguist?

A Machine Learning Data Linguist is a specialist who works at the intersection of linguistics and artificial intelligence. They are responsible for annotating, curating, and analyzing language data to train and improve machine learning models, especially those focused on natural language processing (NLP). Their work often includes tasks like labeling text, refining speech recognition data, and ensuring that language models understand context, grammar, and cultural nuances. This role is essential in developing accurate and inclusive AI systems that interact with human language.

How does a machine learning data linguist typically collaborate with engineers and data scientists on projects?

A Machine Learning Data Linguist works closely with engineers and data scientists by providing linguistic insights and ensuring that language data is accurately annotated and interpreted. They often participate in cross-functional meetings to define project goals, clarify annotation guidelines, and review model outputs for linguistic quality. This collaboration helps bridge the gap between technical development and language-specific nuances, leading to more effective and culturally accurate machine learning models. Effective communication and a strong understanding of both linguistic theory and technical requirements are vital in this collaborative environment.

What are the key skills and qualifications needed to thrive as a machine learning data linguist, and why are they important?

To thrive as a Machine Learning Data Linguist, you need expertise in linguistics, data annotation, and a strong understanding of language structures, often supported by a degree in linguistics or computational linguistics. Familiarity with annotation tools, data labeling platforms, and programming languages like Python is typically required. Strong attention to detail, analytical thinking, and clear communication are essential soft skills for accurately interpreting and conveying linguistic phenomena. These skills ensure high-quality language data, which is critical for developing effective and unbiased machine learning models.

What cities in Missouri are hiring for Machine Learning Data Linguist jobs?

Cities in Missouri with the most Machine Learning Data Linguist job openings:

Senior/Staff Machine Learning Engineer, Data Infrastructure

Jobtailor

California, MO • On-site

$120 - $160/hr

Other

Posted 4 days ago


Job description

  • Develop infrastructure supporting batch and stream big data processing using Flink, Spark, Ray, and similar technologies
  • Design and operate large-scale data pipelines generating training datasets for machine learning training and experimentation
  • Integrate data pipelines with workflow orchestration systems such as Flyte and Airflow for reliable multi-stage training workflows
  • Improve pipeline reproducibility and observability through dataset validation, monitoring, and automated testing
  • Optimize performance and resource utilization across distributed compute systems
  • Partner with ML engineers to enable large-scale experimentation and model iteration
  • Lead architectural improvements to keep offline data pipelines scalable, reliable, and cost-efficient
Requirements
  • Experience working with distributed computing frameworks such as Flink, Spark, and Ray for distributed data processing
  • Experience building infrastructure for training data generation, dataset preparation, or ML feature pipelines
  • Experience optimizing big data pipelines and infrastructure for cost efficiency
  • Strong programming skills in Python and experience with large-scale distributed workloads
  • Experience with modern data infrastructure, including data lakes, warehouses, orchestration systems, and streaming platforms
  • Strong systems thinking and ability to reason about performance, scalability, reliability, and cost tradeoffs in distributed systems
  • Proven ability to lead technical direction and influence architectural decisions across teams without formal authority
  • Sufficient knowledge of English for professional verbal and written exchanges
  • Work visa/immigration sponsorship is not available for this position
Core Competencies

Demonstrates expertise in developing and optimizing large-scale data pipelines for machine learning, utilizing distributed computing frameworks like Flink, Spark, and Ray. Proficient in integrating orchestration systems and ensuring pipeline reliability and cost efficiency.

Highest-signal resume keywords
  • Flink
  • Spark
  • Ray
  • Python Programming
  • Data Pipeline Optimization
ATS Optimization Keywords Hard Skills
  • Distributed Computing
  • Data Pipeline Development
  • Machine Learning Feature Engineering
  • Dataset Validation
  • Automated Testing
Soft Skills
  • Systems Thinking
  • Technical Leadership
  • Influencing Architectural Decisions
Industry Keywords
  • Big Data Processing
  • Performance Optimization
  • Resource Utilization
  • Scalability
  • Reliability
Tools & Technologies
  • Flyte
  • Airflow
  • Data Lakes
  • Data Warehouses
  • Streaming Platforms
#J-18808-Ljbffr