$88K - $106K/yr
You will combine deep ML expertise, systems engineering, and performance analysis to deliver faster, more efficient AI experiences. * Optimize machine learning inference systems to improve latency ...
$88K - $106K/yr
You will combine deep ML expertise, systems engineering, and performance analysis to deliver faster, more efficient AI experiences. * Optimize machine learning inference systems to improve latency ...
$88K - $106K/yr
You will combine deep ML expertise, systems engineering, and performance analysis to deliver faster, more efficient AI experiences. * Optimize machine learning inference systems to improve latency ...
$83K - $113K/yr
Experience self-hosing ML inference. * Hands-on experience with Vertex AI, Kubeflow, or TensorFlow Serving in production. * Background in event-driven architectures and message streaming (e.g., Pub ...
$83K - $113K/yr
Experience self-hosing ML inference. * Hands-on experience with Vertex AI, Kubeflow, or TensorFlow Serving in production. * Background in event-driven architectures and message streaming (e.g., Pub ...
$94K - $124K/yr
Experience with ML inference technologies such as vLLM, TensorRT-LLM, Triton Inference Server, SGLang, or similar platforms is a plus. * Knowledge of runtime optimization techniques including cold ...
$94K - $124K/yr
Experience with ML inference technologies such as vLLM, TensorRT-LLM, Triton Inference Server, SGLang, or similar platforms is a plus. * Knowledge of runtime optimization techniques including cold ...
... ML inference, data pipelines, and platform services. • Establish and scale engineering best practices, standards, and frameworks across teams. • Ensure platform reliability, performance ...
... ML inference, data pipelines, and platform services. • Establish and scale engineering best practices, standards, and frameworks across teams. • Ensure platform reliability, performance ...
Saint Louis, MO · On-site
Work on causal inference techniques such as causal ML, matching models, or uplift modeling to evaluate business interventions. * Collaborate with product, business, and engineering teams to translate ...
Saint Louis, MO · On-site
Work on causal inference techniques such as causal ML, matching models, or uplift modeling to evaluate business interventions. * Collaborate with product, business, and engineering teams to translate ...
Saint Louis, MO · On-site
Work on causal inference techniques such as causal ML, matching models, or uplift modeling to evaluate business interventions. * Collaborate with product, business, and engineering teams to translate ...
Quick apply
Saint Louis, MO · On-site
Work on causal inference techniques such as causal ML, matching models, or uplift modeling to evaluate business interventions. * Collaborate with product, business, and engineering teams to translate ...
Saint Louis, MO · On-site
Work on causal inference techniques such as causal ML, matching models, or uplift modeling to evaluate business interventions. * Collaborate with product, business, and engineering teams to translate ...
Saint Louis, MO · On-site
Work on causal inference techniques such as causal ML, matching models, or uplift modeling to evaluate business interventions. * Collaborate with product, business, and engineering teams to translate ...
Saint Louis, MO · On-site
Design and architect enterprise-grade AI/ML-powered Java applications on AWS cloud infrastructure ... inference optimization, prompt engineering Spring AI, Langgraph, Google ADK, A2A, MCP. * Prompt ...
Saint Louis, MO · On-site
Design and architect enterprise-grade AI/ML-powered Java applications on AWS cloud infrastructure ... inference optimization, prompt engineering Spring AI, Langgraph, Google ADK, A2A, MCP. * Prompt ...
Deliver governed datasets and feature engineering/serving for ML training and real-time inference (online/offline consistency, caching, latency SLOs, backfills). A successful candidate would possess ...
Deliver governed datasets and feature engineering/serving for ML training and real-time inference (online/offline consistency, caching, latency SLOs, backfills). A successful candidate would possess ...
Deliver governed datasets and feature engineering/serving for ML training and real-time inference (online/offline consistency, caching, latency SLOs, backfills). A successful candidate would possess ...
Deliver governed datasets and feature engineering/serving for ML training and real-time inference (online/offline consistency, caching, latency SLOs, backfills). A successful candidate would possess ...
Glenallen, MO · On-site
$200 - $300/hr
Engineer scalable ML pipelines: Develop robust feature engineering, training, inference, and monitoring pipelines built for reliability and scale. * Ship end-to-end: Take models from prototype ...
Glenallen, MO · On-site
$200 - $300/hr
Engineer scalable ML pipelines: Develop robust feature engineering, training, inference, and monitoring pipelines built for reliability and scale. * Ship end-to-end: Take models from prototype ...
Noel, MO · On-site
$110K - $220K/yr
Causal Inference & Elasticity: Identification of treatment effects beyond simple log-log approaches (Double ML, Instrumental Variables, Uplift modeling); Optimization & Reinforcement Learning: Multi ...
New
Noel, MO · On-site
$110K - $220K/yr
Causal Inference & Elasticity: Identification of treatment effects beyond simple log-log approaches (Double ML, Instrumental Variables, Uplift modeling); Optimization & Reinforcement Learning: Multi ...
New
Cassville, MO · On-site
$110K - $220K/yr
Causal Inference & Elasticity: Identification of treatment effects beyond simple log-log approaches (Double ML, Instrumental Variables, Uplift modeling); Optimization & Reinforcement Learning: Multi ...
New
Cassville, MO · On-site
$110K - $220K/yr
Causal Inference & Elasticity: Identification of treatment effects beyond simple log-log approaches (Double ML, Instrumental Variables, Uplift modeling); Optimization & Reinforcement Learning: Multi ...
New
Anderson, MO · On-site
$110K - $220K/yr
Causal Inference & Elasticity: Identification of treatment effects beyond simple log-log approaches (Double ML, Instrumental Variables, Uplift modeling); Optimization & Reinforcement Learning: Multi ...
New
Anderson, MO · On-site
$110K - $220K/yr
Causal Inference & Elasticity: Identification of treatment effects beyond simple log-log approaches (Double ML, Instrumental Variables, Uplift modeling); Optimization & Reinforcement Learning: Multi ...
New
California, MO · On-site
$180 - $250/hr
Partner with AI/ML engineering teams to optimize inference performance in frameworks such as PyTorch and TensorFlow. * Establish benchmarking frameworks and lead performance tuning efforts for ...
California, MO · On-site
$180 - $250/hr
Partner with AI/ML engineering teams to optimize inference performance in frameworks such as PyTorch and TensorFlow. * Establish benchmarking frameworks and lead performance tuning efforts for ...
Saint Louis, MO · On-site
Design and architect enterprise-grade AI/ML-powered Java applications on AWS cloud infrastructure ... inference optimization, prompt engineering Spring AI, Langgraph, Google ADK, A2A, MCP. Prompt ...
Saint Louis, MO · On-site
Design and architect enterprise-grade AI/ML-powered Java applications on AWS cloud infrastructure ... inference optimization, prompt engineering Spring AI, Langgraph, Google ADK, A2A, MCP. Prompt ...
You will own critical areas including model fine-tuning, inference optimization, ML infrastructure, evaluation frameworks, and deployment processes. Working within a small, highly skilled engineering ...
You will own critical areas including model fine-tuning, inference optimization, ML infrastructure, evaluation frameworks, and deployment processes. Working within a small, highly skilled engineering ...
Cassville, MO · On-site
$110K - $220K/yr
Architect end‑to‑end ML systems--from feature engineering through production deployment and ... Engineer computer vision systems for real‑time inference (YOLO, RT‑DETR, CLIP) with multi‑GPU ...
Cassville, MO · On-site
$110K - $220K/yr
Architect end‑to‑end ML systems--from feature engineering through production deployment and ... Engineer computer vision systems for real‑time inference (YOLO, RT‑DETR, CLIP) with multi‑GPU ...
Noel, MO · On-site
$110K - $220K/yr
Architect end‑to‑end ML systems--from feature engineering through production deployment and ... Engineer computer vision systems for real‑time inference (YOLO, RT‑DETR, CLIP) with multi‑GPU ...
Noel, MO · On-site
$110K - $220K/yr
Architect end‑to‑end ML systems--from feature engineering through production deployment and ... Engineer computer vision systems for real‑time inference (YOLO, RT‑DETR, CLIP) with multi‑GPU ...
O Fallon, MO · Hybrid
$97K - $134K/yr
We are looking for a Senior AI Engineer, Generative AI & ML Engineering for the Operational ... Develop high-performance Python backend services for LLM inference orchestration, async job ...
O Fallon, MO · Hybrid
$97K - $134K/yr
We are looking for a Senior AI Engineer, Generative AI & ML Engineering for the Operational ... Develop high-performance Python backend services for LLM inference orchestration, async job ...
| Aspect | ML Inference | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of machine learning models, programming skills | Degree in data science, statistics, or related fields |
| Work Environment | Deploying models in production, real-time data processing | Data analysis, model development, research |
| Industry Usage | AI product deployment, software companies | Research institutions, tech firms, consulting |
ML Inference focuses on deploying trained models to make predictions on new data, often in real-time. Data Scientists develop and analyze models, working primarily in research and development. While both roles require understanding of machine learning, ML Inference emphasizes deployment and operationalization, whereas Data Scientists focus on model creation and analysis.

$88K - $106K/yr
Full-time
Posted 14 days ago
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Machine Learning Engineer - Inference Optimization based in Netherlands.
This role offers the opportunity to optimize the performance of advanced machine learning systems used in real-world production environments.
You will work at the intersection of research and engineering, transforming cutting-edge models into fast, reliable, and cost-efficient solutions.
Your work will directly impact model scalability, user experience, and the efficiency of AI-powered products.
You will dive deep into performance optimization, from model architecture and GPU execution to large-scale inference infrastructure.
Working with talented research, infrastructure, and product teams, you will help push the boundaries of what AI systems can achieve.
This position is ideal for an engineer who enjoys solving complex technical challenges and building high-performance ML systems from the ground up.
As a Machine Learning Engineer specializing in inference optimization, you will own the performance and scalability of machine learning models in production. You will combine deep ML expertise, systems engineering, and performance analysis to deliver faster, more efficient AI experiences.
The ideal candidate is a technically strong machine learning engineer with experience optimizing production inference systems and a passion for high-performance AI engineering. You should enjoy working on complex technical problems, experimenting with new approaches, and taking ownership of critical systems.