Lead Machine Learning Inference Engineer, Advertising
$101K - $133K/yr
Contributions to open-source ML or systems projects #LI-DH2
$101K - $133K/yr
Contributions to open-source ML or systems projects #LI-DH2
$101K - $133K/yr
Contributions to open-source ML or systems projects #LI-DH2
Manor, TX · On-site
$110K - $145K/yr
The role focuses on real-time inference, feature engineering, APIs, graph-based fraud detection ... Integrate ML models with REST APIs and microservices. * Support graph-based fraud detection using ...
Manor, TX · On-site
$110K - $145K/yr
The role focuses on real-time inference, feature engineering, APIs, graph-based fraud detection ... Integrate ML models with REST APIs and microservices. * Support graph-based fraud detection using ...
Irving, TX · On-site
$85 - $107/hr
This role is pivotal in enabling enterprise-scale ML and generative AI capabilities by building ... Design highly available and performant serving environments for LLM inference using Azure ...
Irving, TX · On-site
$85 - $107/hr
This role is pivotal in enabling enterprise-scale ML and generative AI capabilities by building ... Design highly available and performant serving environments for LLM inference using Azure ...
$85K - $107K/yr
This role is pivotal in enabling enterprise-scale ML and generative AI capabilities by building ... Design highly available and performant serving environments for LLM inference using Azure ...
$85K - $107K/yr
This role is pivotal in enabling enterprise-scale ML and generative AI capabilities by building ... Design highly available and performant serving environments for LLM inference using Azure ...
Austin, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
Austin, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
San Antonio, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
San Antonio, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
Fort Worth, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
Fort Worth, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
Houston, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
Houston, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
Austin, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
Austin, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
Austin, TX · On-site
$120K - $225K/yr
Enable early firmware, driver, runtime, and ML stack bring-up on emulated hardware * Support execution of AI inference workloads (e.g., CNNs, transformers) on emulated Mythic accelerators
Austin, TX · On-site
$120K - $225K/yr
Enable early firmware, driver, runtime, and ML stack bring-up on emulated hardware * Support execution of AI inference workloads (e.g., CNNs, transformers) on emulated Mythic accelerators
Dallas, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
Dallas, TX · On-site
... inference, and monitoring in production environments. • Participate in continuous improvement of the ML infrastructure and processes for scalability and performance. Qualifications : Required : • ...
Dallas, TX · On-site
$85K - $107K/yr
This role is pivotal in enabling enterprise-scale ML and generative AI capabilities by building ... Design highly available and performant serving environments for LLM inference using Azure ...
Dallas, TX · On-site
$85K - $107K/yr
This role is pivotal in enabling enterprise-scale ML and generative AI capabilities by building ... Design highly available and performant serving environments for LLM inference using Azure ...
$150K - $225K/yr
We build and operate ML, GenAI, Inference and Data Platforms and Services to provide a comprehensive suite of capabilities-serving business-critical needs across Apple's enterprise. We work on ...
$150K - $225K/yr
We build and operate ML, GenAI, Inference and Data Platforms and Services to provide a comprehensive suite of capabilities-serving business-critical needs across Apple's enterprise. We work on ...
Austin, TX · On-site
$120K - $225K/yr
Enable early firmware, driver, runtime, and ML stack bring-up on emulated hardware * Support execution of AI inference workloads (e.g., CNNs, transformers) on emulated Mythic accelerators
Quick apply
Austin, TX · On-site
$120K - $225K/yr
Enable early firmware, driver, runtime, and ML stack bring-up on emulated hardware * Support execution of AI inference workloads (e.g., CNNs, transformers) on emulated Mythic accelerators
Plano, TX · On-site
Core Responsibilities (AI/ML, Python, AWS, GenAI) * Design and implement end-to-end AI/ML and ... Build robust MLOps workflows, including model versioning, containerized training/inference ...
Quick apply
Plano, TX · On-site
Core Responsibilities (AI/ML, Python, AWS, GenAI) * Design and implement end-to-end AI/ML and ... Build robust MLOps workflows, including model versioning, containerized training/inference ...
Austin, TX · On-site
$150K - $225K/yr
We build and operate ML, GenAI, Inference and Data Platforms and Services to provide a comprehensive suite of capabilities-serving business-critical needs across Apple's enterprise. We work on ...
Austin, TX · On-site
$150K - $225K/yr
We build and operate ML, GenAI, Inference and Data Platforms and Services to provide a comprehensive suite of capabilities-serving business-critical needs across Apple's enterprise. We work on ...
Austin, TX · On-site
$120K - $225K/yr
Enable early firmware, driver, runtime, and ML stack bring-up on emulated hardware * Support execution of AI inference workloads (e.g., CNNs, transformers) on emulated Mythic accelerators
Austin, TX · On-site
$120K - $225K/yr
Enable early firmware, driver, runtime, and ML stack bring-up on emulated hardware * Support execution of AI inference workloads (e.g., CNNs, transformers) on emulated Mythic accelerators
We're looking for a strong technical leader with deep experience in ML serving, high-performance ... Lead the design and development of a SOTA Inference platform * Oversee the development of ...
We're looking for a strong technical leader with deep experience in ML serving, high-performance ... Lead the design and development of a SOTA Inference platform * Oversee the development of ...
Austin, TX · On-site
$120 - $180/hr
We build and operate ML, GenAI, Inference and Data Platforms and Services to provide a comprehensive suite of capabilities--serving business-critical needs across Apple's enterprise. We work on ...
Austin, TX · On-site
$120 - $180/hr
We build and operate ML, GenAI, Inference and Data Platforms and Services to provide a comprehensive suite of capabilities--serving business-critical needs across Apple's enterprise. We work on ...
Irving, TX · On-site
Preferred : • Financial domain expertise (risk, fraud, forecasting, customer intelligence). • Advanced ML topics: time series, graph ML, optimization, causal inference. • ONNX/TensorRT model ...
Irving, TX · On-site
Preferred : • Financial domain expertise (risk, fraud, forecasting, customer intelligence). • Advanced ML topics: time series, graph ML, optimization, causal inference. • ONNX/TensorRT model ...
| Aspect | ML Inference | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of machine learning models, programming skills | Degree in data science, statistics, or related fields |
| Work Environment | Deploying models in production, real-time data processing | Data analysis, model development, research |
| Industry Usage | AI product deployment, software companies | Research institutions, tech firms, consulting |
ML Inference focuses on deploying trained models to make predictions on new data, often in real-time. Data Scientists develop and analyze models, working primarily in research and development. While both roles require understanding of machine learning, ML Inference emphasizes deployment and operationalization, whereas Data Scientists focus on model creation and analysis.
For Ml Inference jobs in Texas, the most frequently searched job titles are:
The top searched job categories for Ml Inference jobs in Texas are:
Cities in Texas with the most Ml Inference job openings:

$101K - $133K/yr
Full-time
Re-posted yesterday
The Advertising Performance group focuses on performance for all participants in the Advertising ecosystem - Advertisers, Publishers, and Roku. The systems and solutions span multiple disciplines and technologies to perform real-time multi-objective optimization across distributed systems at large scale and with low latency. We use Machine Learning, Reinforcement Learning, AI, Control and Optimization Systems, and Auction Dynamics to solve a large set of complex problems. At the core of this is our Machine Learning and Inference Platform that powers the entire landscape.
About the roleIn this role, you will architect, design, and lead the development of a SOTA Inference platform that can handle Advertising-level low latencies, scale, throughput, and availability with optimizations that span across hardware, software, and models. We're looking for a strong technical leader with deep experience in ML serving, high-performance computing, and industry standard frameworks - someone excited to mentor engineers, innovate at scale, and shape the future of machine learning at Roku.
What you'll be doingSourced by ZipRecruiter
Manufacturing
1,001 - 5,000 Employees
San Jose, CA, US
2002