ML Research Scientist
Austin, TX · On-site
SEMRON is redefining what's possible in AI hardware, and they are seeking an ML Research Scientist to design algorithms and quantization schemes for efficient inference on their analog in-memory ...
Austin, TX · On-site
SEMRON is redefining what's possible in AI hardware, and they are seeking an ML Research Scientist to design algorithms and quantization schemes for efficient inference on their analog in-memory ...
Austin, TX · On-site
SEMRON is redefining what's possible in AI hardware, and they are seeking an ML Research Scientist to design algorithms and quantization schemes for efficient inference on their analog in-memory ...
Plano, TX · On-site
$100K - $137K/yr
As a Senior AI/ML Platform Engineer, you will design, build, and support scalable platform ... Develop reusable patterns for inference services, prompt flow integration, and performance tuning ...
Plano, TX · On-site
$100K - $137K/yr
As a Senior AI/ML Platform Engineer, you will design, build, and support scalable platform ... Develop reusable patterns for inference services, prompt flow integration, and performance tuning ...
Austin, TX · On-site
$128K - $174K/yr
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements. * Knowledge of storage networking (NVMe-oF, GPUDirect Storage, S3). * Background of ...
Austin, TX · On-site
$128K - $174K/yr
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements. * Knowledge of storage networking (NVMe-oF, GPUDirect Storage, S3). * Background of ...
Plano, TX · On-site
$98K - $129K/yr
You will help enable secure, production-ready MLOps and LLMOps infrastructure that supports model training, inference, orchestration, and retrieval-augmented generation. The Lead AI/ML Platform ...
Plano, TX · On-site
$98K - $129K/yr
You will help enable secure, production-ready MLOps and LLMOps infrastructure that supports model training, inference, orchestration, and retrieval-augmented generation. The Lead AI/ML Platform ...
Edge AI / ML inference (CNN- and transformer-based workloads) * Functional-safety implementation (safe motion, redundancy, diagnostics) Engage in system-level discussions involving: * Perception ...
Edge AI / ML inference (CNN- and transformer-based workloads) * Functional-safety implementation (safe motion, redundancy, diagnostics) Engage in system-level discussions involving: * Perception ...
$98K - $129K/yr
You will help enable secure, production-ready MLOps and LLMOps infrastructure that supports model training, inference, orchestration, and retrieval-augmented generation. The Lead AI/ML Platform ...
$98K - $129K/yr
You will help enable secure, production-ready MLOps and LLMOps infrastructure that supports model training, inference, orchestration, and retrieval-augmented generation. The Lead AI/ML Platform ...
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements. * CUDA programming and NVIDIA GPU architecture expertise. * Proved experience influencing ...
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements. * CUDA programming and NVIDIA GPU architecture expertise. * Proved experience influencing ...
... inference. AI/ML & GenAI Systems • Design and implement AI/ML pipelines integrating: o Large Language Models (LLMs) o Small Language Models (SMLs) o Reasoning and task specific models • Build ...
... inference. AI/ML & GenAI Systems • Design and implement AI/ML pipelines integrating: o Large Language Models (LLMs) o Small Language Models (SMLs) o Reasoning and task specific models • Build ...
Plano, TX · On-site
ML Engineer Location: Remote -ML Engineer Job Summary We are seeking a talented AI/ML Engineer to ... Optimize model inference performance for scalability and reliability. · Document model ...
Plano, TX · On-site
ML Engineer Location: Remote -ML Engineer Job Summary We are seeking a talented AI/ML Engineer to ... Optimize model inference performance for scalability and reliability. · Document model ...
Austin, TX · On-site
$140 - $180/hr
Experience securing AI/ML inference platforms, model‑serving infrastructure, accelerator‑based systems, or confidential AI workloads. * Experience with trusted execution environments, platform ...
Austin, TX · On-site
$140 - $180/hr
Experience securing AI/ML inference platforms, model‑serving infrastructure, accelerator‑based systems, or confidential AI workloads. * Experience with trusted execution environments, platform ...
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
Quick apply
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
Houston, TX · On-site
$128K - $172K/yr
Strong foundation in statistics, A/B testing, causal inference, and experimental design • ... ML engineering, or related roles • 3+ years building NLP/generative AI applications and ...
Houston, TX · On-site
$128K - $172K/yr
Strong foundation in statistics, A/B testing, causal inference, and experimental design • ... ML engineering, or related roles • 3+ years building NLP/generative AI applications and ...
Austin, TX · On-site
We own the compiler that turns highlevel models into fast, reliable inference across GPUs powering ... Experience with ML frameworks (e.g.,PyTorch, TensorFlow, JAX) and software stack (e.g.,ONNX,MLIR ...
Austin, TX · On-site
We own the compiler that turns highlevel models into fast, reliable inference across GPUs powering ... Experience with ML frameworks (e.g.,PyTorch, TensorFlow, JAX) and software stack (e.g.,ONNX,MLIR ...
$103K - $142K/yr
We own the compiler that turns highlevel models into fast, reliable inference across GPUs powering ... Experience with ML frameworks (e.g.,PyTorch, TensorFlow, JAX) and software stack (e.g.,ONNX,MLIR ...
$103K - $142K/yr
We own the compiler that turns highlevel models into fast, reliable inference across GPUs powering ... Experience with ML frameworks (e.g.,PyTorch, TensorFlow, JAX) and software stack (e.g.,ONNX,MLIR ...
Austin, TX · On-site
... with AI/ML inference stacks (ONNX Runtime, PyTorch, TensorRT-equivalent ecosystems, etc.) • Experience with GPU computing frameworks (ROCm strongly preferred; CUDA familiarity useful) • ...
Austin, TX · On-site
... with AI/ML inference stacks (ONNX Runtime, PyTorch, TensorRT-equivalent ecosystems, etc.) • Experience with GPU computing frameworks (ROCm strongly preferred; CUDA familiarity useful) • ...
Austin, TX · Remote
$139K - $174K/yr
AI/ML Domain Knowledge: Hands-on experience hosting large language or multimodal models using inference engines like vLLM, SGLang, or TensorRT. * Inference Frameworks: Familiarity with distributed ...
Austin, TX · Remote
$139K - $174K/yr
AI/ML Domain Knowledge: Hands-on experience hosting large language or multimodal models using inference engines like vLLM, SGLang, or TensorRT. * Inference Frameworks: Familiarity with distributed ...
Austin, TX · On-site
$272 - $431.25/hr
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements.* CUDA programming and NVIDIA GPU architecture expertise.* Proved experience influencing product ...
Austin, TX · On-site
$272 - $431.25/hr
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements.* CUDA programming and NVIDIA GPU architecture expertise.* Proved experience influencing product ...
Austin, TX · On-site
They are seeking an experienced CV/ML Platform Engineer to design, build, and manage the data ... and model inference metrics. • Strong Linux systems knowledge (Debian/Ubuntu), including ...
Austin, TX · On-site
They are seeking an experienced CV/ML Platform Engineer to design, build, and manage the data ... and model inference metrics. • Strong Linux systems knowledge (Debian/Ubuntu), including ...
| Aspect | ML Inference | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of machine learning models, programming skills | Degree in data science, statistics, or related fields |
| Work Environment | Deploying models in production, real-time data processing | Data analysis, model development, research |
| Industry Usage | AI product deployment, software companies | Research institutions, tech firms, consulting |
ML Inference focuses on deploying trained models to make predictions on new data, often in real-time. Data Scientists develop and analyze models, working primarily in research and development. While both roles require understanding of machine learning, ML Inference emphasizes deployment and operationalization, whereas Data Scientists focus on model creation and analysis.
For Ml Inference jobs in Texas, the most frequently searched job titles are:
The top searched job categories for Ml Inference jobs in Texas are:
Cities in Texas with the most Ml Inference job openings:

Full-time
Re-posted 20 days ago