ML Research Scientist
Austin, TX · On-site
SEMRON is redefining what's possible in AI hardware, and they are seeking an ML Research Scientist to design algorithms and quantization schemes for efficient inference on their analog in-memory ...
Austin, TX · On-site
SEMRON is redefining what's possible in AI hardware, and they are seeking an ML Research Scientist to design algorithms and quantization schemes for efficient inference on their analog in-memory ...
Austin, TX · On-site
SEMRON is redefining what's possible in AI hardware, and they are seeking an ML Research Scientist to design algorithms and quantization schemes for efficient inference on their analog in-memory ...
Profile, quantize, and/or distill ML models (RL policies or action heads) to reduce inference latency and memory footprint for deployment on robot hardware. * Hardware-Aware Evaluation: Build ...
Profile, quantize, and/or distill ML models (RL policies or action heads) to reduce inference latency and memory footprint for deployment on robot hardware. * Hardware-Aware Evaluation: Build ...
$128K - $174K/yr
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements. * Knowledge of storage networking (NVMe-oF, GPUDirect Storage, S3). * Background of ...
$128K - $174K/yr
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements. * Knowledge of storage networking (NVMe-oF, GPUDirect Storage, S3). * Background of ...
Edge AI / ML inference (CNN- and transformer-based workloads) * Functional-safety implementation (safe motion, redundancy, diagnostics) Engage in system-level discussions involving: * Perception ...
Edge AI / ML inference (CNN- and transformer-based workloads) * Functional-safety implementation (safe motion, redundancy, diagnostics) Engage in system-level discussions involving: * Perception ...
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements. * CUDA programming and NVIDIA GPU architecture expertise. * Proved experience influencing ...
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements. * CUDA programming and NVIDIA GPU architecture expertise. * Proved experience influencing ...
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
Quick apply
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
We own the compiler that turns highlevel models into fast, reliable inference across GPUs powering ... Experience with ML frameworks (e.g.,PyTorch, TensorFlow, JAX) and software stack (e.g.,ONNX,MLIR ...
We own the compiler that turns highlevel models into fast, reliable inference across GPUs powering ... Experience with ML frameworks (e.g.,PyTorch, TensorFlow, JAX) and software stack (e.g.,ONNX,MLIR ...
$103K - $142K/yr
We own the compiler that turns highlevel models into fast, reliable inference across GPUs powering ... Experience with ML frameworks (e.g.,PyTorch, TensorFlow, JAX) and software stack (e.g.,ONNX,MLIR ...
$103K - $142K/yr
We own the compiler that turns highlevel models into fast, reliable inference across GPUs powering ... Experience with ML frameworks (e.g.,PyTorch, TensorFlow, JAX) and software stack (e.g.,ONNX,MLIR ...
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
... ML and inference product(s) * Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement ...
Austin, TX · On-site
$272 - $431.25/hr
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements.* CUDA programming and NVIDIA GPU architecture expertise.* Proved experience influencing product ...
Austin, TX · On-site
$272 - $431.25/hr
Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements.* CUDA programming and NVIDIA GPU architecture expertise.* Proved experience influencing product ...
Austin, TX · On-site
... with AI/ML inference stacks (ONNX Runtime, PyTorch, TensorRT-equivalent ecosystems, etc.) • Experience with GPU computing frameworks (ROCm strongly preferred; CUDA familiarity useful) • ...
Austin, TX · On-site
... with AI/ML inference stacks (ONNX Runtime, PyTorch, TensorRT-equivalent ecosystems, etc.) • Experience with GPU computing frameworks (ROCm strongly preferred; CUDA familiarity useful) • ...
Austin, TX · Remote
$139K - $174K/yr
AI/ML Domain Knowledge: Hands-on experience hosting large language or multimodal models using inference engines like vLLM, SGLang, or TensorRT. * Inference Frameworks: Familiarity with distributed ...
Austin, TX · Remote
$139K - $174K/yr
AI/ML Domain Knowledge: Hands-on experience hosting large language or multimodal models using inference engines like vLLM, SGLang, or TensorRT. * Inference Frameworks: Familiarity with distributed ...
Austin, TX · On-site
They are seeking an experienced CV/ML Platform Engineer to design, build, and manage the data ... and model inference metrics. • Strong Linux systems knowledge (Debian/Ubuntu), including ...
Austin, TX · On-site
They are seeking an experienced CV/ML Platform Engineer to design, build, and manage the data ... and model inference metrics. • Strong Linux systems knowledge (Debian/Ubuntu), including ...
Austin, TX · Remote
$200K - $230K/yr
Familiarity with ontology/asset modeling, multimodal ML pipelines, and production ML inference integration. * Defense tech or other high-stakes, complex systems background. NOT A FIT IF: * Experience ...
Quick apply
Austin, TX · Remote
$200K - $230K/yr
Familiarity with ontology/asset modeling, multimodal ML pipelines, and production ML inference integration. * Defense tech or other high-stakes, complex systems background. NOT A FIT IF: * Experience ...
Austin, TX · On-site
They are seeking an experienced CV/ML Platform Engineer to design and maintain the data, model, and ... and model inference metrics. • Strong Linux systems knowledge (Debian/Ubuntu), including ...
Austin, TX · On-site
They are seeking an experienced CV/ML Platform Engineer to design and maintain the data, model, and ... and model inference metrics. • Strong Linux systems knowledge (Debian/Ubuntu), including ...
Austin, TX · On-site
They are seeking an experienced CV/ML Platform Engineer to design and manage the infrastructure for ... and model inference metrics. • Strong Linux systems knowledge (Debian/Ubuntu), including ...
Austin, TX · On-site
They are seeking an experienced CV/ML Platform Engineer to design and manage the infrastructure for ... and model inference metrics. • Strong Linux systems knowledge (Debian/Ubuntu), including ...
Austin, TX · On-site
They are seeking an experienced CV/ML Platform Engineer to design, build, and manage the ... and model inference metrics. • Strong Linux systems knowledge (Debian/Ubuntu), including ...
Austin, TX · On-site
They are seeking an experienced CV/ML Platform Engineer to design, build, and manage the ... and model inference metrics. • Strong Linux systems knowledge (Debian/Ubuntu), including ...
Austin, TX · On-site
Implement retrieval strategies, prompt chaining, and inference orchestration for production use ... Strong Python skills and familiarity with ML/LLM frameworks (PyTorch, Transformers, LangChain ...
Austin, TX · On-site
Implement retrieval strategies, prompt chaining, and inference orchestration for production use ... Strong Python skills and familiarity with ML/LLM frameworks (PyTorch, Transformers, LangChain ...
Austin, TX · On-site
Position Overview We are seeking an experienced CV/ML Platform Engineer with specialization in Computer Vision and Machine Learning (CV/ML) to design, build, and own the data, model, and compute ...
Quick apply
Austin, TX · On-site
Position Overview We are seeking an experienced CV/ML Platform Engineer with specialization in Computer Vision and Machine Learning (CV/ML) to design, build, and own the data, model, and compute ...
$35.3K - $48.9K
2% of jobs
$48.9K - $62.5K
3% of jobs
$62.5K - $76.1K
6% of jobs
$76.1K - $89.7K
9% of jobs
$94K is the 25th percentile. Wages below this are outliers.
$89.7K - $103.3K
15% of jobs
The median wage is $112.3K / yr.
$103.3K - $116.9K
22% of jobs
$124.3K is the 75th percentile. Wages above this are outliers.
$116.9K - $130.4K
32% of jobs
$130.4K - $144K
3% of jobs
$144K - $157.6K
4% of jobs
$157.6K - $171.2K
1% of jobs
$171.2K - $184.8K
2% of jobs
$35.3K
$115.5K
$184.8K
| Aspect | ML Inference | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of machine learning models, programming skills | Degree in data science, statistics, or related fields |
| Work Environment | Deploying models in production, real-time data processing | Data analysis, model development, research |
| Industry Usage | AI product deployment, software companies | Research institutions, tech firms, consulting |
ML Inference focuses on deploying trained models to make predictions on new data, often in real-time. Data Scientists develop and analyze models, working primarily in research and development. While both roles require understanding of machine learning, ML Inference emphasizes deployment and operationalization, whereas Data Scientists focus on model creation and analysis.

Full-time
Re-posted 11 days ago