Machine Learning Engineer
Palo Alto, CA · On-site
As a Machine Learning Engineer, you will play a central role in translating cutting-edge machine ... quantization, deployment optimization). * Experienced in inference time optimization, deep ...
Palo Alto, CA · On-site
As a Machine Learning Engineer, you will play a central role in translating cutting-edge machine ... quantization, deployment optimization). * Experienced in inference time optimization, deep ...
Palo Alto, CA · On-site
As a Machine Learning Engineer, you will play a central role in translating cutting-edge machine ... quantization, deployment optimization). * Experienced in inference time optimization, deep ...
Palo Alto, CA · On-site
$206K - $258K/yr
Strong understanding of deep learning software models. * Experience in compiler pipeline ... Deep knowledge of numerical precision trade-offs, quantization-aware training (QAT), and dynamic ...
Palo Alto, CA · On-site
$206K - $258K/yr
Strong understanding of deep learning software models. * Experience in compiler pipeline ... Deep knowledge of numerical precision trade-offs, quantization-aware training (QAT), and dynamic ...
Palo Alto, CA · On-site
$206K - $258K/yr
Strong understanding of deep learning software models. * Experience in compiler pipeline ... Deep knowledge of numerical precision trade-offs, quantization-aware training (QAT), and dynamic ...
Palo Alto, CA · On-site
$206K - $258K/yr
Strong understanding of deep learning software models. * Experience in compiler pipeline ... Deep knowledge of numerical precision trade-offs, quantization-aware training (QAT), and dynamic ...
Palo Alto, CA · On-site
Nace AI is a company focused on machine learning solutions, and they are seeking a Machine Learning ... quantization, deployment optimization). • Experienced in inference time optimization, deep ...
Palo Alto, CA · On-site
Nace AI is a company focused on machine learning solutions, and they are seeking a Machine Learning ... quantization, deployment optimization). • Experienced in inference time optimization, deep ...
Palo Alto, CA · On-site
$206K - $258K/yr
Deep knowledge of numerical precision trade-offs, quantization-aware training (QAT), and dynamic ... Strong understanding of deep learning software models. * Experience in compiler pipeline ...
Palo Alto, CA · On-site
$206K - $258K/yr
Deep knowledge of numerical precision trade-offs, quantization-aware training (QAT), and dynamic ... Strong understanding of deep learning software models. * Experience in compiler pipeline ...
Cupertino, CA · On-site
$151K - $199K/yr
You'll stay current with developments across computer vision, deep learning, and adjacent ML fields ... Experience optimizing models for on-device/edge inference (quantization, pruning, distillation, or ...
Cupertino, CA · On-site
$151K - $199K/yr
You'll stay current with developments across computer vision, deep learning, and adjacent ML fields ... Experience optimizing models for on-device/edge inference (quantization, pruning, distillation, or ...
Palo Alto, CA · On-site
They are seeking a Machine Learning Engineer to translate research into scalable solutions ... quantization, deployment optimization). • Experienced in inference time optimization, deep ...
Palo Alto, CA · On-site
They are seeking a Machine Learning Engineer to translate research into scalable solutions ... quantization, deployment optimization). • Experienced in inference time optimization, deep ...
Palo Alto, CA · On-site
$122K - $164K/yr
... of deep learning modern architectures, optimization, quantization, and model alignment • Ability to think 'out-of-the-box' to push performance beyond apparent limitations • Excellent problem ...
Palo Alto, CA · On-site
$122K - $164K/yr
... of deep learning modern architectures, optimization, quantization, and model alignment • Ability to think 'out-of-the-box' to push performance beyond apparent limitations • Excellent problem ...
Sunnyvale, CA · On-site
$122K - $167K/yr
... deep learning approaches. • Expertise in model acceleration, quantization, or compression (TensorRT, ONNX Runtime). • Familiarity with real-time frameworks and middleware such as ROS 2, GStreamer ...
Sunnyvale, CA · On-site
$122K - $167K/yr
... deep learning approaches. • Expertise in model acceleration, quantization, or compression (TensorRT, ONNX Runtime). • Familiarity with real-time frameworks and middleware such as ROS 2, GStreamer ...
... deep learning workloads to heterogeneous device backends You will also partner up with peer science teams to innovate on model quantization and compression techniques for efficient execution on ...
... deep learning workloads to heterogeneous device backends You will also partner up with peer science teams to innovate on model quantization and compression techniques for efficient execution on ...
... deep learning workloads to heterogeneous device backends You will also partner up with peer science teams to innovate on model quantization and compression techniques for efficient execution on ...
... deep learning workloads to heterogeneous device backends You will also partner up with peer science teams to innovate on model quantization and compression techniques for efficient execution on ...
$190K - $235K/yr
Strong classical computer vision skills (geometry-based methods, feature extraction) complementing deep learning approaches. * Expertise in model acceleration, quantization, or compression (TensorRT ...
$190K - $235K/yr
Strong classical computer vision skills (geometry-based methods, feature extraction) complementing deep learning approaches. * Expertise in model acceleration, quantization, or compression (TensorRT ...
You will feel a deep sense of responsibility in proactively protecting our community thoughtfully ... learning, quantization, LoRA, distillation). * Drive End-to-End Product Development: You will not ...
You will feel a deep sense of responsibility in proactively protecting our community thoughtfully ... learning, quantization, LoRA, distillation). * Drive End-to-End Product Development: You will not ...
$184K - $324K/yr
You'll stay current with developments across computer vision, deep learning, and adjacent ML fields ... Experience optimizing models for on-device/edge inference (quantization, pruning, distillation, or ...
$184K - $324K/yr
You'll stay current with developments across computer vision, deep learning, and adjacent ML fields ... Experience optimizing models for on-device/edge inference (quantization, pruning, distillation, or ...
Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training frameworks (Ray, Horovod, etc.) * Knowledge of model optimization (quantization, pruning) and CUDA is a ...
Quick apply
Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training frameworks (Ray, Horovod, etc.) * Knowledge of model optimization (quantization, pruning) and CUDA is a ...
Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training frameworks (Ray, Horovod, etc.) * Knowledge of model optimization (quantization, pruning) and CUDA is a ...
Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training frameworks (Ray, Horovod, etc.) * Knowledge of model optimization (quantization, pruning) and CUDA is a ...
Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training frameworks (Ray, Horovod, etc.) * Knowledge of model optimization (quantization, pruning) and CUDA is a ...
Proficiency in deep learning frameworks such as PyTorch * Experience with distributed training frameworks (Ray, Horovod, etc.) * Knowledge of model optimization (quantization, pruning) and CUDA is a ...
Sunnyvale, CA · On-site
$122K - $168K/yr
Strong classical computer vision skills (geometry-based methods, feature extraction) complementing deep learning approaches. * Expertise in model acceleration, quantization, or compression (TensorRT ...
Sunnyvale, CA · On-site
$122K - $168K/yr
Strong classical computer vision skills (geometry-based methods, feature extraction) complementing deep learning approaches. * Expertise in model acceleration, quantization, or compression (TensorRT ...
Proficiency in Python and deep learning frameworks such as PyTorch or TensorFlow. * Solid ... Experience with ONNX, TensorRT, model quantization, C++ inference pipelines, CUDA , or edge ...
Proficiency in Python and deep learning frameworks such as PyTorch or TensorFlow. * Solid ... Experience with ONNX, TensorRT, model quantization, C++ inference pipelines, CUDA , or edge ...
Proficiency in Python and deep learning frameworks such as PyTorch or TensorFlow. * Solid ... Experience with ONNX, TensorRT, model quantization, C++ inference pipelines, CUDA , or edge ...
Proficiency in Python and deep learning frameworks such as PyTorch or TensorFlow. * Solid ... Experience with ONNX, TensorRT, model quantization, C++ inference pipelines, CUDA , or edge ...
$25.6K is the 25th percentile. Wages below this are outliers.
$12.9K - $26.7K
27% of jobs
$26.7K - $40.5K
0% of jobs
$40.5K - $54.2K
0% of jobs
$54.2K - $68K
0% of jobs
$68K - $81.8K
0% of jobs
The median wage is $94.4K / yr.
$81.8K - $95.6K
25% of jobs
$95.6K - $109.3K
18% of jobs
$119.2K is the 75th percentile. Wages above this are outliers.
$109.3K - $123.1K
7% of jobs
$123.1K - $136.9K
2% of jobs
$136.9K - $150.6K
0% of jobs
$150.6K - $164.4K
21% of jobs
$12.9K
$98.5K
$164.4K
| Aspect | Deep Learning Quantization | Machine Learning Engineer |
|---|---|---|
| Required Credentials | Advanced degrees in AI, Computer Science, or related fields; knowledge of neural networks | Bachelor's or Master's in CS, Data Science, or related fields; programming skills |
| Work Environment | Research labs, AI development teams, hardware optimization settings | Software development teams, data-driven projects, product-focused environments |
| Industry Usage | AI hardware optimization, model deployment, edge computing | Model development, data analysis, software solutions across industries |
Deep Learning Quantization focuses on reducing model size and improving inference speed through techniques like weight and activation quantization, often in hardware or embedded systems. Machine Learning Engineers develop, implement, and optimize machine learning models for various applications. While both roles require knowledge of AI and programming, Deep Learning Quantization is more specialized in model optimization techniques, whereas Machine Learning Engineers work broadly on model development and deployment.

Full-time
Re-posted 20 days ago