Principal AI/ML Engineer
Boston, NY ยท On-site
Lead the deployment of ML models into scalable, reliable production systems ... Architect training, inference, and evaluation pipelines for structured and unstructured data * Own ...
Boston, NY ยท On-site
Lead the deployment of ML models into scalable, reliable production systems ... Architect training, inference, and evaluation pipelines for structured and unstructured data * Own ...
Boston, NY ยท On-site
Lead the deployment of ML models into scalable, reliable production systems ... Architect training, inference, and evaluation pipelines for structured and unstructured data * Own ...
Buffalo, NY ยท On-site
$150 - $175/hr
... performance inference. DevOps & Architecture * Set up, maintain, and secure automated CI/CD ... Integrate AI/ML capabilities into existing PHP backend API. * Write and implement software tooling ...
New
Buffalo, NY ยท On-site
$150 - $175/hr
... performance inference. DevOps & Architecture * Set up, maintain, and secure automated CI/CD ... Integrate AI/ML capabilities into existing PHP backend API. * Write and implement software tooling ...
New
Niagara Falls, NY ยท On-site
$235 - $260/hr
Optimize inference workflows for latency, cost, and scalability * Enable LLM-driven workflows that ... Partner with ML teams to improve model performance through better grounding * Mentor engineers and ...
Niagara Falls, NY ยท On-site
$235 - $260/hr
Optimize inference workflows for latency, cost, and scalability * Enable LLM-driven workflows that ... Partner with ML teams to improve model performance through better grounding * Mentor engineers and ...
Buffalo, NY ยท On-site
$120 - $150/hr
Experience optimizing highโlatency models for realโtime inference. * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Buffalo, NY ยท On-site
$120 - $150/hr
Experience optimizing highโlatency models for realโtime inference. * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Buffalo, NY ยท On-site
$110K - $133K/yr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Buffalo, NY ยท On-site
$110K - $133K/yr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
$110K - $133K/yr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
$110K - $133K/yr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
$110K - $133K/yr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
$110K - $133K/yr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Buffalo, NY ยท On-site
$120 - $150/hr
Experience optimizing high-latency models for real-time inference. * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Buffalo, NY ยท On-site
$120 - $150/hr
Experience optimizing high-latency models for real-time inference. * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
$54 - $71.50/hr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
$54 - $71.50/hr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Buffalo, NY ยท On-site
$54 - $71.50/hr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Buffalo, NY ยท On-site
$54 - $71.50/hr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
$54 - $71.50/hr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
$54 - $71.50/hr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Boston, NY ยท On-site
$200K - $350K/yr
Work with LLMs, generative AI, and modern ML frameworks. * Build inference and evaluation pipelines. * Optimize model performance, latency, and cost. * Integrate AI capabilities into production ...
Boston, NY ยท On-site
$200K - $350K/yr
Work with LLMs, generative AI, and modern ML frameworks. * Build inference and evaluation pipelines. * Optimize model performance, latency, and cost. * Integrate AI capabilities into production ...
The ML Observability team builds cutting-edge tools to monitor, explain, and improve AI systems in ... Understanding of model internals, inference pipelines, evaluation techniques, and prompt ...
The ML Observability team builds cutting-edge tools to monitor, explain, and improve AI systems in ... Understanding of model internals, inference pipelines, evaluation techniques, and prompt ...
Boston, NY ยท On-site
$200K - $350K/yr
Deep Python and ML engineering expertise. * Strong production LLM experience. * Strong distributed systems and system design skills. * Experience optimizing inference systems. * Strong technical ...
Boston, NY ยท On-site
$200K - $350K/yr
Deep Python and ML engineering expertise. * Strong production LLM experience. * Strong distributed systems and system design skills. * Experience optimizing inference systems. * Strong technical ...
Boston, NY ยท On-site
$200K - $350K/yr
... AI/ML or software engineering experience. * Deep expertise in production AI systems. * Strong distributed systems and architecture skills. * Significant LLM and inference experience. * Strong ...
Boston, NY ยท On-site
$200K - $350K/yr
... AI/ML or software engineering experience. * Deep expertise in production AI systems. * Strong distributed systems and architecture skills. * Significant LLM and inference experience. * Strong ...
Buffalo, NY ยท On-site
$110 - $140/hr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Buffalo, NY ยท On-site
$110 - $140/hr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Boston, NY ยท On-site +1
$250K/yr
Early-stage operator experience with a proven track record of building, selling, and navigating the edge ML and on-device inference world. * You may have shipped models across multiple NPU or edge ...
Boston, NY ยท On-site +1
$250K/yr
Early-stage operator experience with a proven track record of building, selling, and navigating the edge ML and on-device inference world. * You may have shipped models across multiple NPU or edge ...
Boston, NY ยท On-site +1
$250K/yr
You may have run ML platform or infrastructure teams deploying inference across cloud, neocloud, and owned GPU or edge hardware. * You may have sold into VP Platform, data center, or infrastructure ...
Boston, NY ยท On-site +1
$250K/yr
You may have run ML platform or infrastructure teams deploying inference across cloud, neocloud, and owned GPU or edge hardware. * You may have sold into VP Platform, data center, or infrastructure ...
Buffalo, NY ยท On-site
$110 - $140/hr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (Docker) and the ML model ...
Buffalo, NY ยท On-site
$110 - $140/hr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (Docker) and the ML model ...
Buffalo, NY ยท On-site
$140K - $180K/yr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
Buffalo, NY ยท On-site
$140K - $180K/yr
Experience optimizing high-latency models for real-time inference * Backend software engineering experience in the cloud (AWS / GCP) with a focus on microservices (docker) and the ML model ...
$36.3K - $50.3K
2% of jobs
$50.3K - $64.3K
3% of jobs
$64.3K - $78.3K
6% of jobs
$78.3K - $92.3K
9% of jobs
$96.8K is the 25th percentile. Wages below this are outliers.
$92.3K - $106.3K
15% of jobs
The median wage is $115.7K / yr.
$106.3K - $120.3K
22% of jobs
$128K is the 75th percentile. Wages above this are outliers.
$120.3K - $134.3K
32% of jobs
$134.3K - $148.3K
3% of jobs
$148.3K - $162.3K
4% of jobs
$162.3K - $176.3K
1% of jobs
$176.3K - $190.3K
2% of jobs
$36.3K
$118.9K
$190.3K
| Aspect | ML Inference | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of machine learning models, programming skills | Degree in data science, statistics, or related fields |
| Work Environment | Deploying models in production, real-time data processing | Data analysis, model development, research |
| Industry Usage | AI product deployment, software companies | Research institutions, tech firms, consulting |
ML Inference focuses on deploying trained models to make predictions on new data, often in real-time. Data Scientists develop and analyze models, working primarily in research and development. While both roles require understanding of machine learning, ML Inference emphasizes deployment and operationalization, whereas Data Scientists focus on model creation and analysis.
For Ml Inference jobs in Buffalo, NY, the most frequently searched job titles are:
The top searched job categories for Ml Inference jobs in Buffalo, NY are:
Cities near Buffalo, NY with the most Ml Inference job openings:

Full-time
Medical, Dental, Vision, Retirement
Re-posted 24 days ago