Responsibilities : • Contribute to building AI/ML & Analytics platform, services, and tools across dev, test, and prod environments to accelerate model training, inference, and deployment. • ...
Responsibilities : • Contribute to building AI/ML & Analytics platform, services, and tools across dev, test, and prod environments to accelerate model training, inference, and deployment. • ...
AI/ML Engineer
Philadelphia, PA · On-site
$109K - $131K/yr
Optimize model selection, prompt size, token usage, batching, caching, inference frequency, and ... or ML engineering. * Experience with the below tech stack is required: * Python (advanced ...
Quick apply
AI/ML Engineer
Philadelphia, PA · On-site
$109K - $131K/yr
Optimize model selection, prompt size, token usage, batching, caching, inference frequency, and ... or ML engineering. * Experience with the below tech stack is required: * Python (advanced ...
Machine Learning Engineer- Inference Optimization | Experienced Hire
Bala Cynwyd, PA · On-site
$110 - $150/hr
You will work on inference workloads where latency, throughput, reliability, and hardware ... Solid understanding of modern ML frameworks such as PyTorch, including model execution, export ...
Machine Learning Engineer- Inference Optimization | Experienced Hire
Bala Cynwyd, PA · On-site
$110 - $150/hr
You will work on inference workloads where latency, throughput, reliability, and hardware ... Solid understanding of modern ML frameworks such as PyTorch, including model execution, export ...
Urgent to Fill- GenAI Solutions Architect -Mount Laurel, NJ -Onsite
Mount Laurel, NJ · On-site
$62.50 - $82.25/hr
* Design end-to-end AI/ML architecture including data ingestion, model training, inference, and monitoring systems * Define technical standards, best practices, and governance frameworks for AI/ML ...
Quick apply
Urgent to Fill- GenAI Solutions Architect -Mount Laurel, NJ -Onsite
Mount Laurel, NJ · On-site
$62.50 - $82.25/hr
* Design end-to-end AI/ML architecture including data ingestion, model training, inference, and monitoring systems * Define technical standards, best practices, and governance frameworks for AI/ML ...
Deliver governed datasets and feature engineering/serving for ML training and real-time inference (online/offline consistency, caching, latency SLOs, backfills). A successful candidate would possess ...
Deliver governed datasets and feature engineering/serving for ML training and real-time inference (online/offline consistency, caching, latency SLOs, backfills). A successful candidate would possess ...
Strong experience with Python for AI/ML and backend development * Hands-on experience with open-source LLM deployment (Llama 3, Mistral, Mixtral) * Experience with CPU-based inference and ...
Strong experience with Python for AI/ML and backend development * Hands-on experience with open-source LLM deployment (Llama 3, Mistral, Mixtral) * Experience with CPU-based inference and ...
GenAI Architect
Mount Laurel, NJ · On-site
Responsibilities : • Build scalable AI/ML systems; including data pipelines; model training workflows; and inference services. • Evaluate and integrate open‐source and commercial LLMs (e.g.
GenAI Architect
Mount Laurel, NJ · On-site
Responsibilities : • Build scalable AI/ML systems; including data pipelines; model training workflows; and inference services. • Evaluate and integrate open‐source and commercial LLMs (e.g.
This role is heavily focused on econometric modeling, time series analysis, and causal inference ... Exposure to ML pipelines and MLOps concepts
This role is heavily focused on econometric modeling, time series analysis, and causal inference ... Exposure to ML pipelines and MLOps concepts
Senior Software Engineer Applied AI
Trenton, NJ · On-site
$150 - $260/hr
End‑to‑end ML pipelines: feature engineering, model training, and scheduled inference * Imbalanced, messy real‑world data; calibration and explainability for non‑technical consumers * Turning ...
Senior Software Engineer Applied AI
Trenton, NJ · On-site
$150 - $260/hr
End‑to‑end ML pipelines: feature engineering, model training, and scheduled inference * Imbalanced, messy real‑world data; calibration and explainability for non‑technical consumers * Turning ...
Solutions Architect - AI
Philadelphia, PA · On-site
$60.25 - $79.25/hr
... ML systems in production. * Define and enforce secure-by-design standards for model development, training data handling, inference APIs, and GenAI integrations. * Architect defenses against AI ...
Solutions Architect - AI
Philadelphia, PA · On-site
$60.25 - $79.25/hr
... ML systems in production. * Define and enforce secure-by-design standards for model development, training data handling, inference APIs, and GenAI integrations. * Architect defenses against AI ...
Solutions Architect - AI
Philadelphia, PA · On-site
$63.50 - $83.75/hr
... ML systems in production. * Define and enforce secure-by-design standards for model development, training data handling, inference APIs, and GenAI integrations. * Architect defenses against AI ...
Solutions Architect - AI
Philadelphia, PA · On-site
$63.50 - $83.75/hr
... ML systems in production. * Define and enforce secure-by-design standards for model development, training data handling, inference APIs, and GenAI integrations. * Architect defenses against AI ...
Solutions Architect - AI
Philadelphia, PA · On-site +1
$60.25 - $79.25/hr
... ML systems in production. * Define and enforce secure-by-design standards for model development, training data handling, inference APIs, and GenAI integrations. * Architect defenses against AI ...
Solutions Architect - AI
Philadelphia, PA · On-site +1
$60.25 - $79.25/hr
... ML systems in production. * Define and enforce secure-by-design standards for model development, training data handling, inference APIs, and GenAI integrations. * Architect defenses against AI ...
Senior Software Engineer Applied AI
Trenton, NJ · Remote
$122K - $162K/yr
End-to-end ML pipelines: feature engineering, model training, and scheduled inference * Imbalanced, messy real-world data; calibration and explainability for non-technical consumers * Turning ...
Senior Software Engineer Applied AI
Trenton, NJ · Remote
$122K - $162K/yr
End-to-end ML pipelines: feature engineering, model training, and scheduled inference * Imbalanced, messy real-world data; calibration and explainability for non-technical consumers * Turning ...
Build scalable AI/ML systems; including data pipelines; model training workflows; and inference services. * Evaluate and integrate open‑source and commercial LLMs (e.g.; GPT; Llama; Claude; Mistral)
Quick apply
Build scalable AI/ML systems; including data pipelines; model training workflows; and inference services. * Evaluate and integrate open‑source and commercial LLMs (e.g.; GPT; Llama; Claude; Mistral)
Develop practical ML models that balance predictive performance, explainability, stability ... Help define data pipelines, feature pipelines, inference flows, model outputs, feedback loops, and ...
Develop practical ML models that balance predictive performance, explainability, stability ... Help define data pipelines, feature pipelines, inference flows, model outputs, feedback loops, and ...
Google AI Lead Architect
$55.75 - $76.50/hr
Integrate and fine-tune Large Language Models (LLMs) and other AI/ML models into enterprise applications. Develop and implement strategies for model deployment, inference, and monitoring, with an ...
Google AI Lead Architect
$55.75 - $76.50/hr
Integrate and fine-tune Large Language Models (LLMs) and other AI/ML models into enterprise applications. Develop and implement strategies for model deployment, inference, and monitoring, with an ...
Senior AI Engineer, Video Search (Applied Research & Product)
Conshohocken, PA · On-site
$120 - $160/hr
On-device or edge inference; WebRTC/RTSP ingest; FFmpeg/GStreamer pipelines.* Experience in regulated or high-assurance environments (FedRAMP/HIPAA/CJIS) and privacy-preserving ML.## **Values*** **No ...
Senior AI Engineer, Video Search (Applied Research & Product)
Conshohocken, PA · On-site
$120 - $160/hr
On-device or edge inference; WebRTC/RTSP ingest; FFmpeg/GStreamer pipelines.* Experience in regulated or high-assurance environments (FedRAMP/HIPAA/CJIS) and privacy-preserving ML.## **Values*** **No ...
Integrate and fine-tune Large Language Models (LLMs) and other AI/ML models into enterprise applications. Develop and implement strategies for model deployment, inference, and monitoring, with an ...
Integrate and fine-tune Large Language Models (LLMs) and other AI/ML models into enterprise applications. Develop and implement strategies for model deployment, inference, and monitoring, with an ...
Demonstrate strong expertise in designing and executing A/B tests, analyzing experiments, and troubleshooting ML model behavior. Conduct causal inference studies and exploratory analyses to measure ...
Demonstrate strong expertise in designing and executing A/B tests, analyzing experiments, and troubleshooting ML model behavior. Conduct causal inference studies and exploratory analyses to measure ...
Demonstrate strong expertise in designing and executing A/B tests, analyzing experiments, and troubleshooting ML model behavior. Conduct causal inference studies and exploratory analyses to measure ...
Demonstrate strong expertise in designing and executing A/B tests, analyzing experiments, and troubleshooting ML model behavior. Conduct causal inference studies and exploratory analyses to measure ...
Ml Inference information
See Burlington, NJ salary details
$36.7K - $50.8K
2% of jobs
$50.8K - $64.9K
3% of jobs
$64.9K - $79.1K
6% of jobs
$79.1K - $93.2K
9% of jobs
$97.8K is the 25th percentile. Wages below this are outliers.
$93.2K - $107.3K
15% of jobs
The median wage is $116.8K / yr.
$107.3K - $121.5K
22% of jobs
$129.3K is the 75th percentile. Wages above this are outliers.
$121.5K - $135.6K
32% of jobs
$135.6K - $149.7K
3% of jobs
$149.7K - $163.9K
4% of jobs
$163.9K - $178K
1% of jobs
$178K - $192.2K
2% of jobs
$36.7K
$120K
$192.2K
How much do ml inference jobs pay per year?
What is ML inference?
What is the difference between Ml Inference vs Data Scientist?
| Aspect | ML Inference | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of machine learning models, programming skills | Degree in data science, statistics, or related fields |
| Work Environment | Deploying models in production, real-time data processing | Data analysis, model development, research |
| Industry Usage | AI product deployment, software companies | Research institutions, tech firms, consulting |
ML Inference focuses on deploying trained models to make predictions on new data, often in real-time. Data Scientists develop and analyze models, working primarily in research and development. While both roles require understanding of machine learning, ML Inference emphasizes deployment and operationalization, whereas Data Scientists focus on model creation and analysis.
What are some common challenges faced by ML inference engineers when deploying models to production?
What are the key skills and qualifications needed to thrive in ML inference?
Is ML inference a high paying job?

Full-time
Re-posted 17 days ago
Job description
MDAEdge is a company specializing in AI and analytics solutions, seeking an AI/ML & Analytics Platform Engineer to contribute to their platform development. The role involves building capabilities for scalable AI/ML workflows and collaborating with cross-functional teams to enhance platform performance and deployment efficiency.
Responsibilities:
• Contribute to building AI/ML & Analytics platform, services, and tools across dev, test, and prod environments to accelerate model training, inference, and deployment.
• Build capabilities for batch and real-time workflows at scale with flexible deployment strategies for use cases like low-latency predictions and offline inference.
• Improve platform performance, reduce manual intervention, scale compute, and increase deployment efficiency.
• Collaborate with cloud teams to ensure operational effectiveness, reliability, security, and efficiency.
• Provide technical guidance on monitoring systems like registries and alerting, plus governance frameworks for regulatory compliance.
• Work with cross-functional teams on AI/ML system architecture, deployment pipelines, and solution scaling.
• Champion self-service patterns, IaC, and GitOps for platform development.
Qualifications:
Required:
• Bachelor's or Master's in Computer Science, Engineering, Data Science, Mathematics, Statistics, Operations Research, or related field.
• Experience building scalable AI/ML & Analytics platforms for ML Researchers, Engineers, Data Scientists, and Analysts.
• Proficiency in Python, Spark, SQL, and ML frameworks like PyTorch or TensorFlow.
• Strong AWS knowledge, including AI/ML services like SageMaker.
• IaC tools such as Terraform, OpenTofu, CDK, or Pulumi, plus CI/CD pipelines.
• Containerization with Docker or Podman, and orchestration with Kubernetes or Rancher.
• VCS like GitHub or GitLab, CI/CD tools like GitHub Actions or Jenkins, and JIRA.
• Ops fundamentals including registries, observability, monitoring, performance analysis, and cost optimization.
• Hands-on problem-solving for technical and architectural challenges in scalable, secure platforms.
• Automation-first mindset with security consciousness and focus on developer experience.
• Strong communication to engage stakeholders effectively.
• Ability to work collaboratively in cross-functional, agile teams valuing individual development.
Preferred:
• Pharma/biotech domain experience.
• Strongly typed languages like C/C++, Java, Go, or Rust.
• Large-scale distributed systems like Ray, Dask, Spark, or HPC like Slurm.
• Data platforms like Databricks, Snowflake, or dbt with Delta, Iceberg, Hudi.
• Real-time streaming like Kafka or Spark Streaming.
• GitOps tools like ArgoCD or Crossplane.
• Multi-cloud (AWS, GCP, Azure).
• High-performance inference frameworks like ONNX Runtime, TensorRT, or Triton.
• Large-scale CPU/GPU infrastructure with CUDA knowledge.
Company:
The world doesn't have a talent shortage. It has a talent alignment problem. MDA Edge exists to fix that. Founded in , the company is headquartered in Sheridan, WY, US, , with a team of 51-200 employees. The company is currently Growth Stage.