AI/ML Platform Engineer
Alexandria, VA ยท On-site
FastAPI and microservices for ML inference * InfrastructureasCode (Terraform) * Kubernetes and Docker for scalable ML workloads * Distributed/cloud systems design with AWS * Edgetocloud system ...
Alexandria, VA ยท On-site
FastAPI and microservices for ML inference * InfrastructureasCode (Terraform) * Kubernetes and Docker for scalable ML workloads * Distributed/cloud systems design with AWS * Edgetocloud system ...
Alexandria, VA ยท On-site
FastAPI and microservices for ML inference * InfrastructureasCode (Terraform) * Kubernetes and Docker for scalable ML workloads * Distributed/cloud systems design with AWS * Edgetocloud system ...
$100K - $160K/yr
Integration of on-device ML inference frameworks (Core ML, ONNX Runtime, WhisperKit) - loading, lifecycle, error recovery - not training the models * Real-time audio pipeline: AVAudioEngine, voice ...
$100K - $160K/yr
Integration of on-device ML inference frameworks (Core ML, ONNX Runtime, WhisperKit) - loading, lifecycle, error recovery - not training the models * Real-time audio pipeline: AVAudioEngine, voice ...
Optimize highly parallel numerical operations and ML inference algorithms for specialized hardware ... accelerators. * Lead technical direction and provide engineering mentorship for groups developing ...
Optimize highly parallel numerical operations and ML inference algorithms for specialized hardware ... accelerators. * Lead technical direction and provide engineering mentorship for groups developing ...
$142K - $213K/yr
Golang** for ML and AI platform use cases. * Develop **REST and gRPC APIs** for inference, processing pipelines, orchestration, and platform services. * Implement asynchronous and distributed ...
$142K - $213K/yr
Golang** for ML and AI platform use cases. * Develop **REST and gRPC APIs** for inference, processing pipelines, orchestration, and platform services. * Implement asynchronous and distributed ...
Washington, DC ยท On-site
$142.65 - $213.98/hr
Integrate ML inference services into backend workflows with attention to latency, throughput, and cost. * Work closely with ML engineers and data scientists to productionize models and pipelines.
Washington, DC ยท On-site
$142.65 - $213.98/hr
Integrate ML inference services into backend workflows with attention to latency, throughput, and cost. * Work closely with ML engineers and data scientists to productionize models and pipelines.
Reston, VA ยท On-site
$67.25 - $88.50/hr
... ML inference solutions using microservices and event-driven patterns (batch and real-time) Define CI/CD patterns for ML and integrate with enterprise DevOps tooling Partner with business and ...
Reston, VA ยท On-site
$67.25 - $88.50/hr
... ML inference solutions using microservices and event-driven patterns (batch and real-time) Define CI/CD patterns for ML and integrate with enterprise DevOps tooling Partner with business and ...
Reston, VA ยท On-site
AI/ML Engineer Location: Reston VA Core Responsibilities (AI/ML, Python, AWS, GenAI) Design and ... Build robust MLOps workflows, including model versioning, containerized training/inference ...
Quick apply
Reston, VA ยท On-site
AI/ML Engineer Location: Reston VA Core Responsibilities (AI/ML, Python, AWS, GenAI) Design and ... Build robust MLOps workflows, including model versioning, containerized training/inference ...
Reston, VA ยท On-site
AI/ML Engineer Location: Reston VA - In person interviews so need Local In EAST coast onlyโ Core ... Build robust MLOps workflows, including model versioning, containerized training/inference ...
Quick apply
Reston, VA ยท On-site
AI/ML Engineer Location: Reston VA - In person interviews so need Local In EAST coast onlyโ Core ... Build robust MLOps workflows, including model versioning, containerized training/inference ...
Mclean, VA ยท On-site +1
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site +1
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site +1
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site +1
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Reston, VA ยท On-site
AI/ML Engineer (Python, AWS, GenAI) Location: Reston, VA (In-person interviews required) Candidate ... Build and support MLOps workflows: model versioning, containerized training/inference, automated ...
Quick apply
Reston, VA ยท On-site
AI/ML Engineer (Python, AWS, GenAI) Location: Reston, VA (In-person interviews required) Candidate ... Build and support MLOps workflows: model versioning, containerized training/inference, automated ...
Mclean, VA ยท On-site +1
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site +1
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site +1
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site +1
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Mclean, VA ยท On-site
Build, harden, and operate ML platforms and inference services, including low-latency real-time scoring, batch inference, model packaging and containerization, autoscaling, canary and shadow ...
Optimize highly parallel numerical operations and ML inference algorithms for specialized hardware ... accelerators. * Lead technical direction and provide engineering mentorship for groups developing ...
Optimize highly parallel numerical operations and ML inference algorithms for specialized hardware ... accelerators. * Lead technical direction and provide engineering mentorship for groups developing ...
Reston, VA ยท On-site
Core Responsibilities (AI/ML, Python, AWS, GenAI) * Design and implement end-to-end AI/ML and ... Build robust MLOps workflows, including model versioning, containerized training/inference ...
Quick apply
Reston, VA ยท On-site
Core Responsibilities (AI/ML, Python, AWS, GenAI) * Design and implement end-to-end AI/ML and ... Build robust MLOps workflows, including model versioning, containerized training/inference ...
Integrate ML inference services into backend workflows with attention to latency, throughput, and cost . * Work closely with ML engineers and data scientists to productionize models and pipelines.
Integrate ML inference services into backend workflows with attention to latency, throughput, and cost . * Work closely with ML engineers and data scientists to productionize models and pipelines.
$38.6K - $53.5K
2% of jobs
$53.5K - $68.4K
3% of jobs
$68.4K - $83.3K
6% of jobs
$83.3K - $98.2K
9% of jobs
$103K is the 25th percentile. Wages below this are outliers.
$98.2K - $113.1K
15% of jobs
The median wage is $123K / yr.
$113.1K - $128K
22% of jobs
$136.1K is the 75th percentile. Wages above this are outliers.
$128K - $142.8K
32% of jobs
$142.8K - $157.7K
3% of jobs
$157.7K - $172.6K
4% of jobs
$172.6K - $187.5K
1% of jobs
$187.5K - $202.4K
2% of jobs
$38.6K
$126.4K
$202.4K
| Aspect | ML Inference | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of machine learning models, programming skills | Degree in data science, statistics, or related fields |
| Work Environment | Deploying models in production, real-time data processing | Data analysis, model development, research |
| Industry Usage | AI product deployment, software companies | Research institutions, tech firms, consulting |
ML Inference focuses on deploying trained models to make predictions on new data, often in real-time. Data Scientists develop and analyze models, working primarily in research and development. While both roles require understanding of machine learning, ML Inference emphasizes deployment and operationalization, whereas Data Scientists focus on model creation and analysis.
Cities near Fort Washington, MD with the most Ml Inference job openings:
Full-time
Re-posted 16 days ago
We are seeking a hands-on Senior AI/ML Platform Engineer with 10+ years of IT experience and a strong track record of building, deploying, and operationalizing AI/ML systems. The ideal candidate is a doer who excels in implementing scalable, production-grade AI/ML solutions across cloud environments.
Core Requirements