AI/ML Engineer (MLops)
Woodbridge, NJ ยท On-site
Deploying real-time ML inference pipelines processing millions of records at high throughput * Experience building end to end automated MLOps capabilities along with model and feature drift ...
Woodbridge, NJ ยท On-site
Deploying real-time ML inference pipelines processing millions of records at high throughput * Experience building end to end automated MLOps capabilities along with model and feature drift ...
Woodbridge, NJ ยท On-site
Deploying real-time ML inference pipelines processing millions of records at high throughput * Experience building end to end automated MLOps capabilities along with model and feature drift ...
New York, NY ยท On-site
$220K - $485K/yr
ML compilers and framework internals: PyTorch internals, torch.compile, custom operators ... Low-precision inference: INT8/FP8/FP4 quantization, mixed-precision serving. * Profiling and ...
New York, NY ยท On-site
$220K - $485K/yr
ML compilers and framework internals: PyTorch internals, torch.compile, custom operators ... Low-precision inference: INT8/FP8/FP4 quantization, mixed-precision serving. * Profiling and ...
New York, NY ยท On-site +1
$165K - $330K/yr
Develop infrastructure components for our ML inference platform using Python and Go * Implement and maintain Kubernetes deployments for model serving * Contribute to our inference orchestration layer ...
New York, NY ยท On-site +1
$165K - $330K/yr
Develop infrastructure components for our ML inference platform using Python and Go * Implement and maintain Kubernetes deployments for model serving * Contribute to our inference orchestration layer ...
Manhattan, NY ยท Hybrid
$145K - $180K/yr
The ML Engineer has hands-on experience building and optimizing ML inference systems that run in production environments. This role will develop and tune pipelines that transform millions of photos ...
Manhattan, NY ยท Hybrid
$145K - $180K/yr
The ML Engineer has hands-on experience building and optimizing ML inference systems that run in production environments. This role will develop and tune pipelines that transform millions of photos ...
Manhattan, NY ยท On-site
$145K - $180K/yr
The ML Engineer has hands-on experience building and optimizing ML inference systems that run in production environments. This role will develop and tune pipelines that transform millions of photos ...
Manhattan, NY ยท On-site
$145K - $180K/yr
The ML Engineer has hands-on experience building and optimizing ML inference systems that run in production environments. This role will develop and tune pipelines that transform millions of photos ...
New York, NY ยท On-site
$118K - $161K/yr
Hands-on experience with managed ML inference and serving platforms such as AWS SageMaker and GCP Vertex AI. * A proven track record operating inference at large scale across a range of model types ...
New York, NY ยท On-site
$118K - $161K/yr
Hands-on experience with managed ML inference and serving platforms such as AWS SageMaker and GCP Vertex AI. * A proven track record operating inference at large scale across a range of model types ...
Manhattan, NY ยท On-site
$213K - $263K/yr
Work with model developers to tune their neural networks for better inference efficiency and ... working with ML inference or linear algebra computation * C++ programming skills, including ...
Manhattan, NY ยท On-site
$213K - $263K/yr
Work with model developers to tune their neural networks for better inference efficiency and ... working with ML inference or linear algebra computation * C++ programming skills, including ...
Oversee technical strategy and architecture in ML Inference Platform services * Design, implement, and scale critical engineering components and services to support ML inference and deployment * Work ...
Oversee technical strategy and architecture in ML Inference Platform services * Design, implement, and scale critical engineering components and services to support ML inference and deployment * Work ...
New York, NY ยท On-site
$147K - $198K/yr
Oversee technical strategy and architecture in ML Inference Platform services * Design, implement, and scale critical engineering components and services to support ML inference and deployment * Work ...
New York, NY ยท On-site
$147K - $198K/yr
Oversee technical strategy and architecture in ML Inference Platform services * Design, implement, and scale critical engineering components and services to support ML inference and deployment * Work ...
New York, NY ยท On-site
$165K - $225K/yr
Can be anywhere in that lifecycle, from training machine learning systems, to creating the interfaces users use to navigate inference results. * An interest and passion for AI/ML systems, if you ...
New York, NY ยท On-site
$165K - $225K/yr
Can be anywhere in that lifecycle, from training machine learning systems, to creating the interfaces users use to navigate inference results. * An interest and passion for AI/ML systems, if you ...
New York, NY ยท On-site
$165K - $225K/yr
Can be anywhere in that lifecycle, from training machine learning systems, to creating the interfaces users use to navigate inference results. * An interest and passion for AI/ML systems, if you ...
Quick apply
New York, NY ยท On-site
$165K - $225K/yr
Can be anywhere in that lifecycle, from training machine learning systems, to creating the interfaces users use to navigate inference results. * An interest and passion for AI/ML systems, if you ...
Optimize and tune kernels and compiled code to achieve latency targets for ML inference * Conduct design and code reviews. Evaluate code performance, debug, diagnose and drive resolution of compiler ...
Optimize and tune kernels and compiled code to achieve latency targets for ML inference * Conduct design and code reviews. Evaluate code performance, debug, diagnose and drive resolution of compiler ...
New York, NY ยท On-site
$180K - $360K/yr
Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure. * Deep dive into ...
New York, NY ยท On-site
$180K - $360K/yr
Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure. * Deep dive into ...
Livingston, NJ ยท On-site
$140K - $182K/yr
The AI/ML TPM team owns delivery and execution across CoreWeave's AI/ML Platform Services ... The Inference team is responsible for building and operating highly scalable, reliable production ...
Livingston, NJ ยท On-site
$140K - $182K/yr
The AI/ML TPM team owns delivery and execution across CoreWeave's AI/ML Platform Services ... The Inference team is responsible for building and operating highly scalable, reliable production ...
$141K - $182K/yr
The AI/ML TPM team owns delivery and execution across CoreWeave's AI/ML Platform Services ... The Inference team is responsible for building and operating highly scalable, reliable production ...
Quick apply
$141K - $182K/yr
The AI/ML TPM team owns delivery and execution across CoreWeave's AI/ML Platform Services ... The Inference team is responsible for building and operating highly scalable, reliable production ...
$140K - $182K/yr
The AI/ML TPM team owns delivery and execution across CoreWeave's AI/ML Platform Services ... The Inference team is responsible for building and operating highly scalable, reliable production ...
Quick apply
$140K - $182K/yr
The AI/ML TPM team owns delivery and execution across CoreWeave's AI/ML Platform Services ... The Inference team is responsible for building and operating highly scalable, reliable production ...
New York, NY ยท On-site
$134K - $176K/yr
What You'll Own - Real-Time Proxy Bidding Engine (Tier 1) - Define and lead the engine that computes per-request bids at sub-millisecond latency, build the ML inference infrastructure serving model ...
New York, NY ยท On-site
$134K - $176K/yr
What You'll Own - Real-Time Proxy Bidding Engine (Tier 1) - Define and lead the engine that computes per-request bids at sub-millisecond latency, build the ML inference infrastructure serving model ...
New York, NY ยท On-site
$134K - $176K/yr
What You'll Own - Real-Time Proxy Bidding Engine (Tier 1) - Define and lead the engine that computes per-request bids at sub-millisecond latency, build the ML inference infrastructure serving model ...
New York, NY ยท On-site
$134K - $176K/yr
What You'll Own - Real-Time Proxy Bidding Engine (Tier 1) - Define and lead the engine that computes per-request bids at sub-millisecond latency, build the ML inference infrastructure serving model ...
New York, NY ยท On-site
$134K - $176K/yr
What You'll Own - Real-Time Proxy Bidding Engine (Tier 1) - Define and lead the engine that computes per-request bids at sub-millisecond latency, build the ML inference infrastructure serving model ...
New York, NY ยท On-site
$134K - $176K/yr
What You'll Own - Real-Time Proxy Bidding Engine (Tier 1) - Define and lead the engine that computes per-request bids at sub-millisecond latency, build the ML inference infrastructure serving model ...
About the Role We're looking for a AI/ML Engineer (Senior/Staff/Principal) - Threat Detection who will design, build, and operationalize the detection algorithms, ML inference pipelines, and risk ...
Quick apply
About the Role We're looking for a AI/ML Engineer (Senior/Staff/Principal) - Threat Detection who will design, build, and operationalize the detection algorithms, ML inference pipelines, and risk ...
| Aspect | ML Inference | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of machine learning models, programming skills | Degree in data science, statistics, or related fields |
| Work Environment | Deploying models in production, real-time data processing | Data analysis, model development, research |
| Industry Usage | AI product deployment, software companies | Research institutions, tech firms, consulting |
ML Inference focuses on deploying trained models to make predictions on new data, often in real-time. Data Scientists develop and analyze models, working primarily in research and development. While both roles require understanding of machine learning, ML Inference emphasizes deployment and operationalization, whereas Data Scientists focus on model creation and analysis.
AI/ML Engineer (MLOps)
Iselin, NJ
12 months
Technical Expertise
ML/Al Frameworks: PyTorch, TensorFlow, JAX, HuggingFace, LangChain, LangGraph, Llamalndex, DSPy, ONNX Runtime, TensorRT
GenAl & LLMs: GPT-4/Claude API, LoRA/QLoRA fine-tuning, RAG (FAISS, Pinecone, ChromaDB), prompt engineering, agentic orchestration Languages: Python, C/C++, Java, SQL, Scala, GoLang, JavaScript
Cloud & Infra: AWS (SageMaker, S3, Lambda), Google Cloud Platform (GKE, Vertex Al, BigQuery), Kubernetes, Terraform, Docker
Databases: PostgreSQL, MySQL, MongoDB, Neo4j, BigQuery, Pinecone, ChromaDB, Redis Libraries: Pandas, NumPy, Scikit-learn, OpenCV, Keras, Spark, Kafka Dev
Tools: Linux, Git, Docker, Kubernetes, Jenkins
Experience Required
Education : At least a bachelor s degree (or equivalent experience) in Computer Science, Software Engineering, Electronics Engineering, Information Systems, or a closely related field is required for the project