2

Remote Machine Learning Ops Engineer Jobs in New York

DevOps Engineer

Newark, NJ · Remote

$79.21 - $104.97/hr

Day - 08 Hour (United States of America) We are seeking a high-caliber Senior AI Platform & ML Ops ... This group is tasked to innovate, build, deploy and monitor production grade AI, machine learning ...

Career Renew is recruiting for one of its clients a Senior Machine Learning Engineer - this is a fully remote role for US/Canada based candidates. Salary range: 165-225K USD yearly plus benefits plus ...

Senior Machine Learning Engineer

New York, NY · On-site +1

$180K - $250K/yr

The Role As a Senior Machine Learning Engineer at Orita, you will: * Build and Productionize Models : Design, train, and deploy models that directly power our marketing-focused products, primarily ...

We are seeking a Senior Machine Learning Engineer to join our AI & ML team in New York City. You will play a key role in maturing and scaling our machine learning infrastructure, ensuring the ...

Showing results 21-40

Remote Machine Learning Ops Engineer information

What job categories do people searching Remote Machine Learning Ops Engineer jobs in New York look for? The top searched job categories for Remote Machine Learning Ops Engineer jobs in New York are:
What cities in New York are hiring for Remote Machine Learning Ops Engineer jobs? Cities in New York with the most Remote Machine Learning Ops Engineer job openings:
Machine Learning Operations Engineer - Remote

Machine Learning Operations Engineer - Remote

NAVA Software Solutions

Jersey City, NJ • On-site, Remote

$76K - $102K/yr

Full-time

Posted 22 days ago


Job description

NAVA Software solutions is looking for a Machine Learning Operations Engineer
Details:
Machine Learning Operations (MLOps) Engineer - AWS (with LLM Focus)
Location: Remote work
Duration: 12 months

Responsibilities:
  • LLM-Optimized MLOps Infrastructure: Design and implement MLOps infrastructure on AWS tailored for LLMs, leveraging services like SageMaker, EC2 (with GPU instances), S3, ECS/EKS, Lambda, and more.
  • LLM Deployment Pipelines: Build and manage CI/CD pipelines specifically for LLM deployment, addressing unique challenges like model size, inference optimization, and versioning.
  • LLMOps Practices: Implement LLMOps best practices for monitoring model performance, drift detection, prompt management, and feedback loops for continuous improvement.
  • RESTful API Development: Design and develop RESTful APIs to expose LLM capabilities to other applications and services, ensuring scalability, security, and optimal performance.
  • Model Optimization: Apply techniques like quantization, distillation, and pruning to optimize LLM models for efficient inference on AWS infrastructure.
  • Monitoring and Observability: Establish comprehensive monitoring and alerting mechanisms to track LLM performance, latency, resource utilization, and potential biases.
  • Prompt Engineering and Management: Develop strategies for prompt engineering and management to enhance LLM outputs and ensure consistency and safety.
  • Collaboration: Work closely with data scientists, researchers, and software engineers to integrate LLM models into production systems effectively.
  • Cost Optimization: Continuously optimize LLMOps processes and infrastructure for cost-efficiency while maintaining high performance and reliability.

Qualifications:
  • Experience: 3+ years of experience in MLOps or a related field, with hands-on experience in deploying and managing LLMs.
  • AWS Expertise: Strong proficiency in AWS services relevant to MLOps and LLMs, including SageMaker, EC2 (with GPU instances), S3, ECS/EKS, Lambda, and API Gateway.
  • LLM Knowledge: Deep understanding of LLM architectures (e.g., Transformers), training techniques, and inference optimization strategies.
  • Programming Skills: Proficiency in Python and experience with infrastructure-as-code tools (e.g., Terraform, CloudFormation), REST API frameworks (e.g., Flask, FastAPI), and LLM libraries (e.g., Hugging Face Transformers).
  • Monitoring: Familiarity with monitoring and logging tools for LLMs, such as Prometheus, Grafana, and CloudWatch.
  • Containerization: Experience with Docker and container orchestration (e.g., Kubernetes, ECS) for LLM deployment.
  • Problem Solving: Excellent problem-solving and troubleshooting skills in the context of LLMs and MLOps.
  • Communication: Strong communication and collaboration skills to effectively work with cross-functional teams

NAVA Software Solutions logo

About NAVA Software Solutions

Sourced by ZipRecruiter

NAVA is a strategic partner for companies seeking to develop or customize software and products. Our team of experts leverages cutting-edge technology and deep industry knowledge to provide customized solutions that drive business success. Whether you're looking to improve your operations, increase efficiency, or bring a new product to market, NAVA has the expertise and resources to help you achieve your goals. Trust us to be your partner in software and product development.

Industry

It services

Company size

51 - 200 Employees

Headquarters location

Rocky Hill, CT, US

Social media