Infrastructure Engineering Location / Remote Policy: Remote Role Type: Contract - 6-month initial ... This role is central to supporting AI/ML infrastructure at scale - enabling efficient training and ...
Infrastructure Engineering Location / Remote Policy: Remote Role Type: Contract - 6-month initial ... This role is central to supporting AI/ML infrastructure at scale - enabling efficient training and ...
Develop and maintain high-performance ML infrastructure components in C++ and Python, ensuring ... Mentor other engineers on ML infrastructure best practices, debugging methodologies, and ...
Develop and maintain high-performance ML infrastructure components in C++ and Python, ensuring ... Mentor other engineers on ML infrastructure best practices, debugging methodologies, and ...
ML Engineer
Manhattan, NY · On-site
... the ML infrastructure and processes for scalability and performance. Qualifications : Required ... ML engineering. • Strong programming skills in Python (TypeScript experience is a plus). • ...
ML Engineer
Manhattan, NY · On-site
... the ML infrastructure and processes for scalability and performance. Qualifications : Required ... ML engineering. • Strong programming skills in Python (TypeScript experience is a plus). • ...
ML Engineer
Manhattan, NY · On-site
... the ML infrastructure and processes for scalability and performance. Qualifications : Required ... ML engineering. • Strong programming skills in Python (TypeScript experience is a plus). • ...
ML Engineer
Manhattan, NY · On-site
... the ML infrastructure and processes for scalability and performance. Qualifications : Required ... ML engineering. • Strong programming skills in Python (TypeScript experience is a plus). • ...
AI Infrastructure Engineer
New York, NY · Remote
$140K - $165K/yr
As an AI Infrastructure Engineer, your role will include: * Lead Technical Deployments: Drive end ... AI/ML Familiarity: Experience with inference serving, GPU scheduling, and the tooling around LLM ...
Quick apply
AI Infrastructure Engineer
New York, NY · Remote
$140K - $165K/yr
As an AI Infrastructure Engineer, your role will include: * Lead Technical Deployments: Drive end ... AI/ML Familiarity: Experience with inference serving, GPU scheduling, and the tooling around LLM ...
AI/ML Engineer
New York, NY · On-site
What you will do As our ML Engineer, you'll be the technical backbone powering our content platform ... Designing scalable ML infrastructure and pipelines that handle massive media datasets
AI/ML Engineer
New York, NY · On-site
What you will do As our ML Engineer, you'll be the technical backbone powering our content platform ... Designing scalable ML infrastructure and pipelines that handle massive media datasets
Staff Engineer - Content Intelligence Infrastructure
New York, NY · On-site
$203K - $290K/yr
We're now looking for a Staff Engineer to help build and scale foundational infrastructure powering ... This role sits at the intersection of backend systems, ML infrastructure, distributed data systems ...
Staff Engineer - Content Intelligence Infrastructure
New York, NY · On-site
$203K - $290K/yr
We're now looking for a Staff Engineer to help build and scale foundational infrastructure powering ... This role sits at the intersection of backend systems, ML infrastructure, distributed data systems ...
... learning infrastructure. Responsibilities : • Design, develop, test, deploy, maintain, and ... engineers. • Lead the design of GenAI solutions, optimize ML infrastructure, and guide the ...
... learning infrastructure. Responsibilities : • Design, develop, test, deploy, maintain, and ... engineers. • Lead the design of GenAI solutions, optimize ML infrastructure, and guide the ...
Staff Engineer - Content Intelligence Infrastructure
New York, NY · On-site
$203K - $290K/yr
We're now looking for a Staff Engineer to help build and scale foundational infrastructure powering ... This role sits at the intersection of backend systems, ML infrastructure, distributed data systems ...
Staff Engineer - Content Intelligence Infrastructure
New York, NY · On-site
$203K - $290K/yr
We're now looking for a Staff Engineer to help build and scale foundational infrastructure powering ... This role sits at the intersection of backend systems, ML infrastructure, distributed data systems ...
Senior Software Engineer, Infrastructure
New York, NY · On-site
$160K - $220K/yr
Our team has built SW, ML, and HW products across Meta, CTRL-labs, Google, Apple, Fitbit, Peloton ... About We're looking for an experienced infrastructure engineer to help build Stream, a new ...
Senior Software Engineer, Infrastructure
New York, NY · On-site
$160K - $220K/yr
Our team has built SW, ML, and HW products across Meta, CTRL-labs, Google, Apple, Fitbit, Peloton ... About We're looking for an experienced infrastructure engineer to help build Stream, a new ...
Azure/Databricks Infrastructure Engineer
$57.25 - $76.75/hr
Experience supporting enterprise analytics, ETL, Snowflake, or modern data engineering ecosystems Experience supporting AI/ML or advanced analytics infrastructure environments Azure and/or Databricks ...
Azure/Databricks Infrastructure Engineer
$57.25 - $76.75/hr
Experience supporting enterprise analytics, ETL, Snowflake, or modern data engineering ecosystems Experience supporting AI/ML or advanced analytics infrastructure environments Azure and/or Databricks ...
Azure/Databricks Infrastructure Engineer
$61 - $81.50/hr
Experience supporting enterprise analytics, ETL, Snowflake, or modern data engineering ecosystems Experience supporting AI/ML or advanced analytics infrastructure environments Azure and/or Databricks ...
Azure/Databricks Infrastructure Engineer
$61 - $81.50/hr
Experience supporting enterprise analytics, ETL, Snowflake, or modern data engineering ecosystems Experience supporting AI/ML or advanced analytics infrastructure environments Azure and/or Databricks ...
Simulation Infrastructure Engineer
New York, NY · On-site
$117K - $154K/yr
... infrastructure projects. We're not here debating the future of AI. We're deploying it in the real ... the ML engineers who depend on sim are as efficient as they can possibly be. You'll sit on the ...
Simulation Infrastructure Engineer
New York, NY · On-site
$117K - $154K/yr
... infrastructure projects. We're not here debating the future of AI. We're deploying it in the real ... the ML engineers who depend on sim are as efficient as they can possibly be. You'll sit on the ...
Forward Deployed Engineer - ML
New York, NY · On-site
$180K - $250K/yr
AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new ... The Role: We're looking for Forward Deployed ML Engineers who want to work at the intersection of ...
Forward Deployed Engineer - ML
New York, NY · On-site
$180K - $250K/yr
AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new ... The Role: We're looking for Forward Deployed ML Engineers who want to work at the intersection of ...
Staff Backend Engineer
New York, NY · On-site
$200K - $250K/yr
Join an elite team as a Senior Backend Engineer architecting the high-performance data backbone and ML infrastructure that powers autonomous manufacturing. You will build the low-latency systems that ...
Quick apply
Staff Backend Engineer
New York, NY · On-site
$200K - $250K/yr
Join an elite team as a Senior Backend Engineer architecting the high-performance data backbone and ML infrastructure that powers autonomous manufacturing. You will build the low-latency systems that ...
Senior AI/ML Engineer
New York, NY · On-site
$240K - $270K/yr
As an AI/ML Engineer, you'll join a growing team focused on building the AI foundation that will ... Develop and scale AI/ML infrastructure that powers both internal tooling and customer-facing ...
Senior AI/ML Engineer
New York, NY · On-site
$240K - $270K/yr
As an AI/ML Engineer, you'll join a growing team focused on building the AI foundation that will ... Develop and scale AI/ML infrastructure that powers both internal tooling and customer-facing ...
SRE Manager, ML Operations
New York, NY · On-site
$62.25 - $82.75/hr
... Engineering team, with a focus on ML Operations ... This team owns the reliability, performance, and scalability of the Ad Serving infrastructure that ...
SRE Manager, ML Operations
New York, NY · On-site
$62.25 - $82.75/hr
... Engineering team, with a focus on ML Operations ... This team owns the reliability, performance, and scalability of the Ad Serving infrastructure that ...
Staff AI/ML Engineer
$240K - $270K/yr
As an AI/ML Engineer, you'll join a growing team focused on building the AI foundation that will ... Develop and scale AI/ML infrastructure that powers both internal tooling and customer-facing ...
Staff AI/ML Engineer
$240K - $270K/yr
As an AI/ML Engineer, you'll join a growing team focused on building the AI foundation that will ... Develop and scale AI/ML infrastructure that powers both internal tooling and customer-facing ...
AI/ML Engineer Intern
New York, NY · On-site
$18.25 - $23.75/hr
Designing scalable ML infrastructure and pipelines that handle massive media datasets ... Hands-on ML engineering experience building production systems at Big Tech companies, high-growth ...
AI/ML Engineer Intern
New York, NY · On-site
$18.25 - $23.75/hr
Designing scalable ML infrastructure and pipelines that handle massive media datasets ... Hands-on ML engineering experience building production systems at Big Tech companies, high-growth ...
Staff Infrastructure Engineer
New York, NY · On-site
$225K - $450K/yr
... engineering. Today, we help our customers reduce Snowflake and Databricks SQL compute costs by up ... Build infrastructure and data pipelines that support ML modeling and product features * Improve ...
Staff Infrastructure Engineer
New York, NY · On-site
$225K - $450K/yr
... engineering. Today, we help our customers reduce Snowflake and Databricks SQL compute costs by up ... Build infrastructure and data pipelines that support ML modeling and product features * Improve ...
Ml Infrastructure Engineer information
See Edison, NJ salary details
$47.2K - $59.6K
1% of jobs
$59.6K - $72.1K
1% of jobs
$72.1K - $84.6K
4% of jobs
$84.6K - $97.1K
9% of jobs
$97.1K - $109.6K
10% of jobs
$110.3K is the 25th percentile. Wages below this are outliers.
$109.6K - $122.1K
10% of jobs
The median wage is $127.2K / yr.
$122.1K - $134.6K
39% of jobs
$137.3K is the 75th percentile. Wages above this are outliers.
$134.6K - $147.1K
7% of jobs
$147.1K - $159.6K
10% of jobs
$159.6K - $172.1K
9% of jobs
$172.1K - $184.6K
1% of jobs
$47.2K
$128.9K
$184.6K
How much do ml infrastructure engineer jobs pay per year?
What is the difference between Ml Infrastructure Engineer vs Data Engineer?
| Aspect | ML Infrastructure Engineer | Data Engineer |
|---|---|---|
| Required Credentials | Bachelor's/Master's in CS, experience with cloud platforms, scripting, and ML tools | Bachelor's/Master's in CS, experience with databases, ETL, and data pipelines |
| Work Environment | Focus on deploying and maintaining ML systems, cloud infrastructure, and automation | Designing and building data pipelines, managing large datasets, and data storage |
| Employer & Industry Usage | Tech companies, AI startups, research labs | Finance, healthcare, e-commerce, and data-driven industries |
The ML Infrastructure Engineer specializes in building and maintaining the infrastructure that supports machine learning models, focusing on deployment, scalability, and automation. In contrast, Data Engineers primarily develop data pipelines and manage large datasets to enable data analysis and business intelligence. Both roles require strong technical skills and often overlap, but their core focus areas differ significantly.
What are popular job titles related to Ml Infrastructure Engineer jobs in Edison, NJ?
For Ml Infrastructure Engineer jobs in Edison, NJ, the most frequently searched job titles are:
What job categories do people searching Ml Infrastructure Engineer jobs in Edison, NJ look for?
The top searched job categories for Ml Infrastructure Engineer jobs in Edison, NJ are:
What cities near Edison, NJ are hiring for Ml Infrastructure Engineer jobs?
Cities near Edison, NJ with the most Ml Infrastructure Engineer job openings:

NVIDIA AI Infrastructure & Kubernetes Platform Engineer (DGX Systems)
New York, NY • On-site
$125/hr
Contractor
Medical, Retirement, PTO
Posted 11 days ago
Job description
Department: Infrastructure Engineering
Location / Remote Policy: Remote
Role Type: Contract - 6-month initial engagement
About Our Client
Our client is a technology and professional services firm founded in 2015 on the strength of its founders' 30 years of industry experience. They set out to bridge a gap in professional services - to be a true partner rather than just a vendor - delivering expert guidance, innovative solutions, and personalized service at a cost-effective rate. Their mission is to empower businesses to succeed in the digital era, harnessing technology to drive transformation, innovation, and growth. Guided by a "make a customer, not a sale" philosophy, they lead with a customer-first approach and a team of senior-level engineers sourced from the world's leading OEMs, including AWS, Palo Alto Networks, Cisco, and Microsoft.
Job Description
Our client is seeking a highly skilled AI Infrastructure & Kubernetes Platform Engineer with a proven track record deploying and managing NVIDIA DGX-based AI clusters, orchestrating containerized AI workloads on Kubernetes, and ensuring secure, high-throughput operations across InfiniBand-powered networks. You'll bring a strong certification foundation across both Kubernetes (CKA, CKAD, CKS) and NVIDIA's AI infrastructure stack, paired with hands-on experience across DGX, BlueField, and high-speed networking.
This role is central to supporting AI/ML infrastructure at scale - enabling efficient training and inference for complex models and integrating NVIDIA's compute, storage, and fabric solutions with modern DevOps practices. Day to day, you'll own DGX cluster operations, architect GPU-accelerated Kubernetes platforms, tune InfiniBand fabric for throughput, and harden the environment through DPU-enhanced security.
You'll work at the intersection of infrastructure, DevOps, and AI/ML, keeping the platform reliable and cost-efficient for the teams that depend on it. The ideal candidate is deeply hands-on, obsessed with performance and security, and energized by operating some of the most advanced AI compute available.
Duties and Responsibilities
AI Infrastructure Operations
- Deploy and manage NVIDIA DGX BasePODs and SuperPODs for high-performance AI workloads.
- Oversee DGX system lifecycle operations, including provisioning, monitoring, firmware upgrades, and capacity planning.
- Operate Base Command Manager to manage GPU clusters, schedule workloads, and integrate with MLOps tools.
- Perform DGX node health validation, NCCL interconnect testing, and NVLink topology verification after deployments or hardware changes.
Kubernetes Platform Engineering
- Architect secure, scalable Kubernetes clusters optimized for GPU-accelerated workloads using the NVIDIA GPU Operator.
- Apply CKA/CKAD/CKS expertise to develop, deploy, and secure AI applications on Kubernetes.
- Implement CI/CD pipelines and GitOps methodologies for deploying and managing ML workflows.
High-Performance Networking & DPUs
- Administer InfiniBand networks and BlueField DPUs using Unified Fabric Manager (UFM).
- Enable NVLink/NVSwitch performance across GPU nodes and tune fabric configurations for minimal latency and maximum throughput.
- Use BlueField to offload storage, firewalling, and telemetry, strengthening AI workload security and performance.
Security & Compliance
- Apply CKS best practices to secure containerized AI environments.
- Configure runtime security, secrets management, network segmentation, and auditing across DPU-enhanced Kubernetes deployments.
- Support zero-trust initiatives by enforcing workload identity, RBAC policies, and supply-chain integrity across AI container images and model artifacts.
Monitoring, Telemetry & Optimization
- Monitor GPU, CPU, and I/O performance using NVIDIA DCGM, Prometheus, Grafana, and Base Command APIs.
- Tune system performance and model-training pipelines for cost-efficiency and throughput.
- Build and maintain operational runbooks, incident-response playbooks, and SLA dashboards covering GPU utilization, thermal thresholds, and fabric health.
Required Experience/Skills
Certifications
- Certified Kubernetes Administrator (CKA)
- Certified Kubernetes Application Developer (CKAD)
- Certified Kubernetes Security Specialist (CKS)
- NVIDIA Certified Associate: AI Infrastructure & Operations (NCA-AIIO)
- NVIDIA Certified Professional: AI Infrastructure (NCP-AII)
- NVIDIA Certified Professional: AI Operations (NCP-AIO)
- NVIDIA Certified Professional: AI Networking (NCP-AIN)
Hands-On Expertise
- DGX System, BasePOD, and SuperPOD administration
- BlueField DPU configuration and operations
- InfiniBand fabric and UFM management
- Base Command Manager for workload orchestration
Technical Skills
- Kubernetes, Helm, and the NVIDIA GPU Operator
- DevOps tooling: Ansible, Terraform, GitOps, CI/CD pipelines
- Programming/scripting: Python, YAML, Bash
Nice-to-Haves
- Kubeflow and broader MLOps pipeline experience.
- Parallel/HPC storage: NFS, BeeGFS, Lustre.
- Advanced networking: RoCE, RDMA, gRPC, and DPU offload tuning.
Education
Bachelor's degree in Computer Science, Engineering, or a related field - or equivalent hands-on experience.
Pay & Benefits Summary
- Pay Rate: $125/hr
- Benefits: Eligible for a comprehensive benefits package, including medical/health coverage, paid time off, and 401(k) retirement savings.
Call-to-Action
Operate the cutting edge of AI compute. Apply today and put your DGX, Kubernetes, and NVIDIA expertise to work.
Keywords: NVIDIA DGX | SuperPOD | Kubernetes | CKA / CKAD / CKS | NCA-AIIO | NCP-AII | NCP-AIO | NCP-AIN | InfiniBand | BlueField DPU | UFM | GPU Operator | Kubeflow | RDMA | RoCE | MLOps
About Catapult Solutions Group
Sourced by ZipRecruiter
Industry
Recruiting and staffing services
Company size
201 - 500 Employees
Headquarters location
Plano, TX, US
Year founded
2013