1

Ai Reliability Engineer Jobs in Randolph, NJ (NOW HIRING)

Site Reliability Engineer (SRE)

Parsippany, NJ ยท On-site

$57.25 - $76.25/hr

We are looking for a talented Site Reliability Engineer (SRE) with a strong background in Google ... Familiarity with Google BI and AI/ML tools a plus (Looker, BigQuery ML, Vertex AI, etc.) Experience ...

GCP Site Reliability Engineer Interview Mode: candidates local to Parsippany, Nj who can attend an ... Familiarity with Google BI and AI/ML tools (Looker, BigQuery ML, Vertex AI, etc.) Experience with ...

Senior SRE Engineer

Parsippany, NJ ยท On-site

$57.25 - $76.25/hr

Utilize AI-assisted engineering tools (GitHub Copilot, Cursor, ChatGPT, Claude Code, or similar) to ... Required Qualifications 8+ years of experience in Site Reliability Engineering, DevOps Engineering ...

Applied AI SRE III - PxE GPS

Morristown, NJ ยท On-site

$58.75 - $78/hr

Applied AI Site Reliability Engineer III Role Overview: As an Applied AI Site Reliability Engineer III , you will actively engage in your engineering craft, taking a hands-on approach to the ...

AppOpps Engineer

Berkeley Heights, NJ ยท On-site

$59.50 - $79/hr

Site Reliability Engineer, Google Cloud Engine AI SRE at Google: Focus specifically on AI workload health, and GCE visibility _____ Mandatory Technical Skills & Competencies * Experience: 8+ years in ...

Engineer

Morristown, NJ ยท On-site

$70K - $90K/yr

The Platform Reliability Engineer operates and improves enterprise platforms that support IT operations, automation, observability, AI-enabled workflows, and business-critical services. * The role ...

Data Reliability Engineer II

Basking Ridge, NJ ยท On-site

$58.75 - $78/hr

... AI platform for banking 5. Pixel , India's first digital-native credit card, launched in ... As a Data Reliability Engineer II, you will play a crucial role in developing, optimizing, and ...

... AI platform for banking 5. Pixel , India's first digital-native credit card, launched in ... As a Data Reliability Engineer II, you will play a crucial role in developing, optimizing, and ...

Data Reliability Engineer II

Basking Ridge, NJ ยท On-site

$58.75 - $78/hr

... AI platform for banking 5. Pixel , India's first digital-native credit card, launched in ... As a Data Reliability Engineer II, you will play a crucial role in developing, optimizing, and ...

next page

Showing results 1-20

Ai Reliability Engineer information

See Randolph, NJ salary details

$62.7K

$121.3K

$144.9K

How much do ai reliability engineer jobs pay per year?

As of Aug 10, 2026, the average yearly pay for ai reliability engineer in Randolph, NJ is $121,254.00, according to ZipRecruiter salary data. Most workers in this role earn between $105,400.00 and $132,600.00 per year, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive as an AI reliability engineer, and why are they important?

To thrive as an AI Reliability Engineer, you need a solid background in computer science or engineering, expertise in AI/ML concepts, and experience with software testing and reliability methodologies. Familiarity with tools like TensorFlow, PyTorch, CI/CD pipelines, and reliability testing frameworks, along with certifications in cloud platforms (e.g., AWS Certified Machine Learning), is highly valuable. Analytical thinking, problem-solving abilities, and strong collaboration skills set top performers apart in this role. These skills ensure robust, dependable AI systems that meet performance standards and maintain trust in critical applications.

What is the difference between Ai Reliability Engineer vs Data Scientist?

AspectAi Reliability EngineerData Scientist
Required CredentialsBachelor's or master's in CS, engineering, or related; certifications in AI/MLBachelor's or master's in CS, statistics, or related; certifications in data analysis or ML
Work EnvironmentTech companies, AI-focused teams, engineering departmentsResearch labs, tech firms, analytics teams
Employer & Industry UsageAI product development, machine learning systems, reliability testingData analysis, predictive modeling, business insights

While both roles involve AI and ML, Ai Reliability Engineers focus on ensuring AI system robustness and uptime, whereas Data Scientists analyze data to generate insights and models. The roles often collaborate but serve different primary functions within AI projects.

What is an AI reliability engineer?

AI Reliability Engineers are professionals responsible for ensuring that artificial intelligence systems function reliably, safely, and effectively over time. They work on monitoring AI models in production, identifying and mitigating potential failures, and improving the robustness of AI systems. Their tasks often include testing, validation, performance monitoring, and implementing best practices for maintaining AI infrastructure. By focusing on reliability, they help organizations deploy AI solutions that are dependable and trustworthy in real-world environments.

What are some common challenges AI reliability engineers face when ensuring model robustness in production environments?

Ai Reliability Engineers often encounter challenges such as monitoring AI model performance for drift or unexpected behavior, managing data quality issues, and implementing automated alerting systems for anomalies. In production, it's crucial to ensure that AI models operate consistently and remain reliable under varying conditions and data inputs. Collaborating closely with data scientists, software engineers, and DevOps teams is essential to address these challenges and to continuously improve model reliability and uptime.
What are popular job titles related to Ai Reliability Engineer jobs in Randolph, NJ? For Ai Reliability Engineer jobs in Randolph, NJ, the most frequently searched job titles are:
What cities near Randolph, NJ are hiring for Ai Reliability Engineer jobs? Cities near Randolph, NJ with the most Ai Reliability Engineer job openings:

Site Reliability Engineer (SRE)

SmartIPlace

Parsippany, NJ โ€ข On-site

$57.25 - $76.25/hr

Other

Re-posted 19 hours ago


Job description

We are looking for a talented Site Reliability Engineer (SRE) with a strong background in Google Cloud Platform (GCP), and RedHat OpenShift administration. The ideal candidate will be responsible for ensuring the reliability, performance, and scalability of our on-premise and cloud-based systems along with focus on reducing costs for Google Cloud.
System Reliability: Ensure the reliability and uptime of critical services and infrastructure.
Google Cloud Expertise: Design, implement, and manage cloud infrastructure using Google Cloud services.
Automation: Develop and maintain automation scripts and tools to improve system efficiency and reduce manual intervention.
Monitoring and Incident Response: Implement monitoring solutions and respond to incidents to minimize downtime and ensure quick recovery.
Collaboration: Work closely with development and operations teams to improve system reliability and performance.
Capacity Planning: Conduct capacity planning and performance tuning to ensure systems can handle future growth.
Documentation: Create and maintain comprehensive documentation for system configurations, processes, and procedures.

Qualifications:
Education: Bachelor's degree in computer science, Engineering, or a related field.
Experience: 3+ years of experience in site reliability engineering or a similar role.

Skills:
Proficiency in Google Cloud services (Compute Engine, Kubernetes Engine, Cloud Storage, AlloyDB, Pub/Sub, etc.).
Familiarity with Google BI and AI/ML tools a plus (Looker, BigQuery ML, Vertex AI, etc.)
Experience with automation tools (Terraform, Ansible, Puppet).
Familiarity with CI/CD pipelines and tools (Azure pipelines Jenkins, GitLab CI, etc.).
Strong scripting skills (Python, Bash, etc.).
Knowledge of networking concepts and protocols.
Experience with monitoring tools (Prometheus, Grafana, etc.).

Preferred Certifications:
Google Cloud Professional DevOps Engineer
Google Cloud Professional Cloud Architect
Red Hat Certified Engineer (RHCE) or similar Linux certification


Smart-iPlace logo

About Smart-iPlace

Sourced by ZipRecruiter

SMART-iPLACE provides innovative staffing and consulting solutions that help our clients achieve their business objectives. We can understand and support all areas of your IT systems from back-end infrastructure to front-end personal productivity. Our goal is create innovative IT solutions that enable your business to be more agile and competitive.

Industry

It services

Company size

51 - 200 Employees

Headquarters location

Irving, TX, US

Year founded

2021

Social media