1

Reliability Analyst Jobs in New York (NOW HIRING)

Showing results 41-60

Reliability Analyst information

See New York salary details

$38.8K

$108.5K

$138.9K

How much do reliability analyst jobs pay per year?

As of Aug 17, 2026, the average yearly pay for reliability analyst in New York is $108,482.00, according to ZipRecruiter salary data. Most workers in this role earn between $78,800.00 and $138,400.00 per year, depending on experience, location, and employer.

What is a reliability analyst?

A Reliability Analyst evaluates system performance, identifies potential failures, and develops strategies to improve reliability. They analyze data, monitor equipment or processes, and recommend maintenance or design improvements. Their goal is to minimize downtime, reduce costs, and enhance efficiency by predicting and preventing failures. This role is common in industries such as manufacturing, energy, and technology. Strong analytical skills, knowledge of reliability methodologies, and proficiency in data analysis tools are essential for success in this position.

What does a reliability analyst do?

A Reliability Analyst typically spends their day analyzing equipment performance data, identifying trends or potential failure points, and developing strategies to enhance reliability. They often collaborate with engineering, maintenance, and operations teams to implement reliability improvement initiatives and conduct root cause analysis of failures. The role may also involve preparing technical reports, tracking key performance indicators, and recommending preventive maintenance schedules. This hands-on and data-driven position offers opportunities to directly impact operational efficiency and asset performance.

What skills and qualifications are needed to be a reliability analyst?

To thrive as a Reliability Analyst, you need strong analytical skills, a background in engineering or statistics, and experience with reliability modeling and data analysis. Familiarity with tools like Weibull analysis software, reliability-centered maintenance (RCM) systems, and certifications such as Certified Reliability Engineer (CRE) are often valuable. Strong problem-solving abilities, effective communication, and attention to detail help professionals excel in cross-functional teams. These skills are crucial for accurately identifying risks, improving system reliability, and ensuring the efficient operation of assets.

Infographic showing various Reliability Analyst job openings in New York as of August 2026, with employment types broken down into 1% As Needed, 79% Full Time, 15% Part Time, 1% Temporary, 3% Contract, and 1% Nights. Highlights an 92% Physical, 3% Hybrid, and 5% Remote job distribution, with an average salary of $108,482 per year, or $52.2 per hour.

Site Reliability Engineer (SRE)

Long Finch Technologies

Wallington, NJ

$59 - $78.50/hr

Full-time

Posted 5 days ago


Job description

Overview

We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.

The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.

Key Responsibilities

  • Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
  • Optimize, and support highly available VDI environments on Hyper-V.
  • Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
  • Disaster recovery, backup, patch management, and business continuity strategies.
  • Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
  • Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
  • Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
  • Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
  • Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
  • Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
  • Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
  • Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
  • Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
  • Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.


Experience & Qualifications

  • 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
  • Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
  • Proven experience implementing automation to reduce operational overhead and improve service reliability.
  • Experience supporting enterprise private cloud and VDI environments.
  • Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
  • Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
  • Experience in Banking or Financial Services environments is advantageous.

    Preferred Skills

    • Windows Server 2016/2019/2022 administration.
    • Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
    • Exposure to hybrid cloud and private cloud platforms.
    • Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
    • Experience supporting enterprise VDI environments.
    • Understanding of ITIL Incident, Problem, Change, and Release Management.
    • Experience working in regulated industries such as Banking or Financial Services.