1

Reliability Manager Jobs in Delray Beach, FL (NOW HIRING)

Staff Site Reliability Engineer

Boca Raton, FL · On-site

$54 - $72/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work on the platform. As a Staff Site ...

Staff Site Reliability Engineer

Boca Raton, FL · On-site

$54 - $72/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work on the platform. As a Staff Site ...

Showing results 21-40

Reliability Manager information

See Delray Beach, FL salary details

$58.2K

$110.3K

$158.2K

How much do reliability manager jobs pay per year?

As of Sep 7, 2026, the average yearly pay for reliability manager in Delray Beach, FL is $110,321.00, according to ZipRecruiter salary data. Most workers in this role earn between $88,700.00 and $131,500.00 per year, depending on experience, location, and employer.

What does a reliability manager do?

A Reliability Manager is responsible for ensuring that equipment, processes, and systems operate efficiently and consistently to minimize downtime and maximize performance. They develop and implement reliability strategies, conduct root cause analyses, and oversee preventive and predictive maintenance programs. Their role involves working closely with maintenance teams, engineers, and production staff to improve asset reliability and extend equipment lifespan. Additionally, they analyze failure data, recommend improvements, and help optimize operational costs through reliability-centered maintenance practices.

What are the key skills and qualifications needed to thrive as a reliability manager?

A Reliability Manager needs strong analytical skills, a solid background in engineering or maintenance, and experience with reliability-centered maintenance methodologies. Familiarity with tools like Failure Mode and Effects Analysis (FMEA), Root Cause Analysis (RCA), and certifications such as Certified Reliability Engineer (CRE) are often required. Leadership, problem-solving, and the ability to communicate complex technical information clearly are crucial soft skills for this role. These skills help ensure equipment uptime, optimize maintenance processes, and foster a culture of continuous improvement within the organization.

What job categories do people searching Reliability Manager jobs in Delray Beach, FL look for?

The top searched job categories for Reliability Manager jobs in Delray Beach, FL are:

What cities near Delray Beach, FL are hiring for Reliability Manager jobs?

Cities near Delray Beach, FL with the most Reliability Manager job openings:

Infographic showing various Reliability Manager job openings in Delray Beach, FL as of August 2026, with employment types broken down into 85% Full Time, 14% Part Time, and 1% Contract. Highlights an 79% Physical, 2% Hybrid, and 19% Remote job distribution, with an average salary of $110,321 per year, or $53 per hour.

Senior Observability Engineer (SRE)

NextEra Energy

Plantation, FL • On-site

$54.25 - $72.25/hr

Full-time

Posted 17 days ago


Key responsibilities

  • Assist in designing, implementing, and operating enterprise observability capabilities across metrics, logs, traces, events, synthetic monitoring, dashboards, and service-health views.

  • Partner with infrastructure, application, cloud, database, network, storage, and operations teams to define monitoring requirements, alert thresholds, escalation paths, and service-health indicators.

  • Analyze recurring incidents, monitoring gaps, alert patterns, and operational trends to identify reliability improvement opportunities.


NextEra Energy rating

8.4

Company rating: 8.4 out of 10

Based on 55 frontline employees who took The Breakroom Quiz

19th of 53 rated energy and utility


Job description

Requisition ID:  96990 

Florida Power & Light Company is the largest electric utility in the U.S., providing reliable energy to nearly 12 million Floridians. With one of the nation's most fuel-efficient, cost-effective power generation fleets and industry-leading reliability, we're redefining what's possible in energy. Want to be part of something powerful? Join our outstanding team and help shape the future of energy.

Position Specific Description

The Senior Observability Engineer / Architect will help advance modern SRE practices across enterprise IT operations, with a focus on service-level visibility, observability, actionable alerting, automation, runbook maturity, and proactive service-health management. This role will partner across Observability, Event Management, infrastructure, application, and ServiceNow teams to improve reliability engineering standards, strengthen monitoring coverage, reduce operational noise, and support the transition from reactive incident response to data-driven, service-health operations. 

The position will provide a technical connection point across Information Technology, infrastructure, application, cloud, ServiceNow, and operations teams to improve reliability, observability, and operational readiness. This role will support the development and execution of observability standards, service health practices, alerting improvements, automation opportunities, and reliability engineering patterns that help teams detect, understand, and resolve service issues more effectively.

Project Execution / Analytical Thinking / Problem Solving 

Assist in designing, implementing, and operating enterprise observability capabilities across metrics, logs, traces, events, synthetic monitoring, dashboards, and service-health views. 
Apply modern SRE principles, including SLIs, SLOs, error-budget thinking, toil reduction, automation, incident learning, and reliability-focused engineering practices. 
Partner with infrastructure, application, cloud, database, network, storage, and operations teams to define monitoring requirements, alert thresholds, escalation paths, and service-health indicators. 
Support observability platform capabilities across tools such as ScienceLogic, ServiceNow ITOM/Event Management, Splunk, cloud-native monitoring platforms, AppDynamics, synthetic monitoring, and related technologies. 
Improve alert quality by helping ensure alerts are actionable, properly routed, associated with the correct configuration item or service, and supported by clear response guidance. 
Assist in aligning operational events to ServiceNow Event Management, including event ingestion, alert correlation, suppression logic, incident creation criteria, and notification workflows. 
Contribute to service-level visibility by supporting dashboards, scorecards, service maps, dependency views, ownership models, and operational health reporting. 
Analyze recurring incidents, monitoring gaps, alert patterns, and operational trends to identify reliability improvement opportunities. 
Develop and maintain runbooks, knowledge articles, technical documentation, monitoring standards, and operational handoff materials. 
Support automation opportunities that reduce manual effort, improve triage consistency, and accelerate restoration while maintaining appropriate governance and controls. 
Collaborate with teams during incidents and problem reviews to improve detection, escalation, root-cause analysis, and long-term prevention. 
Help advance observability maturity through practical adoption of standards such as OpenTelemetry where appropriate, along with consistent telemetry collection and platform integration practices. 
Respond to complex operational scenarios where standard procedures have not resolved the issue and provide technical analysis to support restoration and prevention. 
Continuously evaluate observability practices, platform effectiveness, data quality, and service readiness to improve reliability outcomes across the enterprise. 


Skills / Preferred Qualifications 

Strong understanding of SRE principles and reliability engineering practices. 
Experience with metrics, logs, traces, events, and dashboards. 
Knowledge of observability tools such as ScienceLogic, ServiceNow ITOM, Splunk, or AppDynamics. 
Ability to define SLIs, SLOs, and service-health indicators. 
Experience improving alert quality, routing, and correlation. 
Strong troubleshooting and problem-solving skills. 
Scripting or automation experience to reduce manual effort. 
Ability to collaborate across infrastructure, application, cloud, and operations teams. 


What NextEra Energy employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom