1

Reliability Director Jobs (NOW HIRING)

Senior Director - Observability | SRE

Coppell, TX · On-site

$52.50 - $70/hr

About the Role The Senior Director - Observability and SRE, is a strategic leader accountable for ensuring the reliability, availability, and performance of the enterprise technology ecosystem. This ...

Director, Site Reliability Engineering

$58.25 - $77.50/hr

We're looking for a Senior Manager of Site Reliability Engineering to join our team. You'll lead a team of ~10 SREs across North America, UK, HK, and New Zealand - owning both the day-to-day ...

This is a direct-hire position. Key Responsibilities: • Lead and develop a multi-disciplinary maintenance and reliability organization, including engineers, planners, schedulers, supervisors ...

Showing results 41-60

Reliability Director information

See salary details

$62K

$117.5K

$168.5K

How much do reliability director jobs pay per year?

As of Aug 26, 2026, the average yearly pay for reliability director in the United States is $117,488.00, according to ZipRecruiter salary data. Most workers in this role earn between $94,500.00 and $140,000.00 per year, depending on experience, location, and employer.

What does a reliability director do?

A Reliability Director is responsible for overseeing and improving the reliability and performance of systems, equipment, or processes within an organization. They lead teams to develop and implement maintenance strategies, analyze failure data, and ensure that assets run efficiently with minimal downtime. This role often involves collaborating with other departments to promote best practices and drive a culture of continuous improvement. The Reliability Director also monitors key performance indicators and leads initiatives to reduce operational risks and costs.

What are the key skills and qualifications needed to thrive as a reliability director?

To thrive as a Reliability Director, you need deep expertise in reliability engineering principles, asset management, and a relevant engineering degree, often complemented by certifications like Certified Reliability Engineer (CRE). Familiarity with reliability software (such as ReliaSoft or RAM analysis tools), root cause analysis methodologies, and CMMS systems is typically required. Strong leadership, strategic thinking, and excellent communication skills help drive cross-functional teams and foster a culture of continuous improvement. These skills ensure optimal equipment uptime, cost efficiency, and long-term operational success.

How does a reliability director typically collaborate with cross-functional teams to improve organizational performance?

A Reliability Director frequently partners with engineering, operations, maintenance, and safety teams to identify and address potential points of failure across systems and processes. This collaborative approach ensures that reliability strategies are aligned with production goals and safety standards. Regular meetings, data sharing, and joint problem-solving sessions are key aspects of this teamwork, allowing the director to lead initiatives for root cause analysis, preventive maintenance, and process optimization. These relationships are vital to implementing effective reliability programs and driving continuous improvement throughout the organization.

What is the difference between Reliability Director vs Reliability Engineer?

AspectReliability DirectorReliability Engineer
CredentialsBachelor's or Master's in Engineering, certifications like CRC, CMRPBachelor's in Engineering or related field, certifications like CRC, CMRP often preferred
Work EnvironmentLeadership role overseeing reliability strategies across departmentsTechnical role focused on analyzing data and improving equipment reliability
Employer & IndustryManufacturing, energy, oil & gas, utilitiesManufacturing, energy, oil & gas, utilities
Search & Comparison IntentUnderstanding leadership responsibilities and qualificationsTechnical reliability analysis and improvement tasks

The Reliability Director focuses on strategic leadership, overseeing reliability programs and managing teams, while the Reliability Engineer handles technical analysis, equipment maintenance, and reliability improvements. Both roles are vital in ensuring operational efficiency but differ in scope and responsibilities.

More about Reliability Director jobs

What cities are hiring for Reliability Director jobs?

Cities with the most Reliability Director job openings:

What are the most commonly searched types of Reliability jobs?

The most popular types of Reliability jobs are:

What states have the most Reliability Director jobs?

States with the most job openings for Reliability Director jobs include:

Infographic showing various Reliability Director job openings in the United States as of August 2026, with employment types broken down into 60% Full Time, and 40% Part Time. Highlights an 80% In-person, and 20% Remote job distribution, with an average salary of $117,488 per year, or $56.5 per hour.

Director Site Reliability Engineering

Webster Bank

Southington, CT

$58 - $77/hr

Full-time

Re-posted 4 days ago


Webster Bank rating

7.1

Company rating: 7.1 out of 10

Based on 22 frontline employees who took The Breakroom Quiz

123rd of 172 rated banks


Job description

If you're looking for a meaningful career, you'll find it here at Webster Bank, a division of Santander Bank, N.A.. Founded in 1935, our focus has always been to put people first--doing whatever we can to help individuals, families, businesses and our colleagues achieve their financial goals. As a leading commercial bank, we remain passionate about serving our clients and supporting our communities. Integrity, Collaboration, Accountability, Agility, Respect, Excellence are Webster's values, these set us apart as a bank and as an employer.


Come join our team where you can expand your career potential, benefit from our robust development opportunities, and enjoy meaningful work!


The Director of Site Reliability Engineer is a pivotal technical leader within the Software Engineering organization, tasked with transforming how reliability, performance, and availability are achieved across our platforms. This role goes beyond maintaining systems-it reimagines and modernizes operational practices through automation, cloud-native design, and API-driven integration.You will lead initiatives that elevate our AWS cloud architecture and MuleSoft integration ecosystem, ensuring they are secure, scalable, and resilient. By applying advanced software engineering principles and site reliability practices, you will drive a cultural and technical shift toward proactive reliability, continuous improvement, and innovation.This role requires visionary thinking, deep technical expertise in AWS and MuleSoft, and a passion for driving change that results in more reliable, efficient, and future-ready systems.

What you will do

  • Monitoring and Observability: Implement and maintain tools for monitoring, logging, and tracing to gain insights into system performance and health
  • Automation: Write software and scripts to automate repetitive tasks, such as deployment, monitoring, and system management. Advocate for and lead Automation wherever possible. Ensure environments are well-managed, structured appropriately, cost effective, and synchronized as much as possible.
  • Incident Management: Respond to incidents, troubleshoot system-level issues, and perform root cause analysis to prevent recurrence
  • Reliability Engineering: Design and build reliable and scalable systems, define Service Level Objectives (SLOs) and Indicators (SLIs), and implement reliability patterns
  • Collaboration: Work closely with software developers to ensure applications are reliable and to provide feedback on performance in a production environment
  • Documentation: Create and maintain documentation, including runbooks and system diagrams, to ensure knowledge sharing and team efficiency
  • Set a high bar for reliability and availability -- and meet the bar via automation relentless improvement.
  • Improve and sustain services through rigorous development, testing and release procedures.
  • Key player during deliberations on system design, platform management, and capacity planning.
  • Have a strong 'detective' mindset on why things don't work and be among the first to offer and work on solutions.
  • Be a 'link' between technologists and business stakeholders: able to have conversations with Line of Business (LoB) and technical Agile teams to work through challenges.
  • Partner with peers to advance the maturity of the DevOps practice including new/existing technologies, tools, processes, and standards. Clearly communicate expectations on technical direction and provide ongoing guidance.
  • Serve as a sounding board and technical advisor for your team in the analysis, design, and execution of solutions. Help your team anticipate unforeseen dependencies or gaps early in the SDLC.
  • From your domain's viewpoint, provide leadership and technical expertise to your Agile team to validate story points are sized appropriately, sprint plans are achievable, and releases are well-planned.
  • Shared accountability with peers to ensure quality, performance, and security of systems are optimal and meet both customer SLA's and internal/external audit expectations. Contribute and/or support others in the timely remediation of security remediation, audit, or production support issues escalated to the software engineering group. Occasional evening or weekend involvement may be needed for business-critical situations.


Skills and Abilities

  • Deep understanding of systems development life cycle, cloud-based systems, and application architecture.
  • Experience with multiple programming languages (Python, etc.), configuration management tools, and containers strongly preferred.
  • Moderate to advanced proficiency of cloud development, AWS services (ECS, EKS, S3, RDS, VPC), identity management (Okta), authorization frameworks (OAuth2), monitoring (Dynatrace), Agile DevOps (GitLab, Terraform), application security (Owasp, Veracode, AppScan), and API/Integrations (Apigee, MuleSoft, BizTalk).
  • Familiarity with unit testing concepts and test automation frameworks (SpecFlow, SOAPUI), RESTful APIs and micro-services, WCF services, TSQL, SQL queries and stored procedures.
  • Advanced knowledge and strong desire to work with Agile methodology required (SAFe, Scrum) and experience with Agile tools (Jira, Confluence).


Education Qualifications

  • Bachelor's Degree in Arts/Sciences (BA/BS) in related field required


Experience Qualifications

  • 5-7 years of progressive working experience in designing, building, and maintaining business applications across systems and networks of moderate to high complexity in cloud hosted environments required
  • Experience in setting up SLAs/SLOs/SLIs for critical services and establishing monitoring required
  • Hands on experience in cloud migration journey required
  • Experience and knowledge within the financial services or healthcare industries favorable preferred

The estimated salary range for this position is $135,000.00 to $155,000.00. Actual salary may vary up or down depending on job-related factors which may include knowledge, skills, experience, and location. In addition, this position is eligible for incentive compensation.

#LI-Hybrid

#LI-FO1


Santander Holdings USA, Inc. and its subsidiaries ("Santander") are equal opportunity employers committed to sustaining an inclusive environment. All qualified applicants will receive consideration for employment without regard to race, color, religion, age, marital status, national origin, ancestry, citizenship, sex, sexual orientation, gender identity and/or expression, physical or mental disability, protected veteran status, or any other characteristic protected by law.


What Webster Bank employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom