1

Director Reliability Manager Jobs in Illinois (NOW HIRING)

Engineer, Reliability II

North Chicago, IL

$98K - $124K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

... by managing the reliability improvement processes for the site, with the aim of reducing the risk levels to the business. Within the role, it will direct and provide leadership for key equipment ...

Engineer, Reliability II

North Chicago, IL ยท On-site

$98K - $124K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

... by managing the reliability improvement processes for the site, with the aim of reducing the risk levels to the business. Within the role, it will direct and provide leadership for key equipment ...

Engineer, Reliability II

North Chicago, IL ยท On-site

$65K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

... by managing the reliability improvement processes for the site, with the aim of reducing the risk levels to the business. Within the role, it will direct and provide leadership for key equipment ...

Site Reliability Engineer, Observability

Chicago, IL ยท On-site

$160 - $200/hr

  • Retirement

  • PTO

... manage and optimize liquidity in real-time, across traditional and digital assets, under one ... Consultative mindset with the ability to influence and guide teams without direct authority ...

Showing results 21-40

Director Reliability Manager information

What is the difference between Director Reliability Manager vs Reliability Engineer?

AspectDirector Reliability ManagerReliability Engineer
CredentialsBachelor's or Master's in Engineering, certifications like CRC, CMRPBachelor's in Engineering or related field, certifications like CRC, CMRP
Work EnvironmentLeadership roles overseeing teams, strategic planningTechnical roles focused on analysis, testing, and troubleshooting
Industry UsageUsed in manufacturing, energy, aerospace for high-level reliability strategiesCommon in manufacturing, maintenance, and engineering teams
Search & Comparison IntentUnderstanding leadership responsibilities, strategic focusTechnical skills, daily tasks, and hands-on work

The Director Reliability Manager typically oversees reliability strategies and manages teams, focusing on high-level planning and decision-making. Reliability Engineers are more involved in technical analysis, testing, and implementing reliability improvements. Both roles require similar credentials but differ in scope and responsibilities within organizations.

What are the key skills and qualifications needed to thrive as a director reliability manager?

To thrive as a Director Reliability Manager, you need deep expertise in reliability engineering, maintenance management, and a relevant engineering degree, often supported by several years of leadership experience. Familiarity with reliability-centered maintenance (RCM), predictive analytics tools, CMMS software, and certifications like Certified Reliability Engineer (CRE) are typically required. Strong leadership, strategic thinking, and effective communication skills set outstanding candidates apart in this role. These competencies ensure optimized asset performance, minimized downtime, and alignment between technical teams and organizational goals.

What are some common challenges faced by a director reliability manager, and how are they typically addressed?

A Director Reliability Manager often faces challenges such as balancing long-term reliability improvements with immediate operational demands, integrating reliability practices across diverse teams, and ensuring consistent data-driven decision-making. Addressing these challenges typically involves fostering a culture of proactive maintenance, promoting cross-department collaboration, and implementing robust reliability tracking systems. Building strong relationships with engineering, operations, and maintenance teams is crucial for aligning reliability goals with business objectives and achieving sustainable results.

What does a director reliability manager do?

A Director Reliability Manager is responsible for overseeing and improving the reliability and performance of an organization's systems, equipment, or processes. This role typically involves leading teams to identify and mitigate risks, develop maintenance strategies, and implement best practices to reduce downtime and increase efficiency. The Director works closely with other departments to analyze data, drive reliability initiatives, and ensure compliance with industry standards. Their ultimate goal is to enhance operational reliability, reduce costs, and support business objectives.

What are the most commonly searched types of Reliability Manager jobs in Illinois?

The most popular types of Reliability Manager jobs in Illinois are:

What are popular job titles related to Director Reliability Manager jobs in Illinois?

For Director Reliability Manager jobs in Illinois, the most frequently searched job titles are:

What job categories do people searching Director Reliability Manager jobs in Illinois look for?

The top searched job categories for Director Reliability Manager jobs in Illinois are:

What cities in Illinois are hiring for Director Reliability Manager jobs?

Cities in Illinois with the most Director Reliability Manager job openings:

Director Enterprise Platform Governance SRE

Request Technology, LLC

Chicago, IL โ€ข On-site

$58.75 - $78/hr

Other

Re-posted 28 days ago


Job description

***Position is bonus eligible***

Prestigious Financial Institution is currently seeking a Director Enterprise Platform Governance and Strategy with strong SRE leadership experience. Candidate will manage a team of engineering managers and senior technical staff across Site Reliability Engineering, Cloud Architecture, and Metrics & Reporting functions. The leader in this role is accountable for scaling a mature SRE practice, driving cloud architectural standards and multi-year strategy, and ensuring the organization operates with clear, data-driven visibility into platform health and performance. A critical dimension of this role is ownership of the FinOps and SecOps domains as Product Manager, alongside governance of PE compliance obligations spanning incidents, risks, and audit findings. 

Responsibilities:

Site Reliability Engineering

    • Lead the scaling and maturation of the SRE practice, establishing error budgets, SLOs, SLAs, and incident response frameworks across all platform services.
    • Define and enforce reliability standards including on-call models, blameless postmortem processes, and corrective action tracking to drive continuous improvement.
    • Partner with Platform Foundation teams (Kubernetes, Kafka, FinOps/Security) to embed reliability principles into build and operate models.
    • Champion toil reduction through automation, ensuring engineering capacity is redirected from manual operations to higher-value platform capabilities.

Platform Engineering Governance & Compliance

    • Serve as Product Manager for the FinOps and SecOps domains within Platform Engineering, owning the product vision, prioritization, and stakeholder alignment for governance tooling and practices.
    • Establish and maintain a governance framework ensuring Platform Engineering adheres to organizational standards across incident and problem management, SORTs, risk tracking, and audit findings.
    • Own the end-to-end process for PE compliance obligations, ensuring timely resolution and closure of incidents, problem tickets, risk items, and audit observations with clear accountability and tracking.
    • Partner with Risk, Compliance, and Security functions to proactively identify governance gaps, drive remediation, and ensure PE operates within the organization's risk appetite.
    • Maintain visibility and reporting on PE's compliance posture across all obligation types, surfacing trends, aging items, and residual risks to CARE leadership and relevant stakeholders.

Site Reliability Engineering COE

    • Lead the scaling and maturation of the SRE practice, establishing error budgets, SLOs, SLAs, and incident response frameworks across all platform services.
    • Define and enforce reliability standards including on-call models, blameless postmortem processes, and corrective action tracking to drive continuous improvement.
    • Partner with Platform Engineering Product teams (Kubernetes, Kafka, FinOps/Security) to embed reliability principles into build and operate models.
    • Champion toil reduction through automation, ensuring engineering capacity is redirected from manual operations to higher-value platform capabilities.

Cloud Strategy & Architecture

    • Define and execute the multi-year cloud architecture strategy aligned to business growth, scalability, regulatory compliance, and cost optimization goals.
    • Establish cloud architectural standards, reference architectures, and governance frameworks (landing zones, identity, network patterns, service catalog) and drive adoption across engineering.
    • Guide cloud-native architecture decisions including containers/orchestration, IaaS/PaaS adoption, disaster recovery, and multi-region patterns with a steady eye on regulatory requirements (e.g., CIS, NIST).
    • Oversee technology roadmaps and end-of-life planning for cloud platform components, ensuring forward-looking decisions balance innovation with operational stability.
    • Serve as a key technical advisor to senior leadership, translating complex architectural trade-offs into clear business decisions.

Metrics & Reporting

    • Own the platform metrics and reporting function, establishing a consistent framework for measuring platform health, engineering velocity, reliability, and cost efficiency across CARE.
    • Define and track KPIs aligned to internal SLAs, executive reporting needs, and audit/compliance requirements.
    • Ensure Jira and other platform tooling serve as the single source of truth for work visibility, with dashboards and reporting that enable data-driven prioritization.
    • Build and maintain reporting cadences for leadership, including platform health scorecards, capacity forecasting, and risk transparency.

PM Coordination & Platform Delivery

    • Serve as the primary engineering leadership partner to the Platform Engineering Program Management function, ensuring platform initiatives are properly scoped, sequenced, and resourced.
    • Drive alignment between engineering capacity and roadmap commitments, proactively surfacing dependency risks and trade-off decisions to the CARE Executive Director.
    • Coordinate across PE domains to ensure cross-team delivery dependencies are managed and resolved effectively.
    • Partner with Product and Engineering leaders outside of CARE to align platform capabilities to broader organizational roadmaps.

Leadership & Organizational Excellence

    • Lead, develop, and retain a high-performing team of engineering managers and individual contributors with clear ownership, career paths, and accountability frameworks.
    • Foster a culture consistent with CARE operating principles: automation-first, full-stack ownership, stability as a prerequisite for velocity, and transparency through tooling.
    • Manage budget for areas of responsibility; ensure adherence to schedules, work plans, and performance requirements.
    • Oversee remediation of audit findings and observations within areas of responsibility, ensuring root cause is addressed, residual risk is reduced, and remediation is completed timely.
    • Maintain appropriate work/life balance within teams while upholding a high standard of delivery quality

Qualifications:

    • [Required] Proven executive-level leadership of SRE, cloud engineering, or platform reliability organizations in a regulated industry environment.
    • [Required] Demonstrated ability to build and scale SRE practices including SLO/SLA frameworks, on-call models, error budgets, and incident response programs.
    • [Required] Deep expertise in cloud architecture strategy and governance, with experience defining and driving enterprise-wide architectural standards.
    • [Required] Strong track record of cross-functional partnership with Program/Product Management, translating platform capabilities into sequenced, delivery-ready roadmaps.
    • [Required] Demonstrated experience serving in a Product Manager capacity for technical domains such as FinOps, SecOps, or platform tooling, including ownership of roadmap, prioritization, and stakeholder alignment.
    • [Required] Experience establishing and managing governance and compliance frameworks within a platform or infrastructure engineering organization, including oversight of incidents, problem management, risk items, and audit obligations.
    • [Required] Ability to design and maintain metrics and reporting frameworks that provide meaningful visibility into platform health, engineering performance, and compliance posture.
    • [Required] Exceptional written and verbal communication skills; ability to translate technical complexity into executive-level insights and business decisions.
    • [Required] Demonstrated ability to lead high-performing, highly technical teams through accountability, coaching, and clear ownership models.
    • [Required] Experience managing work in Agile/Scrum environments with strong prioritization and deadline management discipline.
    • [Preferred] Experience operating in a production change control process and working directly with audit and compliance functions in a regulated environment.
    • [Preferred] Experience in financial services or similarly regulated industries with exposure to CIS, NIST, and related frameworks.

Technical Skills:

    • [Required] Deep knowledge of SRE tooling and observability platforms (e.g., Prometheus, Grafana, PagerDuty, Datadog, or equivalents). 
    • [Required] Expert-level knowledge of cloud platforms: AWS, Azure, or Google Cloud Platform; experience with multi-cloud or hybrid environments preferred. 
    • [Required] Strong working knowledge of cloud-native architecture patterns and Infrastructure as Code principles. 
    • [Required] Familiarity with container orchestration and streaming platforms (Kubernetes, Kafka) and CI/CD tooling (GitHub Actions, Jenkins, or equivalents). 
    • [Required] Experience with metrics and reporting platforms; ability to design KPI frameworks and reporting dashboards for both technical and executive audiences. 
    • [Required] Working knowledge of FinOps principles and cloud cost governance, with experience driving cost transparency and optimization at an organizational level. 
    • [Required] Familiarity with SecOps tooling and security governance practices within a cloud or platform engineering context. 
    • [Preferred] Experience with GRC tooling or platforms used to manage risk, audit findings, and compliance obligations (e.g., ServiceNow, Archer, or equivalents).

Education and/or Experience:

    • [Required] Bachelor's degree, preferably in a technical discipline (Computer Science, Mathematics, Engineering, or related field), or equivalent combination of education and experience. 
    • [Required] 15+ years of progressive experience in cloud engineering, platform reliability, or infrastructure roles with at least 5 years in senior engineering leadership. 
    • [Preferred] Depth and breadth of experience in a highly regulated industry such as financial services, with demonstrated understanding of applicable rules and regulatory frameworks. 
    • [Required] AWS Solutions Architect Associate Certification or higher strongly desired. 
    • [Preferred] Google Cloud Professional Cloud Architect, Microsoft Azure Solutions Architect, or equivalent certification. 
    • [Preferred] Relevant SRE or reliability-focused certifications a plus.