1

Reliability Manager Jobs in Alsip, IL (NOW HIRING)

Staff Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

As a Staff SRE, you will play a pivotal role in shaping the reliability engineering strategy at WEX ... management, capacity planning, and performance optimization. You'll be hands-on in building ...

Staff Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

As a Staff SRE, you will play a pivotal role in shaping the reliability engineering strategy at WEX ... management, capacity planning, and performance optimization. You'll be hands-on in building ...

Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

Cohesity, Dell PowerProtect Data Manager, Dell Data Domain, Dell Cyber Recovery, Rubrik, Commvault ... Familiarity with SRE tooling and reliability metrics. * Experience implementing AI-assisted ...

New

Site Reliability Engineer

Chicago, IL · On-site

$150 - $200/hr

Cohesity, Dell PowerProtect Data Manager, Dell Data Domain, Dell Cyber Recovery, Rubrik, Commvault ... Familiarity with SRE tooling and reliability metrics. * Experience implementing AI-assisted ...

New

... management strategy; implement AI‑powered operational improvements to improve quality and speed ... E function. * Proven ability to understand and communicate highly complex issues for ...

Manage infrastructure as code using Terraform . * Build and maintain containerized environments ... Reliability & Observability * Own and improve the reliability, availability, performance, and ...

Posted today

Senior Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

Manage infrastructure as code using Terraform . * Build and maintain containerized environments ... Reliability & Observability * Own and improve the reliability, availability, performance, and ...

Manage infrastructure as code using Terraform . * Build and maintain containerized environments ... Reliability & Observability * Own and improve the reliability, availability, performance, and ...

Site Reliability Engineering Lead

Chicago, IL · On-site

$58.75 - $78/hr

Manage Customer Reliability Engineering activities driving Application Monitoring, Metrics, Incident Reviews and Long Term Actions, and support BISO activities / implementing InfoSec changes ...

Site Reliability Engineering Lead

Chicago, IL · On-site

$58.75 - $78/hr

Manage Customer Reliability Engineering activities driving Application Monitoring, Metrics, Incident Reviews and Long Term Actions, and support BISO activities / implementing InfoSec changes ...

Site Reliability Engineer III

Chicago, IL · On-site

$58.75 - $78/hr

Exposure to or hands-on experience in supporting SRE practices for Data management/migration platforms and products, with familiarity in on-prem/public cloud infrastructure components such as Compute ...

Site Reliability Engineer

Chicago, IL · On-site +1

$100K - $120K/yr

Strong knowledge of SRE best practices and incident management protocols * Deep experience using and/or configuring New Relic, Data Dog, SumoLogic or similar observability tools * Proficiency in ...

Site Reliability Engineer III

Chicago, IL · On-site

$58.75 - $78/hr

Exposure to or hands-on experience in supporting SRE practices for Data management/migration platforms and products, with familiarity in on-prem/public cloud infrastructure components such as Compute ...

Site Reliability Engineer III

Chicago, IL · On-site

$58.75 - $78/hr

Exposure to or hands-on experience in supporting SRE practices for Data management/migration platforms and products, with familiarity in on-prem/public cloud infrastructure components such as Compute ...

Reliability Observability Engineer 2

Chicago, IL · On-site

$58.75 - $78/hr

Establish Observability Governance for dashboards, alerts, synthetic monitoring, telemetry standards, and lifecycle management of monitoring assets. * Partner with Product, Engineering, SRE, and ...

New

Site Reliability Engineer III

Chicago, IL · On-site

$58.75 - $78/hr

Exposure to or hands-on experience in supporting SRE practices for Data management/migration platforms and products, with familiarity in on-prem/public cloud infrastructure components such as Compute ...

Showing results 41-60

Reliability Manager information

See Alsip, IL salary details

$63.1K

$119.6K

$171.5K

How much do reliability manager jobs pay per year?

As of Sep 7, 2026, the average yearly pay for reliability manager in Alsip, IL is $119,601.00, according to ZipRecruiter salary data. Most workers in this role earn between $96,200.00 and $142,500.00 per year, depending on experience, location, and employer.

What does a reliability manager do?

A Reliability Manager is responsible for ensuring that equipment, processes, and systems operate efficiently and consistently to minimize downtime and maximize performance. They develop and implement reliability strategies, conduct root cause analyses, and oversee preventive and predictive maintenance programs. Their role involves working closely with maintenance teams, engineers, and production staff to improve asset reliability and extend equipment lifespan. Additionally, they analyze failure data, recommend improvements, and help optimize operational costs through reliability-centered maintenance practices.

What are the key skills and qualifications needed to thrive as a reliability manager?

A Reliability Manager needs strong analytical skills, a solid background in engineering or maintenance, and experience with reliability-centered maintenance methodologies. Familiarity with tools like Failure Mode and Effects Analysis (FMEA), Root Cause Analysis (RCA), and certifications such as Certified Reliability Engineer (CRE) are often required. Leadership, problem-solving, and the ability to communicate complex technical information clearly are crucial soft skills for this role. These skills help ensure equipment uptime, optimize maintenance processes, and foster a culture of continuous improvement within the organization.

What are the most commonly searched types of Reliability jobs in Alsip, IL?

The most popular types of Reliability jobs in Alsip, IL are:

What job categories do people searching Reliability Manager jobs in Alsip, IL look for?

The top searched job categories for Reliability Manager jobs in Alsip, IL are:

What cities near Alsip, IL are hiring for Reliability Manager jobs?

Cities near Alsip, IL with the most Reliability Manager job openings:

Infographic showing various Reliability Manager job openings in Alsip, IL as of August 2026, with employment types broken down into 82% Full Time, 16% Part Time, 1% Contract, and 1% Nights. Highlights an 79% Physical, 3% Hybrid, and 18% Remote job distribution, with an average salary of $119,601 per year, or $57.5 per hour.

Staff Site Reliability Engineer

eNett

Chicago, IL • On-site

$58.75 - $78/hr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Posted 12 days ago


Job description

About the Role

We are looking for a highly motivated, high-potential Staff Site Reliability Engineer (SRE) to join our team as a technical leader and drive transformative impact across WEX's platform reliability and operational excellence.

This is a particularly exciting time to be part of the SRE function at WEX. Our diverse product ecosystem supports a wide array of customer businesses and generates rich, complex telemetry across applications, infrastructure, and platforms. Ensuring these systems are scalable, observable, and resilient is critical to unlocking business value and customer success.

As a Staff SRE, you will play a pivotal role in shaping the reliability engineering strategy at WEX. You'll architect and lead efforts that improve availability, performance, and efficiency at scale, driving initiatives across observability, automation, incident management, problem management, capacity planning, and performance optimization. You'll be hands-on in building foundational tooling and frameworks while also acting as a multiplier, mentoring engineers, aligning cross-functional teams, and influencing platform decisions with a strong reliability lens.

You'll also help define how WEX applies AI to reliability engineering, building agents and reusable skills that automate high-TOIL work, integrating safely into our AI ecosystem, and establishing security and operational guardrails so intelligent automation is trustworthy, measurable, and scalable. Our team embraces agile development, a strong product mindset, and modern engineering practices, including AI-assisted operations and intelligent automation.

You'll take on some of the most complex, high-impact challenges at WEX, supported by a team of highly skilled engineers and technical leaders invested in your success and growth.

If you're a senior technical leader passionate about building reliable systems, leading through influence, and making a meaningful impact with AI-enabled operations, this is a fantastic opportunity for you.

What You'll Do
  • Architect and oversee the implementation of mission-critical systems with a focus on availability, scalability, and operational excellence.

  • Define and enforce SRE best practices and operational standards across engineering and platform teams.

  • Lead cross-functional initiatives to enhance system reliability, performance, and efficiency at scale.

  • Serve as a technical advisor for engineering leadership on reliability, architecture, and operational risk.

  • Develop capacity planning and load testing strategies that proactively identify and mitigate scalability risks.

  • Design self-healing and auto-recovery mechanisms that reduce manual intervention during failures.

  • Drive cloud cost optimization and budgeting initiatives without compromising reliability.

  • Design, build, and govern AI agents and reusable skills that automate operational workflows and reduce TOIL.

  • Evaluate and integrate AI ecosystems, including models, agent frameworks, orchestration, tooling interfaces, and evaluation practices, into SRE and platform workflows.

  • Apply AI security and governance controls, including least-privilege tool access, secure data and prompt handling, auditability, and safe automation boundaries.

  • Lead AI-enabled initiatives for incident response, runbook automation, anomaly detection, and capacity/performance insights, with clear measurement of TOIL reduction and reliability outcomes.

  • Mentor engineers on production-grade agentic solutions and help embed AI into day-to-day reliability practices.

What You'll Bring
  • 8+ years of experience with a focus on large-scale system reliability.

  • Expertise in system architecture, cloud platforms, and automation frameworks.

  • Deep knowledge of Kubernetes, service meshes, and distributed tracing.

  • Experience with monitoring and logging platforms (Grafana, ELK stack, Splunk, etc.).

  • Knowledge of containerization and orchestration (Docker, Kubernetes).

  • Experience designing high-availability, fault-tolerant architectures.

  • Strong understanding of database reliability engineering (MySQL, PostgreSQL, NoSQL), plus networking, databases, and storage architectures.

  • Excellent incident command and crisis management skills.

  • Hands-on experience building AI agents and skills/tools that integrate with operational systems (APIs, observability, ticketing, CI/CD).

  • Working knowledge of AI ecosystems and agent architectures, including orchestration, tool calling, context/memory, evaluation, and human-in-the-loop patterns.

  • Practical understanding of AI security and governance for production use, secure permissions, data leakage prevention, secrets handling, and guarded autonomous actions.

  • Demonstrated ability to reduce TOIL with AI by automating repetitive operational work and delivering measurable efficiency and reliability gains.

Nice to Have

  • Experience with multi-region and multi-cloud deployments.

  • Deep expertise in scalable microservices and event-driven architectures.

  • Strong experience with advanced observability tools (OpenTelemetry, Jaeger, Prometheus).

  • Leadership in driving large-scale SRE transformations.

  • Experience designing and developing AI agents, skills, and copilots for SRE/platform engineering, including evaluation and safe rollout practices.

  • Familiarity with enterprise agent platforms, skill registries, and observability for AI/agent workflows.

  • Ability to influence engineering culture and process improvements, including adoption of AI-assisted operations under change control, safety, and audit requirements.

The base pay range represents the anticipated low and high end of the pay range for this position. Actual pay rates will vary and will be based on various factors, such as your qualifications, skills, competencies, and proficiency for the role. Base pay is one component of WEX's total compensation package. Most sales positions are eligible for commission under the terms of an applicable plan. Non-sales roles are typically eligible for a quarterly or annual bonus based on their role and applicable plan. WEX's comprehensive and market competitive benefits are designed to support your personal and professional well-being. Benefits include health, dental and vision insurances, retirement savings plan, paid time off, health savings account, flexible spending accounts, life insurance, disability insurance, tuition reimbursement, and more. For more information, check out the "About Us" section.Pay Range: $120,600.00 - $150,900.00