1

Senior Reliability Engineer Jobs in Monee, IL (NOW HIRING)

Senior DevOps/SRE Engineer

Chicago, IL · On-site

$140K - $170K/yr

We are looking for aSeniorSite Reliability Engineer to work as part of a lean,productfocusedengineering organization. This role is about building andoperatingreliablecloudbasedsystems by writing code ...

Senior Distinguished Engineer

Chicago, IL · Hybrid

$107K - $147K/yr

You will lead 100+ engineers across multiple directors, grow high-potential senior leaders, and even found a novel SRE function that raises operational excellence across the entire Customer ...

Overview We are seeking a Senior AI Engineer role to build and operate an internal platform that ... Partner with SRE/operations to support availability, latency, performance, change management ...

Overview We are seeking a Senior AI Engineer role to build and operate an internal platform that ... Partner with SRE/operations to support availability, latency, performance, change management ...

Senior DevOps Engineer

Chicago, IL · On-site

$133K - $172K/yr

As a Senior DevOps Engin eer , you will be a key technical contributor responsible for designing ... Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity ...

Showing results 41-60

Senior Reliability Engineer information

See Monee, IL salary details

$20

$62

$89

How much do senior reliability engineer jobs pay per hour?

As of Aug 17, 2026, the average hourly pay for senior reliability engineer in Monee, IL is $62.30, according to ZipRecruiter salary data. Most workers in this role earn between $51.39 and $74.62 per hour, depending on experience, location, and employer.

How much do senior reliability engineers get paid?

Senior reliability engineers typically earn between $90,000 and $130,000 annually, depending on experience, industry, and location. They often have expertise in systems analysis, failure modes, and reliability tools like FMEA and RCM, which can influence compensation levels.

What are the key skills and qualifications needed to thrive as a senior reliability engineer?

To thrive as a Senior Reliability Engineer, you need expertise in reliability engineering principles, root cause analysis, and a relevant engineering degree such as mechanical, electrical, or industrial engineering. Familiarity with tools like FMEA, RCA software, CMMS, and certifications such as Certified Reliability Engineer (CRE) are often required. Strong analytical thinking, communication skills, and the ability to lead cross-functional teams set top performers apart. These skills are essential for minimizing downtime, improving system reliability, and ensuring safe, efficient operations.

What are some common challenges faced by senior reliability engineers, and how are they typically addressed within the team?

Senior Reliability Engineers often encounter challenges such as diagnosing complex system failures, balancing proactive maintenance with urgent reactive fixes, and ensuring consistent communication across multidisciplinary teams. These challenges are typically addressed through root cause analysis, prioritization frameworks, and fostering a culture of knowledge sharing. Regular collaboration with operations, maintenance, and engineering teams helps in developing effective solutions and continuous improvement strategies.

What does a senior reliability engineer do?

A Senior Reliability Engineer is responsible for ensuring that systems, products, or processes operate reliably and efficiently over time. They analyze failure data, design reliability tests, develop maintenance strategies, and work with cross-functional teams to improve system performance and reduce downtime. Their expertise helps organizations minimize risk, optimize lifecycle costs, and maintain high standards of quality and safety. Senior Reliability Engineers often mentor junior team members and play a key role in developing reliability standards and best practices.

What is the difference between Senior Reliability Engineer vs Reliability Engineer?

AspectSenior Reliability EngineerReliability Engineer
CredentialsTypically requires 5+ years experience, certifications like CRE or Six SigmaEntry to mid-level, often with 2-4 years experience, similar certifications
Work EnvironmentDesigns and oversees reliability programs, leads projectsPerforms analysis, supports reliability improvements
Industry UsageUsed across manufacturing, energy, aerospaceCommon in same industries, often as a stepping stone to senior roles

The main difference between a Senior Reliability Engineer and a Reliability Engineer lies in experience, leadership responsibilities, and scope of work. Senior Reliability Engineers typically lead projects and develop strategies, while Reliability Engineers focus on analysis and supporting reliability initiatives. Both roles are vital in ensuring equipment and system dependability across industries.

What job categories do people searching Senior Reliability Engineer jobs in Monee, IL look for?

The top searched job categories for Senior Reliability Engineer jobs in Monee, IL are:

What cities near Monee, IL are hiring for Senior Reliability Engineer jobs?

Cities near Monee, IL with the most Senior Reliability Engineer job openings:

Infographic showing various Senior Reliability Engineer job openings in Monee, IL as of August 2026, with employment types broken down into 80% Full Time, and 20% Contract. Highlights an 90% In-person, 5% Hybrid, and 5% Remote job distribution, with an average salary of $129,585 per year, or $62.3 per hour.

Senior Site Reliability Engineer, Observability

Ripple

Chicago, IL

$58.75 - $78/hr

Full-time

Re-posted 9 days ago


Job description

At Ripple, we're building a world where value moves like information does today. Through our crypto solutions for financial institutions, businesses, governments, and developers, we are improving the global financial system and creating greater economic fairness and opportunity for more people, in more places around the world.

Ripple Treasury, now a Ripple solution acquired in 2025, marks a significant expansion into the multi-trillion-dollar corporate finance arena. With more than 40 years of experience supporting some of the world's largest and most sophisticated companies, Ripple Treasury integrates a treasury command center into Ripple's technology stack-giving corporates the ability to move, manage, and optimize liquidity in real-time, across traditional and digital assets, under one expanded umbrella.

THE WORK:

This is an engineering-first role with a coaching dimension-not the other way around. You will spend the majority of your time doing hands-on observability and reliability engineering work: building instrumentation, designing alert configurations, authoring Terraform, and troubleshooting production systems. Alongside that, you will coach and consult with stream-aligned product teams, helping them build operational maturity over time.

You will join Ripple's Technical Operations team and work across Azure (80%) and AWS (20%) environments supporting infrastructure that is predominantly Windows-based (80%), handling significant payment volume for enterprise treasury customers. The incident management program you will help build is early-stage-you will be establishing practices, not inheriting a mature playbook.


WHAT YOU'LL DO:

  • Observability Engineering
    • Design and implement monitoring, alerting, and dashboards in New Relic (APM, Infrastructure, Logs, Synthetics) across Azure and AWS; write NRQL queries for troubleshooting, analysis, and reporting.
    • Define and implement SLOs/SLIs and error budgets; coach teams on using them to balance feature velocity with reliability and communicate system health to stakeholders.
    • Lead alert noise reduction and signal quality engineering-tune thresholds, eliminate false positives, and ensure every alert is actionable.
    • Optimize observability costs through log ingestion management, pipeline rules, and New Relic configuration governance.
    • Partner with engineering teams to improve observability maturity: structured logging, metrics instrumentation (RED/USE methods), distributed tracing, and effective dashboard patterns.
    Infrastructure & IaC
    • Develop and maintain Terraform infrastructure as code for provisioning and managing monitoring resources, alert configurations, and observability infrastructure-this is a primary engineering responsibility, not an occasional task.
    • Establish and enforce IaC governance standards for observability infrastructure across teams, providing a repeatable, auditable model for how monitoring resources are managed.
    • Author and troubleshoot Azure DevOps pipelines; support teams with deployment visibility, change tracking, and release hygiene as it relates to production reliability.
    Incident Management
    • Administer and configure Incident.IO: alert routing, notification workflows, Slack and OpsGenie integration, and runbook management-operationalizing what exists today and expanding from there.
    • Build out incident management foundations that are largely yours to establish: PIR/postmortem processes, on-call rotation design, escalation policies, incident severity classification, and response playbooks.
    • Track and report on MTTR, MTTD, and incident frequency; identify trends and drive continuous improvement in partnership with engineering teams.
    • Respond to and debrief on production incidents-providing real-time troubleshooting support and facilitating structured post-incident reviews.
    Cross-Functional Enablement
    • Enable stream-aligned engineering teams to adopt improved observability and incident management practices through workshops, consultation, and hands-on guidance.
    • Collaborate with the Subsystems Platform Team to translate common needs into self-service observability and incident management capabilities.
    • Build lasting team competency through documentation, training materials, and knowledge-sharing sessions that outlast any individual engagement.

WHAT YOU'LL BRING: 

  • Core SRE Experience
    • 7+ years in Site Reliability Engineering, DevOps, or Platform Engineering with a strong focus on observability and production operations.
    • Proven ability to deliver hands-on engineering work while coaching and mentoring teams-comfortable switching between builder and consultant modes.
    • Experience working in Agile/Scrum environments and collaborating effectively with cross-functional teams.

    Observability & Incident Management Expertise - Required
    • Expert-level hands-on experience with New Relic (APM, Infrastructure, Logs, Synthetics, Alerts) and strong NRQL proficiency for troubleshooting and analysis.
    • Deep understanding of structured logging, metrics collection (RED/USE methods), distributed tracing, and designing effective dashboards and alerts.
    • Expertise defining and implementing SLOs/SLIs and error budgets for reliability management.
    • Hands-on experience with incident management platforms (Incident.IO, PagerDuty, OpsGenie, or similar).
    • Experience designing incident response workflows, on-call rotations, escalation policies, and facilitating post-incident reviews that drive actionable improvements.
    • Demonstrated ability to troubleshoot complex production issues using observability data across distributed systems.

    Infrastructure & Tools - Required
    • Strong Terraform experience: developing and maintaining IaC for cloud infrastructure and monitoring resources; familiarity with IaC governance patterns.
    • Proficiency with PowerShell scripting (required given the 80% Windows environment).
    • Strong experience with Azure cloud (App Services, Virtual Machines, Azure SQL, networking, monitoring) and working knowledge of AWS.
    • Experience with Azure DevOps for CI/CD pipeline authoring and troubleshooting.
    • Experience with Octopus Deploy for deployment management and release orchestration.
    • Comfort working across both Windows and Linux server environments.
    • Familiarity with Slack for operational workflows, alert routing, and incident communication.

    Desired / Additional
    • Experience with alert noise reduction strategies and observability cost optimization (log ingestion, pipeline rules, cardinality management).
    • Background facilitating chaos engineering, game day exercises, or failure injection to build team resilience.
    • Knowledge of VM-hosted SQL Server monitoring and performance optimization.
    • Familiarity with FinTech compliance requirements (SOC 2, ISO 27001) and audit evidence collection.
    • Experience measuring and improving key reliability metrics (MTTR, MTTD, availability, error budgets) at an organizational level.
    • Python or Bash scripting experience in addition to PowerShell.
    • Familiarity with Jira for incident tracking and workflow automation.
    Other common names for this role: Senior Site Reliability Engineer, Observability Engineer, Incident Management Engineer