2

Remote Reliability Engineer Jobs in Addison, IL (NOW HIRING)

Principal Software Engineer

Chicago, IL ยท On-site +1

$117K - $123K/yr

  • Medical

  • Dental

  • Life

  • Retirement

... reliability, and maintainability (10%). Provide technical leadership across multiple engineering ... Remote work requests will be considered consistent with company's remote work policy. Job ...

DevOps Engineer

Chicago, IL ยท Remote

$54 - $74/hr

Participate in evaluating and integrating new technologies to enhance the scalability, reliability ... We offer a hybrid work schedule to perfectly combine the benefits of remote work and the essential ...

Senior DevOps Engineer

Chicago, IL ยท On-site +1

$160K - $200K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

Reliability & Security * Improve observability, monitoring, and production reliability. * Lead ... , SRE, or platform engineering. * Strong software engineering background with production ...

SVP, Chief Technology Officer (CTO)

Downers Grove, IL ยท On-site +1

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

... IL, Remote-MI, Remote-NJ, St. Louis, Missouri Details Kemper is one of the nation's leading ... Site Reliability Engineering (SRE) * Enterprise observability and monitoring * Identity and shared ...

SVP, Chief Technology Officer (CTO)

Chicago, IL ยท On-site +1

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

... IL, Remote-MI, Remote-NJ, St. Louis, Missouri Details Kemper is one of the nation's leading ... Site Reliability Engineering (SRE) * Enterprise observability and monitoring * Identity and shared ...

Software Engineer

Naperville, IL ยท On-site +1

$132K - $147K/yr

Reliability Engineering: Define and establish company-wide Test-Driven Development (TDD) and ... remote work is not an option for tax reasons: AL, AK, AR, CA, CT, DE, HI, ID, IA, KS, KY, LA, ME ...

Senior DevOps Engineer

Chicago, IL ยท On-site +1

$120K - $160K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

What We're Looking For * 3+ years of experience in a DevOps or SRE role. * Infrastructure as Code expertise (Terraform/Open Tofu, Helm). * A strong Devops SDLC mindset. * Seasoned in orchestrating ...

Design Engineer (PE)

Wheaton, IL ยท On-site +1

$100K - $125K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Remote/Wheaton, IL Hire Type: Direct Hire Salary: $100K-$125K/YR Benefits: Medical, Dental, Vision ... Support electric distribution planning, capacity analysis, reliability improvements, and ...

Substation Physical Engineer - REMOTE

Chicago, IL ยท Remote

$101K - $129K/yr

  • Retirement

Substation Physical Engineer - REMOTE Location: Remote US Ready to make a difference? ICF is ... and reliability> ICF Power Delivery Services #GEA25 #POWERDELIVERY #INDEED #LI-CC1 Bot and Third ...

Showing results 41-60

Remote Reliability Engineer information

See Addison, IL salary details

$61.1K

$118.2K

$141.3K

How much do remote reliability engineer jobs pay per year?

As of Aug 18, 2026, the average yearly pay for remote reliability engineer in Addison, IL is $118,193.00, according to ZipRecruiter salary data. Most workers in this role earn between $102,700.00 and $129,200.00 per year, depending on experience, location, and employer.

What is a remote reliability engineer?

A Remote Reliability Engineer is a professional who works from a remote location to ensure that systems, applications, or infrastructure are reliable, available, and performing well. Their responsibilities typically include monitoring system health, diagnosing issues, implementing preventative measures, and collaborating with teams to improve system reliability. They often use tools for automation, incident response, and performance monitoring, all while working offsite. This role is critical in minimizing downtime and ensuring a smooth user experience, especially for companies with complex technical environments. Remote Reliability Engineers must have strong problem-solving skills and be proficient in cloud technologies, automation, and incident management.

What are the key skills and qualifications needed to thrive as a remote reliability engineer?

To thrive as a Remote Reliability Engineer, you need a strong background in systems engineering, software development, and infrastructure management, often supported by a degree in computer science or a related field. Proficiency with cloud platforms (such as AWS, Azure, or GCP), monitoring tools (like Prometheus, Grafana), and relevant certifications (e.g., AWS Certified DevOps Engineer) is highly valuable. Excellent problem-solving, communication, and collaboration skills are crucial for working effectively across distributed teams and responding to incidents. These abilities ensure system reliability, quick incident resolution, and seamless remote teamwork, which are vital for maintaining high service uptime and user satisfaction.

How do remote reliability engineers typically collaborate with on-site teams to address urgent technical issues?

Remote Reliability Engineers often utilize a combination of video conferencing, instant messaging, and collaborative monitoring tools to stay closely connected with on-site teams. When urgent technical issues arise, they participate in real-time troubleshooting sessions, analyze system logs remotely, and may guide on-site staff through step-by-step resolution procedures. Building strong communication channels and regular check-ins are essential to ensure swift and effective collaboration, even across different time zones. This structure allows Remote Reliability Engineers to contribute significantly to system uptime while working from a distance.

What is the difference between Remote Reliability Engineer vs Remote Site Reliability Engineer?

AspectRemote Reliability EngineerRemote Site Reliability Engineer
CredentialsTypically requires certifications like AWS Certified Solutions Architect, Linux Foundation certificationsSimilar credentials, often with additional focus on site-specific tools and monitoring
Work EnvironmentPrimarily remote, focusing on cloud infrastructure and system reliabilityRemote with some on-site responsibilities, focusing on infrastructure and operational stability
Industry UsageUsed across tech, cloud providers, SaaS companiesCommon in data centers, cloud providers, and large enterprise IT
Search & Comparison IntentOften compared due to overlapping roles in system reliability and cloud infrastructureCompared for on-site vs remote operational responsibilities

The main difference is that Remote Reliability Engineers focus on cloud and system reliability remotely, while Remote Site Reliability Engineers may have some on-site duties related to infrastructure. Both roles require similar skills and certifications but differ in their work environment and specific responsibilities.

What are popular job titles related to Remote Reliability Engineer jobs in Addison, IL?

For Remote Reliability Engineer jobs in Addison, IL, the most frequently searched job titles are:

What cities near Addison, IL are hiring for Remote Reliability Engineer jobs?

Cities near Addison, IL with the most Remote Reliability Engineer job openings:

Production Engineer (IC4) (Remote)

Ontrac Solutions

Chicago, IL โ€ข Remote

$75 - $85/hr

Full-time

Posted 10 days ago


Job description

Overview
Ontrac Solutions is seeking a high-aptitude Production Engineer (IC2) to support a large-scale enterprise OS modernization and infrastructure hardening program for one of our enterprise clients. This role is built for an engineer with real software engineering foundations specifically Python who has since moved into infrastructure and is comfortable working across OS modernizations (RHEL7 EL8/EL9), packaging migrations (Chef CINC), and CI/CD hardening, while simultaneously executing hands-on runbooks, automation, and service onboarding.

This is a genuine hybrid role: roughly half build Python tooling, RPM packaging, pipeline and rollback work and half operate runbooks, service onboarding, and Tier-2 break/fix. You will work directly with the client's SRE organization, internal engineering teams, and customer stakeholders from initial definition through final delivery. Candidates who are pure application developers with no Linux fleet exposure, and candidates who are pure operations with no real software engineering behind them, will not clear screening.

What your application must clearly show
We screen against the requirements below exactly as written your resume should make these easy to find. Specifics matter more than vocabulary: a resume that restates this posting's terminology without the detail below will not advance.

  • Python you actually wrote, described as software, not as a skills keyword. Name the project, what it did, who depended on it, and how it was tested and shipped. "Python (scripting)" in a skills list will not clear this bar. A GitHub, GitLab, or public repo link is strongly preferred we look at code.
  • RPM packaging you personally did. Name the .spec files you authored or maintained, how you handled dependencies and versioning, your build tooling (rpmbuild, mock, Koji, or an internal builder), roughly how many packages you owned, and where they were published.
  • A real OS migration you worked on with version numbers and what actually broke. System Python 23, OpenSSL and crypto-policy changes, systemd unit differences, deprecated or renamed packages. We are more interested in the failure modes you hit than in the name of the program.
  • Configuration management you have run in production Chef (cookbooks, recipes, Ohai, Test Kitchen/InSpec), CINC, Puppet, Ansible, or Salt. Say which resources you wrote and how you tested convergence.
  • Monitoring and logging work you executed. Name the stack, what you onboarded to it, and what you actually instrumented metrics, dashboards, alert rules, log pipelines not just the product name.
  • Tier-2 or on-call experience: the rotation you carried, the scale of the fleet behind it, and one incident you personally drove to resolution.
  • CI/CD pipelines you built or hardened, including how rollout and rollback were handled when a change went wrong.
  • Your certifications, named, with dates and credential IDs or verification links we verify certifications.

Shortly after you apply you will receive a short role-specific questionnaire completing it promptly is the fastest way to move into screening.

Required Qualifications:

  • Software engineering foundation: 3+ years of professional software engineering experience, with strong hands-on Python. You have written and maintained code other engineers depended on modules, packaging, tests, code review, and version control, not just single-file scripts.
  • Enterprise OS modernization: Hands-on experience with enterprise Linux at fleet scale and with large-scale OS upgrade programs RHEL7 EL8/EL9 or an equivalent major-version migration you executed rather than observed.
  • Packaging migrations: Experience building RPM packages to replace legacy configuration, and with large-scale packaging or configuration-management migrations such as Chef CINC.
  • Independent bug ownership: Able to triage, own, and resolve bugs end to end without hand-holding reproduce, isolate, fix, test, and ship.
  • CI/CD and release safety: Ability to harden CI/CD pipelines, observability frameworks, and rollout/rollback mechanisms specifically tailored for legacy-to-modern infrastructure transitions.
  • Tier-2 operational support: Willingness and experience to partner closely with an SRE team providing "follow-the-sun" tier-2 support, including hands-on incident response and break/fix operations on existing platforms.
  • Service onboarding: Experience onboarding services to newly established monitoring and logging stacks.
  • Automation and documentation: A demonstrated habit of automating repetitive operations and documenting technical procedures for others to run.
  • Location and work authorization: Must be located in the United States and authorized to work in the US.

Preferred Qualifications:

  • Proven experience planning and executing logging and monitoring tool rollouts end to end not only operating a stack someone else stood up.
  • Experience supporting a team through cloud cutovers and component migrations to cloud environments.
  • Perl scripting experience legacy tooling in this environment is Perl, and the ability to read and safely modify it is a real advantage.
  • Hands-on exposure to modern observability tooling Chronosphere, Prometheus, or Grafana and to Splunk integrations.
  • Provisioning and image work: Kickstart/PXE, golden images, or repository and mirror management.

Key Responsibilities OS Modernization & Packaging:

  • Drive the technical transition of legacy systems to modern enterprise Linux environments, including RHEL7 EL8/EL9 upgrade paths.
  • Build and maintain RPM packages to replace legacy configuration, and carry the packaging migration from Chef to CINC.
  • Develop and execute automated runbooks that make the migration repeatable rather than manual.
  • Triage, own, and resolve migration bugs independently, from first report through verified fix.

Key Responsibilities Observability & Monitoring Transition:

  • Transition monitoring infrastructure to a modern stack Chronosphere, Prometheus, and Grafana and manage Splunk integrations.
  • Onboard services to the newly established monitoring and logging stacks, including metrics, dashboards, and alert rules.
  • Harden observability frameworks alongside the pipelines they instrument, so regressions surface before users find them.

Key Responsibilities CI/CD, Rollout & Rollback:

  • Harden CI/CD pipelines for legacy-to-modern infrastructure transitions, including build, test, and package promotion stages.
  • Design and maintain rollout and rollback mechanisms that make large-fleet changes reversible.
  • Automate repetitive operational work and replace manual runbook steps with tested, reviewed code.

Key Responsibilities Tier-2 Operations & Migration Support:

  • Provide Tier-2 operational support and incident response under a follow-the-sun model, in close partnership with the client's SRE team.
  • Perform hands-on break/fix operations on existing platforms while the modernization proceeds in parallel.
  • Assist application developers with architectural support and troubleshooting during cloud migration phases.
  • Author and maintain technical documentation runbooks, migration procedures, and package and pipeline ownership notes.

__________________________________

Ontrac Solutions has partnered with PinpointVerify to help genuine applicants rise above the noise. Today, qualified candidates are too often overshadowed by fake and fraudulent applications. PinpointVerify gives our recruiters confidence that you are exactly who you say you are and gives you a portable verification credential you can share with any employer.

Applicants who complete verification are prioritized over non-verified candidates with comparable experience. And if you're hired, Ontrac reimburses the full cost of your verification.
Get verified https://pinpointverify.com/ontrac