1

Reliability Manager Jobs in Springfield, MO (NOW HIRING)

Lead team members engaged in measuring, testing, and tabulating data concerning materials, product, or process quality and reliability. * Manage the gauge system and actively pursue cost reduction ...

Lead team members engaged in measuring, testing, and tabulating data concerning materials, product, or process quality and reliability. * Manage the gauge system and actively pursue cost reduction ...

Web Services Manager

Springfield, MO · On-site

$110 - $140/hr

... standards of performance, reliability, and security across all web service platforms ... Proven ability to manage priorities, allocate resources, delegate work, and meet project deadlines.

Manager, Maintenance in Springfield, MO Build your future at Curia, where our work has the power to ... Essential Responsibilities Maintenance Operations & Reliability * Provide accountable, day-to-day ...

Manager, Maintenance in Springfield, MO Build your future at Curia, where our work has the power to ... Essential Responsibilities Maintenance Operations & Reliability * Provide accountable, day-to-day ...

This role ensures equipment reliability, adherence to preventive and predictive maintenance schedules, and compliance with food safety and regulatory standards. The Maintenance Manager will assemble ...

This role ensures equipment reliability, adherence to preventive and predictive maintenance schedules, and compliance with food safety and regulatory standards. The Maintenance Manager will assemble ...

next page

Showing results 1-20

Reliability Manager information

See Springfield, MO salary details

$56.4K

$106.9K

$153.3K

How much do reliability manager jobs pay per year?

As of Aug 27, 2026, the average yearly pay for reliability manager in Springfield, MO is $106,859.00, according to ZipRecruiter salary data. Most workers in this role earn between $86,000.00 and $127,300.00 per year, depending on experience, location, and employer.

What does a reliability manager do?

A Reliability Manager is responsible for ensuring that equipment, processes, and systems operate efficiently and consistently to minimize downtime and maximize performance. They develop and implement reliability strategies, conduct root cause analyses, and oversee preventive and predictive maintenance programs. Their role involves working closely with maintenance teams, engineers, and production staff to improve asset reliability and extend equipment lifespan. Additionally, they analyze failure data, recommend improvements, and help optimize operational costs through reliability-centered maintenance practices.

What are the key skills and qualifications needed to thrive as a reliability manager?

A Reliability Manager needs strong analytical skills, a solid background in engineering or maintenance, and experience with reliability-centered maintenance methodologies. Familiarity with tools like Failure Mode and Effects Analysis (FMEA), Root Cause Analysis (RCA), and certifications such as Certified Reliability Engineer (CRE) are often required. Leadership, problem-solving, and the ability to communicate complex technical information clearly are crucial soft skills for this role. These skills help ensure equipment uptime, optimize maintenance processes, and foster a culture of continuous improvement within the organization.

What are popular job titles related to Reliability Manager jobs in Springfield, MO?

For Reliability Manager jobs in Springfield, MO, the most frequently searched job titles are:

What job categories do people searching Reliability Manager jobs in Springfield, MO look for?

The top searched job categories for Reliability Manager jobs in Springfield, MO are:

What cities near Springfield, MO are hiring for Reliability Manager jobs?

Cities near Springfield, MO with the most Reliability Manager job openings:

Infographic showing various Reliability Manager job openings in Springfield, MO as of August 2026, with employment types broken down into 1% As Needed, 81% Full Time, 16% Part Time, 1% Temporary, and 1% Contract. Highlights an 79% Physical, 3% Hybrid, and 18% Remote job distribution, with an average salary of $106,859 per year, or $51.4 per hour.

Senior Site Reliability Engineer (Onsite Role)

O'Reilly Auto Parts

Springfield, MO • On-site

$53.75 - $71.25/hr

Full-time

Medical, Dental, Vision, Retirement, PTO

Re-posted 14 days ago


O'Reilly Auto Parts rating

5.2

Company rating: 5.2 out of 10

Based on 1,901 frontline employees who took The Breakroom Quiz

567th of 738 rated retailers


Job description

Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firm's most critical customer-facing microservices that power all eCommerce channels. This role applies Google-inspired SRE principles to balance feature velocity and system reliability using Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.
The role combines software engineering, cloud engineering, automation, and production operations, with a strong emphasis on building systems that are observable, resilient, and operable by default.
This is an on-site position located in Springfield, MO. Remote work is not an option for this position.
Primary Responsibilities
  • Define, implement, and own SLIs, SLOs, and error budgets for critical microservices in collaboration with product and engineering teams.
  • Use error budgets to influence release decisions, prioritize reliability initiatives, and manage operational risk.
  • Design and maintain observability platforms, including metrics, logs, traces, and real-time telemetry.
  • Track, manage, and reduce operational toil by converting repetitive operational tasks into Jira stories and epics with clear ownership and measurable outcomes.
  • Design, implement, and validate resiliency mechanisms such as graceful degradation, redundancy, automated failover, and disaster recovery.
  • Lead incident response efforts, act as an escalation point for high-severity incidents, and drive blameless postmortems.
  • Capture incident action items and reliability improvements in Jira, ensuring accountability, closure, and continuous improvement.
  • Partner with Scrum teams to improve reliability through release readiness reviews, production change validation, and testing strategies.
  • Perform deep root cause analysis, debugging, and performance tuning across distributed systems.
  • Promote shift-left reliability practices by embedding operability, monitoring, and failure testing early in the SDLC.
  • Drive continuous improvement through automation, self-healing systems, chaos engineering, and capacity planning.
  • Maintain runbooks, playbooks, and knowledge repositories, linking documentation to Jira tasks to reduce MTTR.
  • Provide technical leadership and mentoring to junior SREs and engineers.
  • Collaborate with global, distributed teams, leveraging Jira for transparent planning, dependency tracking, and execution.
  • Conduct production readiness reviews and ensure services meet operational excellence standards before deployment.
  • Track and improve operational KPIs such as availability, MTTR, MTTD, deployment success rate, and incident recurrence.
  • Collaborate with security and platform teams to ensure reliability, compliance, and operational security best practices are embedded into systems and deployment pipelines.
  • Explore opportunities to leverage AI-driven observability, anomaly detection, and operational automation to improve system reliability and reduce manual effort.

Core Competencies & Qualifications
  • 4+ years of experience in SRE, software engineering, or production operations supporting large-scale eCommerce platforms.
  • Hands-on experience with Java/J2EE-based distributed systems; React experience is a plus.
  • Proven ability to design and operate systems using SLO-driven reliability models.
  • Experience defining and measuring SLIs, including availability, latency, error rates, throughput, and saturation.
  • Good understanding of NoSQL technologies and RDBMS concepts, with the ability to write and troubleshoot database queries.
  • Experience deploying and operating services on cloud platforms such as AWS, Azure, or Google Cloud Platform (GCP).
  • Expertise with observability, APM, and caching tools such as Dynatrace, Splunk, ELK, Akamai, Quantum Metric, and Tealeaf.
  • Strong experience using Jira for backlog management, incident tracking, toil reduction initiatives, and cross-team coordination.
  • Ability to independently own services and drive reliability initiatives end-to-end.
  • Strong communication skills with the ability to influence engineering and product teams.
  • Experience participating in on-call rotations and handling critical/high-severity incidents.

Desired Skills
  • Experience building and operating microservices architectures using Spring Boot, Groovy, React, or similar technologies.
  • Strong understanding of CI/CD pipelines, release automation, and progressive delivery practices.
  • Experience working within eCommerce domains such as Catalog, Customer Data, and Order Management.
  • Familiarity with search platforms including Endeca, Solr, Lucene, and Elasticsearch.
  • Proficiency in scripting and automation using Python, Bash, Ruby, Perl, or PowerShell.
  • Experience with ITSM tools integrated with Jira workflows.
  • Exposure to capacity planning, load testing, and chaos engineering practices.
  • Experience with containerization and orchestration technologies such as Docker and Kubernetes (EKS, AKS, or GKE).
  • Familiarity with Infrastructure as Code (IaC) tools such as Terraform, CloudFormation, or Ansible.
  • Understanding of operational KPIs including availability, MTTR, MTTD, deployment success rate, and incident recurrence metrics.
  • Experience conducting production readiness reviews and implementing operational governance processes.
  • Ability to collaborate with security and platform engineering teams to ensure reliability, compliance, and operational security best practices.
  • Exposure to AI-assisted operations, anomaly detection, intelligent alerting, and automated remediation solutions.
  • Experience designing scalable, self-healing platforms and automation frameworks for cloud-native environments.

O'Reilly Auto Parts has a proven track record of growth and stability. O'Reilly is full of successful career stories and believes in a strong promote-from-within philosophy, encouraging you to grow your career along with the organization.
Total Compensation Package:
  • Competitive Wages & Paid Time Off
  • Stock Purchase Plan & 401k with Employer Contributions Starting Day One
  • Medical, Dental, & Vision Insurance with Optional Flexible Spending Account (FSA)
  • Team Member Health/Wellbeing Programs
  • Tuition Educational Assistance Programs
  • Opportunities for Career Growth

O'Reilly Auto Parts is an equal opportunity employer. The Company does not discriminate on the basis of race, religion, color, national origin or ancestry (including immigration status or citizenship), sex, sexual orientation, gender identity, pregnancy (including childbirth, lactation, and related medical conditions,) age (40 and over), veteran status, uniformed service member status, physical or mental disability, genetic information (including testing or characteristics) or another protected status as defined by local, state, or federal law, as applicable.
Qualified individuals with a disability may be entitled to reasonable accommodation under the Americans with Disabilities Act. If you require a reasonable accommodation during the application or employment process, please send an email to: rar@oreillyauto.com or call (800) 471-7431 option , and provide your requested accommodation, and position details.

What O'Reilly Auto Parts employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom