1

Systems Reliability Engineer Jobs (NOW HIRING)

Systems Reliability Engineer II

Durham, NC · On-site

$99K - $124K/yr

Our customer base is growing rapidly and we're looking for top-tier Systems Reliability Engineers who share our passion for customer success to join our team! We drive the success of our customers ...

New

next page

Showing results 1-20

Systems Reliability Engineer information

See salary details

$61K

$118K

$141K

How much do systems reliability engineer jobs pay per year?

As of Aug 5, 2026, the average yearly pay for systems reliability engineer in the United States is $117,973.00, according to ZipRecruiter salary data. Most workers in this role earn between $102,500.00 and $129,000.00 per year, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive as a systems reliability engineer?

To thrive as a Systems Reliability Engineer, you need expertise in infrastructure management, automation, and software engineering, often supported by a degree in computer science or a related field. Familiarity with tools like Kubernetes, Docker, CI/CD pipelines, monitoring systems (e.g., Prometheus, Grafana), and relevant cloud certifications (AWS, GCP, or Azure) is typically required. Strong problem-solving abilities, communication skills, and a proactive mindset help you prevent and resolve incidents efficiently. These skills ensure systems remain robust, scalable, and highly available, which is critical for maintaining business continuity and user trust.

What is a systems reliability engineer?

Systems Reliability Engineers (SREs) are IT professionals responsible for ensuring the reliability, availability, and performance of software systems and infrastructure. They combine software engineering and systems administration skills to automate processes, monitor system health, respond to incidents, and improve system resilience. SREs work closely with development and operations teams to optimize deployment pipelines, manage outages, and implement best practices for scalability and reliability. Their goal is to minimize downtime and ensure a seamless user experience.

How does a systems reliability engineer typically collaborate with development and operations teams to ensure system stability?

Systems Reliability Engineers (SREs) work closely with both development and operations teams to bridge the gap between software engineering and IT operations. They participate in design reviews to ensure reliability is built into new features, coordinate with developers to automate deployments, and work with operations to monitor system health and respond to incidents. By fostering a culture of shared responsibility for uptime and performance, SREs help streamline troubleshooting and drive improvements across the organization. Regular communication and joint post-incident reviews are key practices in this collaborative environment.

What is the difference between Systems Reliability Engineer vs DevOps Engineer?

AspectSystems Reliability EngineerDevOps Engineer
Primary FocusEnsuring system reliability, availability, and performanceAutomating deployment, integration, and continuous delivery
Skills & CertificationsSRE certifications, Linux, scripting, monitoring toolsCI/CD tools, cloud platforms, scripting, automation
Work EnvironmentOperations, infrastructure, and reliability teamsDevelopment and operations collaboration
Industry UsageTech, finance, e-commerceTech, startups, cloud services

While both roles focus on improving system performance, Systems Reliability Engineers primarily concentrate on maintaining system uptime and reliability, whereas DevOps Engineers focus on streamlining development and deployment processes. Both roles often collaborate but serve different core functions within an organization.

More about Systems Reliability Engineer jobs
What cities are hiring for Systems Reliability Engineer jobs? Cities with the most Systems Reliability Engineer job openings:
What states have the most Systems Reliability Engineer jobs? States with the most job openings for Systems Reliability Engineer jobs include:
What job categories do people searching Systems Reliability Engineer jobs look for? The top searched job categories for Systems Reliability Engineer jobs are:
Infographic showing various Systems Reliability Engineer job openings in the United States as of July 2026, with employment types broken down into 96% Full Time, 1% Part Time, and 3% Contract. Highlights an 87% Physical, 5% Hybrid, and 8% Remote job distribution, with an average salary of $117,973 per year, or $56.7 per hour.

Systems Reliability Engineer

Members 1st Federal Credit Union

Enola, PA • On-site

$95K - $119K/yr

Full-time

Posted 7 days ago


Members 1st Federal Credit Union rating

8.9

Company rating: 8.9 out of 10

Based on 15 frontline employees who took The Breakroom Quiz


Job description

Overview
When you join the Members 1st team, you become part of something much bigger than a credit union. You become part of our faM1ly-a tight-knit bunch with big dreams and even bigger values. It is an exciting time for us as we continue to grow, and we hope that you will choose to grow along with us. Wanting the absolute best for our associates means more than just competitive pay. It means fantastic healthcare, paid benefits, opportunities for professional advancement and work-life balance - and best of all, a place where you are accepted and respected for your individuality.
Responsibilities
The Systems Reliability Engineer provides advanced operational support for platform, application, security, and infrastructure issues escalated from Tier 1 support. This role leverages AI-assisted operational tooling to receive and interpret event monitoring alerts, visualize system health and operational planes, and accelerate troubleshooting, impact analysis, and runbook execution. The engineer is responsible for triaging and resolving complex incidents, executing system-level support tasks, and maintaining IT operational runbooks using both traditional documentation and AI-assisted querying and revision techniques.
Key Responsibilities include but are not limited to:
• Triage and resolve Tier II support issues escalated by the Help Desk, including application access issues, system performance degradation, service outages, and infrastructure-level incidents.
• Perform system-level operational tasks to restore service, including IIS resets, application pool recycling, service restarts, and execution of approved PowerShell scripts in accordance with security and change controls.
• Monitor system health and operational telemetry, review logs, and proactively identify and remediate recurring or emerging issues.
• Research, diagnose, and resolve complex user issues across applications, integrations, and infrastructure components.
• Collaborate with Tier III, infrastructure, security, and application teams to escalate, coordinate, and resolve complex technical incidents.
• Manage incident, request, change, and problem tickets within the ServiceNow ITSM platform in alignment with established SLAs, escalation procedures, and governance standards.
• Participate in problem management activities by identifying root causes, documenting known errors, and contributing to permanent corrective actions.
• Support change management efforts, including review, testing coordination, execution of approved changes, and post-change validation.
• Analyze trends in incidents and requests to provide insights that support service reliability, operational maturity, and continuous improvement.
• Create, execute, and maintain IT operational runbooks to ensure consistent, repeatable support and response procedures.
• Keep runbooks current based on changes to systems, processes, or technologies, coordinating updates with engineering, platform, and support teams.
• Document support procedures, troubleshooting steps, and resolutions in ServiceNow Knowledge Base articles.
• Capture successful support outcomes and lessons learned in shared knowledge repositories to improve Tier I efficiency and enable self-service.
• Identify and analyze business needs, gather requirements, and help define scope and objectives for system enhancements or operational improvements.
• Research business requirements and document relationships between users, business processes, data, applications, and devices to support impact analysis.
• Translate business requirements into clear, actionable application or operational requirements.
• Make recommendations for technology-based solutions or process improvements using new or existing tools.
• Assist with product refinement and backlog prioritization by contributing operational insights and support data.
• Use AI-enabled monitoring and event intelligence tools to receive, interpret, and prioritize operational alerts across applications, infrastructure, security, and integrations.
• Leverage AI-assisted dashboards and visualizations to assess system health, service dependencies, and operational risk.
• Query AI tools to identify upstream and downstream impacts during incidents, changes, or performance issues.
• Use AI-assisted analysis to validate, enhance, and continuously improve operational runbooks, identifying gaps and outdated procedures.
• Collaborate with engineering and platform teams to validate AI-suggested remediation actions prior to execution in regulated production environments.
• Provide feedback to improve AI operational tooling accuracy, relevance, and governance over time.
• Participate in on call services providing after-hours support.
SKILLS
• Strong analytical and troubleshooting skills, with experience managing and resolving complex incidents and participating in incident and problem management activities using ITIL-aligned practices
• Solid understanding of software engineering concepts and delivery disciplines, including iterative development, CI/CD pipelines, unit testing, and how applications are built, monitored, supported, and maintained in production environments
• Hands-on technical experience supporting Windows Server environments, IIS administration, PowerShell scripting, and foundational system administration, including working knowledge of TCP/IP networking, DNS, load balancing, and SSL/TLS certificate management
• Experience developing and maintaining IT operational documentation, including runbooks and knowledge base articles, ensuring consistent, repeatable support and continuous improvement
• Working knowledge of ITSM processes and tools, particularly incident, request, change, and problem management within platforms such as ServiceNow
• Experience leveraging AI-assisted tools for IT Operations, including operational monitoring, alert analysis, impact assessment, incident triage, runbook execution, and documentation updates, with demonstrated judgment in validating AI-generated insights before execution in production environments
• Ability to query and prompt AI tools effectively to support operations, troubleshooting, and continuous improvement initiatives
• Strong time management, evaluation, and adaptability skills, with the ability to stay current on emerging IT tools, trends, and operational best practices
• Experience with AI-enabled ITSM, AIOps, or observability platforms (for example ServiceNow AI features, AIOps tools, or intelligent monitoring platforms) preferred
• Familiarity with service dependency mapping, topology visualization, or blast-radius analysis tools preferred
• Experience improving operational workflows through automation, AI augmentation, or decision-support tooling preferred
COMPETENCIES
• Accountability and self-management
• Communication
• Effective knowledge
• Innovation and problem-solving
• Teamwork and leadership
WORKING CONDITIONS/PHYSICAL DEMANDS
• Ability to communicate effectively in English, both orally and in writing
• Visually able to perform activities such as preparing and analyzing data and figures, viewing a computer terminal, and extensive reading
• Ability to sit for extended time periods
• Sufficient manual skill for operation of PC keyboard and other standard office equipment
• Ability to travel, including occasional overnight travel
• Ability to exert minimum amounts of force occasionally to lift, carry, push, pull or move objects
Qualifications
1-3 years of related experience
Does this position require a valid Drivers License?
No
Education Level
General and business knowledge equivalent to a bachelor's degree
Certifications
IIBA certification preferred; PSPO certification preferred
About Us
At Members 1st, we look for individuals who will show up as their whole self because we value diversity, inclusion, and belonging, as well as people who believe in the philosophy of, WE>me. To be sure you align with our company mission, vision, values and culture reference the information below.
Company Culture is at Our Core
If there is one concept, we want you to understand about us, it is this. WE. It is a simple little word but means everything here. We think as one. One faM1ly. One community. One place where everyone belongings. Everything we do is in the best interest of all of us.
What WE Believe
Our Missions: WE serve our members, associates, and communities through support, empowerment, and meaningful relationships.
Our Vision: WE are growing our faM1ly by delivering everything they need to live well financially, through all life's moments and milestones.
Our Values: WE deliver unparalleled experiences through a culture of WE. WE > me. WE are servant leaders- at work and in the communities we serve. WE are financially safe and sound stewards of members dollars. WE are faM1ly.
Join a company that grows with you - personally and professionally
Equal Opportunity Employer
Members 1st provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
Compensation Overview
We are excited to offer a competitive salary for this position. This figure serves as the entry point in our salary range, and there is potential for the actual salary to be higher based on a variety of factors, such as your experience, skills, education, and location. We believe in recognizing and rewarding talent, so our compensation packages are thoughtfully designed to reflect the unique qualifications and contributions of each candidate.
The minimum salary for this position is:
$60,000.00+/yr
Certifications
IIBA certification preferred; PSPO certification preferred

What Members 1st Federal Credit Union employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom