1

System Reliability Engineer Jobs (NOW HIRING)

SRE Engineer (Technical Lead)

Austin, TX · On-site

$56.50 - $75/hr

Info Way Solutions is seeking an experienced SRE Engineer (Technical Lead) to support and enhance system reliability, scalability, and performance across AWS environments. The ideal candidate should ...

Senior SRE Engineer

New York, NY · On-site

$62.25 - $82.75/hr

Lead and manage the SRE team to uphold system reliability, availability, and performance standards. * Design, implement, and optimize scalable infrastructure and automation solutions to support ...

Site Reliability Engineer (SRE)

Plano, TX · On-site

$54.50 - $72.50/hr

Monitor system health, performance, and availability using SRE best practices * Implement automation to reduce manual operational work * Troubleshoot production incidents and perform root cause ...

DeVops SRE

Wilmington, DE · On-site

$55.25 - $73.50/hr

HCL Global Systems Inc is seeking a GIS Team Member to leverage expertise in SRE DevOps and monitoring technologies. The role involves optimizing systems using tools like Splunk, Python, and AWS ...

$42 - $55.75/hr

We're seeking an SRE to ensure the reliability and performance of our clients' critical systems. You'll work on observability, incident response, and platform reliability. Responsibilities

New

Site Reliability Engineer (SRE)

Seattle, WA · On-site

$65 - $86.25/hr

We're seeking an SRE to ensure the reliability and performance of our clients' critical systems. You'll work on observability, incident response, and platform reliability. Responsibilities

New

Site Reliability Engineer (SRE)

Plano, TX · On-site

$54.75 - $72.75/hr

Design and maintain highly available scalable and faulttolerant systems * Implement and manage ... Drive adoption of SRE practices like SLIs SLOs and error budgets * Ensure performance optimisation ...

GPU System Reliability Engineer Lead

San Carlos, CA · On-site

$123K - $155K/yr

... reliability engineering, or silicon validation on server-class compute systems. * Deep understanding of CPU and GPU architecture, including memory subsystems (DDR, HBM), cache hierarchies, and ...

SME SRE Observability

Fremont, CA · On-site

$62.50 - $83.25/hr

... ) with a strong focus on Observability. The ideal candidate will be responsible for designing ... ensure high system reliability, performance, and scalability in a production environment.

Site Reliability Engineer

Austin, TX · On-site

$56.50 - $75/hr

The role involves owning production infrastructure, enhancing system reliability, and collaborating with engineering teams to ensure robust and scalable systems. Responsibilities : • Design, build ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

The SRE Engineer will improve the reliability, availability, performance, and operational resilience of mission-critical systems for a federal enterprise program. Responsibilities : • Define ...

SRE

Charlotte, NC · On-site

$55.75 - $74/hr

Role: SRE Location: Charlotte, NC Skills: Grafana, Python, Splunk, Linux, Scripting. Microsoft 360 ... This role will focus on ensuring the stability, performance, and efficiency of the systems while ...

SRE Engineer

Fremont, CA · On-site

$62.50 - $83.25/hr

The role involves responsibilities related to site reliability engineering, ensuring system uptime and performance, and collaborating with development teams to enhance system reliability. Company

Support production systems including on-call, incident response, and RCA * Collaborate with SRE and Security teams to ensure system reliability and scalability * Drive architectural decisions and ...

Showing results 41-60

System Reliability Engineer information

See salary details

$61K

$118K

$141K

How much do system reliability engineer jobs pay per year?

As of Sep 11, 2026, the average yearly pay for system reliability engineer in the United States is $117,973.00, according to ZipRecruiter salary data. Most workers in this role earn between $102,500.00 and $129,000.00 per year, depending on experience, location, and employer.
More about System Reliability Engineer jobs

Who are the top companies hiring for System Reliability Engineer jobs?

The top employers for System Reliability Engineer jobs are:

What are popular job titles related to System Reliability Engineer jobs?

For System Reliability Engineer jobs, the most frequently searched job titles are:

Infographic showing various System Reliability Engineer job openings in the United States as of August 2026, with employment types broken down into 1% As Needed, 84% Full Time, 12% Part Time, and 3% Contract. Highlights an 91% Physical, 2% Hybrid, and 7% Remote job distribution, with an average salary of $117,973 per year, or $56.7 per hour.

Site Reliability Engineer

Pittsburgh, PA • On-site

System One Holdings, LLC
Business Consulting Services • 5 - 10K employees

$55.25 - $73.50/hr

Full-time

Re-posted 8 hours ago


Job description

Senior Site Reliability Engineer (SRE)
Location: Pittsburgh, PA / Cleveland, OH / Dallas, TX
FTE

Position Overview
We are seeking an experienced Senior Site Reliability Engineer (SRE) to support production operations, application reliability, performance management, and continuous improvement initiatives.
The selected candidate will work closely with production support and engineering teams to ensure critical internal and external applications maintain appropriate levels of availability, reliability, and uptime.
This role requires strong experience in production support, incident management, monitoring, troubleshooting, log analysis, automation identification, infrastructure technologies, databases, and application servers. The SRE will also provide technical leadership and collaborate with geographically distributed teams.
Key Skills
  • Site Reliability Engineering (SRE)
  • Production Support / Application Support
  • Incident & Problem Management
  • Linux
  • Windows Server
  • Oracle / PL/SQL / DB2
  • Dynatrace / DT Managed
  • GlassBox / ITCAM / TrueSight / OEM
  • Tomcat / Apache / WebSphere (WAS) / IIS
  • REST & SOAP Web Services
  • Log Analysis & Troubleshooting
  • AIOps / NLP
  • Monitoring & Performance Management
  • Automation
  • Root Cause Analysis
  • Business Analytics
  • Agile
  • Technical Leadership
  • Client-Facing Production Support

Responsibilities
  • Monitor distributed systems and proactively identify potential production issues.
  • Support troubleshooting and participate in on-call activities.
  • Manage, track, and coordinate production incidents and application outages.
  • Lead incident-analysis and problem-management meetings.
  • Identify opportunities for operational and production-support automation.
  • Monitor applications and related infrastructure to maintain system reliability.
  • Coordinate follow-up activities through incident resolution and closure.
  • Troubleshoot complex application issues using system and application logs.
  • Participate in critical incident calls and contribute technical expertise toward resolution.
  • Perform root cause analysis and recommend corrective actions.
  • Research and reproduce user issues to validate solutions.
  • Resolve technical problems that cannot be handled by junior team members.
  • Provide technical guidance and solutions to the production-support team.
  • Introduce process improvements and innovative solutions for operational challenges.
  • Develop and maintain SOPs, operational procedures, and knowledge documentation.
  • Collaborate with offshore and geographically distributed teams.
  • Work with client technical teams, SMEs, and leadership.
  • Support extended or weekend hours when required during critical production events.
  • Participate in overlapping business-hour shifts for critical meetings and activities.

Required Qualifications
  • 5+ years of overall IT experience.
  • 2-3 years of business analytics and technical leadership experience.
  • Strong experience with production/application support in a client-facing environment.
  • Strong understanding of Site Reliability Engineering and production operations.
  • Hands-on experience troubleshooting production applications and analyzing log files.
  • Strong knowledge of system-management, monitoring, and support analytics tools.
  • Experience with incident management, root cause analysis, and problem resolution.
  • Strong understanding of AIOps and NLP concepts.
  • Experience identifying opportunities for automation and process improvement.
  • Strong problem-solving and analytical capabilities.
  • Ability to recommend efficient and cost-effective technical solutions.
  • Experience working with geographically distributed/onshore-offshore teams.
  • Excellent client-facing verbal and written communication skills.

Database Technologies
Strong knowledge of:
  • Oracle
  • PL/SQL
  • DB2

Web Services
Experience developing and consuming:
  • REST APIs
  • SOAP Web Services

Experience should preferably be within an operational/production environment.
Application Servers / Web Servers
Strong knowledge of:
  • Tomcat
  • Apache
  • WebSphere (WAS)
  • IIS

Operating Systems
  • Extensive experience with Linux
  • Good understanding of Windows Server
  • Linux and Windows server configuration and troubleshooting

Monitoring & Support Tools
Experience with monitoring tools such as:
  • Dynatrace
  • Dynatrace Managed / DT Managed
  • GlassBox
  • ITCAM / ITCAMS
  • TrueSight
  • Oracle Enterprise Manager (OEM)

Additional Skills
  • Agile methodology
  • SOP and technical documentation
  • Performance management
  • System reliability and availability
  • Production incident coordination
  • Technical research and solution evaluation
  • Process improvement
  • Automation opportunity identification
  • Strong stakeholder and client communication

#M1
#DI-CB2
#L1 - KB1
Ref: #404-IT Pittsburgh

System One logo

About System One

Sourced by ZipRecruiter

System One helps employers get work done more efficiently and economically without compromising quality. Over our 35+ year history, we've helped connect thousands of talented people with innovative companies. The excitement of a perfect fit motivates us every single day.

Industry

Business consulting services and recruiting and staffing services

Company size

5,001 - 10,000 Employees

Headquarters location

Pittsburgh, PA, US