1

Site Reliability Engineer Jobs in Draper, UT (NOW HIRING)

Reliability Engineer

Salt Lake City, UT

$98K - $123K/yr

Job Title Reliability Engineer Summary Reliability, Maintenance, and Engineering (RME) is hiring ... Lead the Root Cause Analysis and Permanent Corrective Action activities at the site. * Identify ...

Principal Reliability Engineer

Provo, UT

$97K - $122K/yr

US-AZ-TUCSON-801 ~ 1151 E Hermans Rd ~ BLDG 801 (External Site) Position Role Type: Onsite U.S ... Coach other Reliability engineers and work with functional leadership in identifying and ...

Reliability Engineer

Salt Lake City, UT · On-site

$99K - $125K/yr

Requires collaboration with cross-functional teams, including engineering, maintenance, operations, and quality, to achieve optimal equipment reliability and production efficiency. Requirements ...

US-AZ-TUCSON-801 ~ 1151 E Hermans Rd ~ BLDG 801 (External Site) Position Role Type: Onsite U.S ... We are looking for a Senior Reliability Engineer to join our team located in Tucson, AZ. What You ...

Reliability Engineer II

Provo, UT

$97K - $122K/yr

US-AZ-TUCSON-801 ~ 1151 E Hermans Rd ~ BLDG 801 (External Site) Position Role Type: Onsite U.S ... We are looking for a Reliability Engineer II to join our team located in Tucson, AZ. What You Will ...

Showing results 21-40

Site Reliability Engineer information

See Draper, UT salary details

$10

$59

$85

How much do site reliability engineer jobs pay per hour?

As of Aug 6, 2026, the average hourly pay for site reliability engineer in Draper, UT is $59.59, according to ZipRecruiter salary data. Most workers in this role earn between $51.25 and $68.08 per hour, depending on experience, location, and employer.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often involves on-call duties, troubleshooting complex issues, and working with automation tools, which can contribute to work-related stress but also offers opportunities for skill development and problem-solving.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.
What job categories do people searching Site Reliability Engineer jobs in Draper, UT look for? The top searched job categories for Site Reliability Engineer jobs in Draper, UT are:
What cities near Draper, UT are hiring for Site Reliability Engineer jobs? Cities near Draper, UT with the most Site Reliability Engineer job openings:
Infographic showing various Site Reliability Engineer job openings in Draper, UT as of August 2026, with employment types broken down into 82% Full Time, and 18% Contract. Highlights an 64% In-person, 5% Hybrid, and 31% Remote job distribution, with an average salary of $123,945 per year, or $59.6 per hour.

Site Reliability Engineer I- Operations

Utah Valley University

UT • On-site

$23.95 - $28.18/hr

Part-time

Posted 10 days ago


Utah Valley University rating

6.8

Company rating: 6.8 out of 10

Based on 30 frontline employees who took The Breakroom Quiz

476th of 615 rated colleges and universities


Job description

Salary: $23.95 - $28.18 Hourly
Location : DX Building
Job Type: Part-Time Staff
Job Number: FY2706318
Division: VP Digital Transformation/CIO
Opening Date: 07/27/2026
Closing Date: 8/10/2026 11:59 PM Mountain
First Review Date: 08/03/2026
Required Documents Needed to Apply: Resume
Applicant Support: 1-855-524-5627
Support@schooljobs.com
Position Announcement
Join our team as a Site Reliability Engineer I - Operations and play a critical role in maintaining the reliability, security, and performance of the technology that powers our organization. In this position, you'll work with modern infrastructure, automation tools, and cloud technologies to design and improve resilient systems, troubleshoot complex technical challenges, and help minimize downtime across enterprise platforms. You'll collaborate with cross-functional teams, contribute to continuous improvement initiatives, and leverage industry-leading tools such as Jira, Confluence, Opsgenie, and CI/CD pipelines to enhance operational efficiency. If you enjoy solving technical challenges, building reliable systems, and making a measurable impact through technology, this role offers an excellent opportunity to develop your expertise while supporting mission-critical services.
Summary of Responsibilities
  • Under close supervision, epic plans and executes projects related to the three pillars of IT operations (operational processes, change incident problem, and Ops readiness. Assists in the execution of monitoring systems and alert configurations so that Operations knows about outages before users.
  • Collaborates with leadership on the creation, facilitation, and integration of documentation, including installation steps, standard operating procedures, incident runbooks, and disaster recovery documentation into a curated change/incident/problem management library. Assists Network, Application, database, and systems administrators with the enforcement of standard procedures, acts as a remote hand within a secure data center, and maintains all required supplies and tooling for the deployment of physical enterprise equipment.
  • As an incident commander, participates in business-hour on-call rotation, evaluating incoming alerts for validity and dispatching the appropriate SME to resolve issues. Executes public communications in accordance with Operational standard procedures, informing stakeholders of possible service disruptions. Maintains the integrity of Runbooks.
  • Performs other job-related duties as assigned.

Qualifications / Licenses / Certifications
Graduation from an accredited college or university with an associate's degree and two years' experience OR any combination of education and experience totaling four years.
Licenses or Certifications:
Comptia A+, Comptia network+, Comptia security+, and Comptia Linux+
Knowledge / Skills / Abilities
Knowledge
  • Knowledge of Linux and Windows Operating systems, TCP/IP fundamentals, firewall management, and anti-virus software.
  • Knowledge of best practices for securing operating systems, data center maintenance, and network setup.
  • Knowledge of various Monitoring solutions such as Prometheus, PRTG, Site24x7, TestCafe, Selenium, Splunk, NewRelic, Azure Monitor, and AWS CloudWatch.
  • Knowledge of storage technologies such as SAN or NAS.
  • Knowledge of Azure Active Directory, Active Directory, and LDAP.
  • Knowledge of load balancing, clustering, and enterprise server architecture.
  • Knowledge of Relational Database principles and databases/languages such as PL/SQL, MySQL, SQL Server, Oracle, Microsoft SQL, or MS Access.
  • Knowledge of the Atlassian Suite, including Jira, Confluence, Status Page, and Opsgenie.
  • Knowledge of Scrum/Agile principles as applicable to a DevOps Team.
Skills
  • Communicate effectively in normal and high-pressure situations verbally and through written mediums.
  • Perform basic server, system, and application procedures such as managing user access, performing maintenance, and troubleshooting.
  • Skills in troubleshooting hardware and software problems and researching technical issues.
  • Experience using basic CLI tools in Windows and Linux operating systems to troubleshoot and gather information.
  • Skills in customer service and interpersonal communication, both verbally and written.
  • Basic scripting and programming skills in languages such as Python, JavaScript, JSON, SQL, Bash, TestCafe, and Selenium.
  • Experience with instant communication and team collaboration platforms like MS Teams, Slack, or Jitsi.
  • Skills in working in an ITSM solution such as Jira, ServiceNow, Asana.
Abilities
  • Ability to identify, research, troubleshoot, and implement solutions for hardware and software problems.
  • Ability to work in a customer service, team-oriented, collaborative, Scrum/Agile environment.
  • Highly self motivated with the ability to learn quickly and accept feedback from peers.
  • Ability to learn the implement process, and maintenance procedures for new technologies, equipment, hardware, and software such as operating systems, ITSM tools, monitoring solutions, and data center management.
  • Ability to act as an "on-call" incident commander for communicating outages between customers, subject matter experts, teams, and leaders.
  • Ability to create proposals in visually-pleasing and user-friendly language. Ability to think critically and solve complex problems.
  • Ability to perform tasks in a timely and professional manner.

EEO Statement:
UVU employment decisions are made on the basis of an applicant's qualifications and ability to perform the job without regard to race, color, religion, national origin, sex, sexual orientation, gender identity, gender expression, age (40 and over), disability, veteran status, pregnancy, childbirth, or pregnancy-related conditions, genetic information, or other bases protected by applicable federal, state, or local law.
Utah Valley University's dedication to exceptional care offers quality service and benefits to employees while staying committed to meeting the needs of a diverse workforce.
Although part-time and adjunct (less than 30 hours per week) employees are ineligible for the full-time employee benefits package, the university is pleased to offer the following:
  • Undergraduate tuition remission benefit:
    • Part-time: waives 100 percent of tuition and general student fees for up to 3 credit hours or 1 course after 6 consecutive months of employment and a minimum of 480 hours worked (dependents do not qualify)
    • Adjunct: waives 100 percent of tuition and general student fees for one course per semester in which there is an active teaching assignment at the university (dependents do not qualify)
  • Access to the Employee Assistance Program (EAP)
  • Access to Campus Recreation and Wellness
  • Exclusive cash rewards through Utah Community Credit Union's (UCCU) branch on UVU's campus when becoming a member

01
What is your highest level of education?
  • High School or equivalent
  • Associates
  • Bachelors
  • Masters
  • Ph.D.
  • Ph.D. (abd)
  • Juris Doctorate

02
How many years of experience do you have in this type of position?
  • <1
  • 1
  • 2
  • 3
  • 4
  • 5
  • 6
  • 7
  • 8+

03
Do you hold the required certifications: CompTIA A+, CompTIA Network+, CompTIA Security+, and CompTIA Linux+?
  • Yes
  • No

Required Question

What Utah Valley University employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom