2

Remote Reliability Engineer Jobs in Orem, UT (NOW HIRING)

Senior Database Reliability Engineer

Provo, UT ยท Remote

$220K - $260K/yr

Remote Job Summary: Join our team as a Senior Database Reliability Engineer, where you'll play a central role in architecting, optimizing, and ensuring the stability of mission-critical PostgreSQL ...

Senior Software Engineer

Provo, UT ยท On-site +1

$120K - $140K/yr

With a remote or in office work environment and a schedule of Monday through Friday, 8 a.m. to 5 p ... Your strong understanding of automated testing frameworks will ensure reliability and performance ...

Device Experience Platform Developer

Lehi, UT ยท Remote

$115K - $218K/yr

Device Experience Platform Engineer - NexthinkJob Summary The Device Experience Platform Engineer ... Develop Remote Actions using PowerShell. Build workflows using Nexthink Flow. Create NQL queries.

Senior Software Engineer

Provo, UT ยท Remote

$120K - $140K/yr

With a remote or in office work environment and a schedule of Monday through Friday, 8 a.m. to 5 p ... Your strong understanding of automated testing frameworks will ensure reliability and performance ...

Senior Staff Agentic AI Engineer

Draper, UT ยท On-site +1

$99K - $134K/yr

Measure and improve agent decision reliability using tools such as LangSmith, Arize Phoenix, and ... Normal office environment. (Remote or Hybrid), 3 to 4 days per month are required in office if ...

Senior Machine Learning Engineer

Lehi, UT ยท On-site +1

$144K - $233K/yr

Build evaluation frameworks and benchmarks to measure model quality, safety, reliability, and task ... Flexible and transparent culture with remote and hybrid work options, generous vacation time, and ...

Senior Machine Learning Engineer

Lehi, UT ยท On-site +1

$144K - $233K/yr

Build evaluation frameworks and benchmarks to measure model quality, safety, reliability, and task ... Flexible and transparent culture with remote and hybrid work options, generous vacation time, and ...

Senior AI Engineer - Agentic

Lehi, UT ยท On-site +1

$98K - $134K/yr

Own the full lifecycle: architecture, implementation, deployment, and ongoing reliability ... Collaborate with engineers, product managers, and AI/ML scientists to deliver end-to-end features ...

Senior AI Engineer - Agentic

Lehi, UT ยท On-site +1

$98K - $134K/yr

Own the full lifecycle: architecture, implementation, deployment, and ongoing reliability ... Collaborate with engineers, product managers, and AI/ML scientists to deliver end-to-end features ...

Regression Tester

Orem, UT ยท On-site +1

$110K - $130K/yr

... and QA engineers to maintain platform stability, reliability, and user experience Key ... Flexible remote/hybrid work options. * Contribute to cutting-edge recruitment and analytics ...

Senior Mining Engineer

Sandy, UT ยท On-site +1

$99K - $136K/yr

... to remote locations * Strong network with mining clients is considered a strong asset ... leader in engineering and consultancy across energy and the built environment, helping to unlock ...

Senior Data Scientist

Lehi, UT ยท On-site +1

$133K - $213K/yr

Partner with machine learning engineers to move successful experiments into production. * Develop ... Flexible and transparent culture with remote and hybrid work options, generous vacation time, and ...

Senior Data Scientist

Lehi, UT ยท On-site +1

$133K - $213K/yr

Partner with machine learning engineers to move successful experiments into production. * Develop ... Flexible and transparent culture with remote and hybrid work options, generous vacation time, and ...

next page

Showing results 1-20

Remote Reliability Engineer information

See Orem, UT salary details

$53K

$102.6K

$122.6K

How much do remote reliability engineer jobs pay per year?

As of Sep 5, 2026, the average yearly pay for remote reliability engineer in Orem, UT is $102,562.00, according to ZipRecruiter salary data. Most workers in this role earn between $89,100.00 and $112,100.00 per year, depending on experience, location, and employer.

What is a remote reliability engineer?

A Remote Reliability Engineer is a professional who works from a remote location to ensure that systems, applications, or infrastructure are reliable, available, and performing well. Their responsibilities typically include monitoring system health, diagnosing issues, implementing preventative measures, and collaborating with teams to improve system reliability. They often use tools for automation, incident response, and performance monitoring, all while working offsite. This role is critical in minimizing downtime and ensuring a smooth user experience, especially for companies with complex technical environments. Remote Reliability Engineers must have strong problem-solving skills and be proficient in cloud technologies, automation, and incident management.

What are the key skills and qualifications needed to thrive as a remote reliability engineer?

To thrive as a Remote Reliability Engineer, you need a strong background in systems engineering, software development, and infrastructure management, often supported by a degree in computer science or a related field. Proficiency with cloud platforms (such as AWS, Azure, or GCP), monitoring tools (like Prometheus, Grafana), and relevant certifications (e.g., AWS Certified DevOps Engineer) is highly valuable. Excellent problem-solving, communication, and collaboration skills are crucial for working effectively across distributed teams and responding to incidents. These abilities ensure system reliability, quick incident resolution, and seamless remote teamwork, which are vital for maintaining high service uptime and user satisfaction.

How do remote reliability engineers typically collaborate with on-site teams to address urgent technical issues?

Remote Reliability Engineers often utilize a combination of video conferencing, instant messaging, and collaborative monitoring tools to stay closely connected with on-site teams. When urgent technical issues arise, they participate in real-time troubleshooting sessions, analyze system logs remotely, and may guide on-site staff through step-by-step resolution procedures. Building strong communication channels and regular check-ins are essential to ensure swift and effective collaboration, even across different time zones. This structure allows Remote Reliability Engineers to contribute significantly to system uptime while working from a distance.

What is the difference between Remote Reliability Engineer vs Remote Site Reliability Engineer?

AspectRemote Reliability EngineerRemote Site Reliability Engineer
CredentialsTypically requires certifications like AWS Certified Solutions Architect, Linux Foundation certificationsSimilar credentials, often with additional focus on site-specific tools and monitoring
Work EnvironmentPrimarily remote, focusing on cloud infrastructure and system reliabilityRemote with some on-site responsibilities, focusing on infrastructure and operational stability
Industry UsageUsed across tech, cloud providers, SaaS companiesCommon in data centers, cloud providers, and large enterprise IT
Search & Comparison IntentOften compared due to overlapping roles in system reliability and cloud infrastructureCompared for on-site vs remote operational responsibilities

The main difference is that Remote Reliability Engineers focus on cloud and system reliability remotely, while Remote Site Reliability Engineers may have some on-site duties related to infrastructure. Both roles require similar skills and certifications but differ in their work environment and specific responsibilities.

What job categories do people searching Remote Reliability Engineer jobs in Orem, UT look for?

The top searched job categories for Remote Reliability Engineer jobs in Orem, UT are:

Infographic showing various Remote Reliability Engineer job openings in Orem, UT as of August 2026, with employment types broken down into 92% Full Time, 2% Part Time, and 6% Contract. Highlights an 87% Physical, 4% Hybrid, and 9% Remote job distribution, with an average salary of $102,562 per year, or $49.3 per hour.

Senior Site Reliability Engineer (SRE)

CenCore LLC

Springville, UT โ€ข On-site, Remote

$52.75 - $70/hr

Full-time

Posted 28 days ago


Job description

Description
The Senior Site Reliability Engineer (SRE) will implement, secure, and operate the cloud infrastructure that supports CenCore Group's proprietary enterprise SaaS platform. This role is responsible for maintaining a scalable, highly available, secure, and reliable cloud environment as the platform grows and supports enterprise customers.
Key Responsibilities
  • Manage, maintain, and improve AWS-based cloud infrastructure supporting enterprise SaaS operations.
  • Operate and support Kubernetes environments, including Amazon EKS.
  • Own platform reliability, scalability, availability, disaster recovery readiness, and operational resilience.
  • Design and support cloud networking, load balancing, routing, traffic management, and related infrastructure components.
  • Implement and maintain monitoring, alerting, logging, and observability solutions to support proactive issue detection and response.
  • Establish and document operational standards, Service Level Objectives (SLOs), incident response processes, and reliability best practices.
  • Partner with software engineering and product teams to improve application performance, platform stability, and deployment reliability.
  • Apply security best practices across IAM, secrets management, encryption, vulnerability remediation, access controls, and production operations.
  • Support production operations, troubleshoot critical issues, and participate in incident resolution as needed.

Requirements
Required Qualifications
  • Active Top Secret clearance with SCI eligibility
  • Professional experience supporting cloud infrastructure, site reliability, DevOps, platform engineering, or systems engineering functions.
  • Hands-on experience with AWS cloud services and production cloud operations.
  • Experience administering or operating Kubernetes environments.
  • Working knowledge of infrastructure reliability, availability, scalability, incident response, and operational support practices.
  • Experience implementing monitoring, logging, alerting, or observability tools.
  • Ability to troubleshoot complex production issues and coordinate resolution across technical teams.
  • Strong understanding of cloud security fundamentals, including identity and access management, encryption, secrets management, and vulnerability remediation.
  • Ability to document technical processes, standards, and operational procedures.

Preferred Qualifications
  • Experience with AWS services such as EKS, ALB, VPC, CloudFront, Route 53, RDS/Aurora, S3, and IAM.
  • Experience with Terraform or other Infrastructure as Code tools.
  • Experience with monitoring platforms such as Datadog, CloudWatch, Grafana, Prometheus, or similar tools.
  • PostgreSQL administration, performance tuning, or database operations experience.
  • Experience supporting enterprise SaaS, cloud-native applications, or customer-facing production platforms.
  • Experience developing disaster recovery, operational readiness, or production support documentation.

Skills / Competencies
  • Cloud infrastructure operations and automation
  • Platform reliability, scalability, and performance optimization
  • Kubernetes administration and containerized application support
  • Monitoring, observability, and incident response
  • Cloud security and operational risk awareness
  • Technical troubleshooting and root cause analysis
  • Cross-functional collaboration with engineering, product, and operations teams
  • Clear technical documentation and process improvement

Work Environment and Physical Requirements
This role is primarily performed in a professional office or remote technology environment, depending on business needs and position requirements. Work involves regular use of a computer, collaboration tools, and cloud-based systems. The position may require participation in production support, incident response, or after-hours troubleshooting as needed. Physical requirements are generally sedentary and include prolonged periods of sitting, computer use, and communicating with internal teams.