1

Customer Reliability Engineer Jobs in Utah (NOW HIRING)

Site Reliability Engineer III - Neovest

Orem, UT ยท On-site

$49.50 - $65.75/hr

Understands service level indicators and utilizes service level objectives to proactively resolve issues before they impact customers * Supports the adoption of site reliability engineering best ...

Maintenance & Reliability Engineering The Regional Reliability Manager's main goal is to boost the ... Interprets and communicates customer requirements to plant production and/or support groups.

Sr. Cloud Platform Engineer

Lehi, UT ยท On-site

$52.25 - $70/hr

Architect for Reliability: Design and implement production-grade Kubernetes environments and GCP ... Prevented customer-impacting regressions attributable to platform design * Clear, documented cloud ...

New

Sr. Cloud Platform Engineer

Lehi, UT ยท On-site

$52.25 - $70/hr

Architect for Reliability: Design and implement production-grade Kubernetes environments and GCP ... Prevented customer-impacting regressions attributable to platform design * Clear, documented cloud ...

New

Showing results 21-40

Customer Reliability Engineer information

See Utah salary details

$55.5K

$107.4K

$128.4K

How much do customer reliability engineer jobs pay per year?

As of Aug 13, 2026, the average yearly pay for customer reliability engineer in Utah is $107,399.00, according to ZipRecruiter salary data. Most workers in this role earn between $93,300.00 and $117,400.00 per year, depending on experience, location, and employer.

How does a customer reliability engineer typically interact with clients and internal engineering teams?

Customer Reliability Engineers serve as a vital bridge between clients and internal technical teams. They regularly communicate with customers to understand their needs, troubleshoot issues, and provide technical guidance. Internally, they collaborate closely with product, support, and development teams to relay customer feedback, help prioritize reliability improvements, and ensure seamless incident resolution. This cross-functional role requires strong communication skills and the ability to translate technical information for different audiences, making every day varied and impactful.

What does a customer reliability engineer do?

A customer reliability engineer (CRE) works with clients to ensure the reliability, performance, and availability of products or services. They analyze system issues, develop solutions, and often collaborate with engineering teams to improve infrastructure and customer experience, typically using monitoring tools and technical expertise. CREs may also provide technical support and guidance to help clients optimize their use of the company's offerings.

What is the difference between Customer Reliability Engineer vs Site Reliability Engineer?

AspectCustomer Reliability EngineerSite Reliability Engineer
CredentialsTypically requires engineering degrees, certifications in cloud platforms (AWS, Azure), and knowledge of customer supportRequires engineering degrees, certifications in cloud and systems management, with a focus on infrastructure
Work EnvironmentCustomer-facing, involves direct interaction with clients to resolve issues and improve reliabilityPrimarily internal, focused on maintaining and improving system reliability and scalability
Employer & Industry UsageUsed by cloud service providers and tech companies with a customer support componentCommon in large tech companies managing large-scale infrastructure and services

The main difference is that Customer Reliability Engineers focus on ensuring customer satisfaction and resolving client-specific issues, while Site Reliability Engineers concentrate on internal system stability and scalability. Both roles require technical expertise and cloud knowledge but serve different operational needs.

What is a customer reliability engineer?

A Customer Reliability Engineer (CRE) is a technical professional who works closely with customers to ensure the reliability, performance, and uptime of software products and services. CREs act as a bridge between customers and engineering teams, helping to identify, troubleshoot, and resolve reliability issues. They often collaborate with multiple departments to implement best practices, monitor systems, and proactively address potential problems, ultimately aiming to improve the overall customer experience.

What skills and qualifications are needed to thrive as a customer reliability engineer?

To thrive as a Customer Reliability Engineer, you need a solid background in systems engineering, incident management, and troubleshooting, often supported by a degree in computer science or related field. Familiarity with cloud platforms (such as AWS or GCP), monitoring tools (like Datadog or Prometheus), and automation scripts is typically required. Exceptional communication, problem-solving abilities, and a customer-centric mindset are vital soft skills for this role. These skills ensure efficient incident resolution, strong client relationships, and reliable system performance under pressure.
What job categories do people searching Customer Reliability Engineer jobs in Utah look for? The top searched job categories for Customer Reliability Engineer jobs in Utah are:
What cities in Utah are hiring for Customer Reliability Engineer jobs? Cities in Utah with the most Customer Reliability Engineer job openings:

Manager, Site Reliability Engineering

1 O.C. Tanner Company

Salt Lake City, UT โ€ข On-site

$120 - $180/hr

Other

Posted 3 days ago

New


Job description

O.C. Tanner is the global leader in software and services that improve workplace culture through meaningful employee experiences. Our Culture Cloud is a suite of apps designed to enhance the employee experience with strategic recognition, service awards, wellbeing, leadership, and events that help people thrive at work. Our Culture by Design approach provides expert services to organizations looking to create great workplaces. Our global team of 1,500 people hail from 58 countries and speak 62 languages. As programmers, researchers, designers, client professionals and craftspeople we create the tech, tools and awards that connect employees to purpose at thousands of companies. Join us as we help people all over the world thrive at work.

Location: Salt Lake City, UT

As the Manager of Site Reliability Engineering, you will lead the strategy, execution, and evolution of reliability for our world-class employee recognition platform. You will build, mentor, and empower a team of Site Reliability Engineers while partnering closely with Engineering, Product, and Support organizations to deliver highly available, scalable, and resilient services that serve millions of users. We are seeking a leader who is passionate about operational excellence, continuous improvement, and fostering a reliability-first culture through automation, observability, and shared ownership. In this role, you will champion the development of self-healing platforms, drive incident and operational maturity, and enable engineering teams to innovate faster while delivering exceptional customer experiences.

Key Responsibilities
  • Lead, mentor, and develop a team of Site Reliability Engineers, fostering a culture of reliability, accountability, operational excellence, and continuous improvement.
  • Define and execute the organization's reliability strategy, improving availability, scalability, performance, and resilience through automation and engineering best practices.
  • Establish team priorities, goals, and success metrics aligned with business objectives, customer needs, and platform health.
  • Partner with Engineering, Product, and Support leaders to drive shared ownership of production services and embed reliability, observability, and operational excellence throughout the software development lifecycle.
  • Build and evolve observability capabilities using OpenTelemetry, Datadog, Coralogix, or similar tools, establishing enterprise standards for metrics, logs, traces, alerting, and Service Level Objectives (SLOs).
  • Oversee production triage, incident response, and escalation processes, ensuring timely service restoration, effective root cause analysis, and blameless post-incident reviews.
  • Champion a reliability-first engineering culture focused on automation, proactive risk reduction, operational readiness, shiftโ€‘left quality practices, and continuous improvement.
  • Collaborate with global engineering teams in a followโ€‘theโ€‘sun support model, ensuring seamless 24x7 coverage, effective operational handoffs, and consistent service ownership.
  • Own onโ€‘call programs, incident management practices, and operational health metrics, driving improvements in alert quality, operational efficiency, and toil reduction.
  • Manage team capacity, hiring, performance management, career development, budgeting, and workforce planning to ensure effective support of businessโ€‘critical services.
  • Provide regular reporting to engineering and executive leadership on reliability trends, incidents, risks, performance metrics, and strategic initiatives.
Required Qualifications
  • 5+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or related disciplines, including 2+ years in a technical leadership or people management role.
  • Proven experience leading teams responsible for production operations, reliability engineering, incident management, and operational excellence.
  • Experience designing and implementing SRE practices, reliability programs, or operational maturity initiatives within growing engineering organizations.
  • Experience operating large-scale, customerโ€‘facing SaaS platforms with high availability, performance, and scalability requirements.
  • Strong understanding of modern software engineering practices and partnering with development teams to build reliable, resilient systems.
  • Handsโ€‘on experience with observability platforms such as OpenTelemetry, Datadog, Coralogix, or similar technologies.
  • Strong knowledge of AWS and Kubernetes in production environments.
  • Deep understanding of monitoring, logging, distributed tracing, SLIs, SLOs, error budgets, and reliability engineering principles.
  • Demonstrated ability to lead crossโ€‘functional initiatives and influence stakeholders across Engineering, Product, and Support organizations.
  • Experience developing engineering roadmaps, defining team objectives, aligning reliability investments with business priorities, and driving continuous operational improvement through incident learning and postโ€‘incident reviews.
Preferred Qualifications
  • Experience leading distributed or globally dispersed engineering teams.
  • Experience with multiple cloud providers or cloudโ€‘agnostic platform architectures.
  • Familiarity with security, compliance, governance, and operational risk management frameworks.
  • Proficiency with modern Infrastructureโ€‘asโ€‘Code and technologies such as Terraform, Golang, Python, Playwright, and Performance Monitoring tools.
  • Experience with relational and distributed data technologies such as PostgreSQL, OpenSearch, Redis/ElastiCache, or Aurora.
  • Experience with messaging and streaming platforms such as Kafka, ActiveMQ, SNS/SQS, or similar eventโ€‘driven technologies.
  • Strong understanding of cost optimization, platform sustainability, and engineering efficiency metrics.

We create inspiring workplaces for some of the biggest and best companies in the world. And we do it within our own teams every day. Thatโ€™s one reason we made the Fortune 100 Best Companies to Work For list in 2021. Join us and watch people thrive at workโ€”including you. With seven global offices and employees working around the world, weโ€™re committed to creating an atmosphere where every person can share their talents and reach their potential.

#J-18808-Ljbffr