1

Senior Observability Engineer Jobs in Orem, UT (NOW HIRING)

Sr. Observability Engineer

Lehi, UT · On-site

$98K - $134K/yr

As a Senior Observability Engineer, you build and operate an observability control plane. You scaffold baselines, score coverage, and turn every real incident into the detection the platform should ...

Observability & Incident Response: Implement robust observability, telemetry, and distributed ... Proven Senior Engineering Experience: 5+ years of professional backend software engineering ...

Senior Data Engineer

Lehi, UT · On-site

$99K - $135K/yr

ABOUT THIS POSITION Waystar is seeking a highly skilled Senior Data Engineer to join our Business ... Develop monitoring, alerting, and observability capabilities across the data platform * Investigate ...

Senior Data Engineer

Lehi, UT · On-site

$99K - $135K/yr

ABOUT THIS POSITION Waystar is seeking a highly skilled Senior Data Engineer to join our Business ... Develop monitoring, alerting, and observability capabilities across the data platform * Investigate ...

Senior DevOps Engineer I

American Fork, UT · On-site

$116K - $149K/yr

ABOUT THIS ROLE As a Senior DevOps Engineer at LVT, you will play a pivotal role in designing ... System Observability: Develop, implement, and maintain advanced observability, monitoring, and ...

Senior Application Security Engineer

Draper, UT · Remote

$107K - $146K/yr

Senior Application Security Engineer Canopy, South Jordan, UT About Us Canopy is a fast-growing ... Build out the security observability stack: audit logging, cloud posture monitoring, runtime threat ...

Senior Software Engineer - Backend

American Fork, UT · On-site

$109K - $144K/yr

Observability & Incident Response: Implement robust observability, telemetry, and distributed ... Proven Senior Engineering Experience: 5+ years of professional backend software engineering ...

Senior Software Engineer - Backend

American Fork, UT · On-site

$109K - $144K/yr

Observability & Incident Response: Implement robust observability, telemetry, and distributed ... Proven Senior Engineering Experience: 5+ years of professional backend software engineering ...

Sr. Core Backend Engineer

Lehi, UT · On-site

$171 - $181/hr

Sr. Core Backend Engineer Remote Role Overview The Core Backend Engineering team builds the ... Optimize system performance, reliability, scalability, and observability across production services.

Senior DevOps Engineer I

American Fork, UT · On-site

$116K - $149K/yr

ABOUT THIS ROLE As a Senior DevOps Engineer at LVT, you will play a pivotal role in designing ... System Observability: Develop, implement, and maintain advanced observability, monitoring, and ...

Senior DevOps Engineer I

American Fork, UT · On-site

$116K - $149K/yr

ABOUT THIS ROLE As a Senior DevOps Engineer at LVT, you will play a pivotal role in designing ... System Observability: Develop, implement, and maintain advanced observability, monitoring, and ...

Senior Engineer - Machine Learning

Midvale, UT · Hybrid

$98K - $135K/yr

As a Senior Engineer, Machine Learning at Berkadia, you'll be at the forefront of applying cutting ... Familiarity with cloud infrastructure and MLOps practices (deployment, monitoring, observability)

Senior Software Engineer

Draper, UT · On-site

$114K - $151K/yr

As a Senior Software Engineer , you will be a technical owner of key services supporting our ... Build for Resilience & Observability * Implement proven resilience patterns including circuit ...

Senior Engineer - Machine Learning

Midvale, UT · On-site

$98K - $135K/yr

As a Senior Engineer, Machine Learning at Berkadia, you'll be at the forefront of applying cutting ... Familiarity with cloud infrastructure and MLOps practices (deployment, monitoring, observability)

Senior Software Engineer - Next-Gen

Lehi, UT · On-site +1

$140K - $155K/yr

What You'll Do NetDocuments is seeking a Senior Full Stack Software Engineer to help build the next ... Implement logging, monitoring, and observability to improve production visibility * Optimize ...

next page

Showing results 1-20

Senior Observability Engineer information

See Orem, UT salary details

$51.7K

$110K

$159.5K

How much do senior observability engineer jobs pay per year?

As of Aug 23, 2026, the average yearly pay for senior observability engineer in Orem, UT is $110,025.00, according to ZipRecruiter salary data. Most workers in this role earn between $90,800.00 and $124,800.00 per year, depending on experience, location, and employer.

What is a senior observability engineer?

A Senior Observability Engineer is a seasoned IT professional responsible for designing, implementing, and maintaining systems that monitor and provide insights into the performance, health, and reliability of software applications and infrastructure. They utilize tools for logging, monitoring, tracing, and alerting to ensure that systems are observable and any issues can be quickly detected and resolved. In addition to technical expertise, they often collaborate with development and operations teams to establish best practices, improve incident response, and optimize system performance. Their work is crucial for maintaining uptime, enhancing customer experiences, and supporting the scalability of technology platforms.

How does a senior observability engineer typically collaborate with development and operations teams?

A Senior Observability Engineer works closely with both development and operations teams to ensure robust monitoring, logging, and tracing solutions are in place across all applications and infrastructure. They often participate in architecture discussions to advise on best practices for instrumenting code and systems for observability. By analyzing metrics and alerting patterns, they help teams proactively resolve issues and optimize system performance. This role also involves mentoring engineers on observability tools and fostering a culture of transparency and accountability in incident response.

What are the key skills and qualifications needed to thrive as a senior observability engineer, and why are they important?

To thrive as a Senior Observability Engineer, you need expertise in monitoring, logging, and tracing systems, with a solid background in computer science or a related field. Familiarity with tools like Prometheus, Grafana, ELK stack, and cloud platforms, as well as certifications such as AWS Certified DevOps Engineer, are typically required. Strong problem-solving, collaboration, and communication skills are critical for effectively diagnosing and resolving complex infrastructure issues. These skills ensure reliable system performance, rapid incident response, and continuous improvement of the technology environment.

What is the difference between Senior Observability Engineer vs Site Reliability Engineer?

AspectSenior Observability EngineerSite Reliability Engineer
CredentialsExperience with monitoring tools, scripting, cloud platformsSame as Senior Observability Engineer, often with SRE certifications
Work EnvironmentFocus on monitoring, logging, and tracing systemsFocus on system reliability, automation, and incident response
Industry UsageUsed in tech companies emphasizing system observabilityCommon in large-scale tech and cloud services
Search/Comparison IntentOften compared for monitoring rolesCompared for reliability and system stability roles

While both roles require expertise in cloud platforms and scripting, the Senior Observability Engineer primarily focuses on designing and maintaining monitoring, logging, and tracing systems to ensure system visibility. In contrast, a Site Reliability Engineer emphasizes system reliability, automation, and incident management to maintain service uptime. Both roles are vital in tech environments but serve different core functions related to system health and stability.

How much do senior observability engineers make?

Senior observability engineers typically earn between $110,000 and $160,000 annually, depending on experience, location, and company size. They often work with tools like Prometheus, Grafana, and cloud platforms, and may require advanced knowledge of monitoring, logging, and alerting systems.

What does a senior observability engineer do?

A senior observability engineer designs, implements, and maintains systems to monitor the performance and health of software applications and infrastructure. They utilize tools like Prometheus, Grafana, and ELK stack to analyze metrics, logs, and traces, ensuring system reliability and performance. This role often requires strong scripting skills and knowledge of cloud environments and distributed systems.

What are the most commonly searched types of Observability Engineer jobs in Orem, UT?

The most popular types of Observability Engineer jobs in Orem, UT are:

What are popular job titles related to Senior Observability Engineer jobs in Orem, UT?

For Senior Observability Engineer jobs in Orem, UT, the most frequently searched job titles are:

What job categories do people searching Senior Observability Engineer jobs in Orem, UT look for?

The top searched job categories for Senior Observability Engineer jobs in Orem, UT are:

Infographic showing various Senior Observability Engineer job openings in Orem, UT as of August 2026, with employment types broken down into 92% Full Time, 2% Part Time, and 6% Contract. Highlights an 87% Physical, 4% Hybrid, and 9% Remote job distribution, with an average salary of $110,025 per year, or $52.9 per hour.

Full-time

Posted 22 days ago


Job description

At MX, reliability is a product. Our infrastructure powers financial applications used by millions of people and processes billions of transactions for major financial institutions, and customers feel every second of downtime.

We're building a new observability function that runs the way we run incident response: the system does the heavy lifting, and people handle judgment, customers, and the exceptions. As a Senior Observability Engineer, you build and operate an observability control plane. You scaffold baselines, score coverage, and turn every real incident into the detection the platform should have caught. This is a multiplier role: you raise the bar for every team through standards and automation instead of building each team's dashboards by hand.

We call it the shepherd model. You shepherd Datadog and partner with our product engineering teams so they observe the right signals for their products. Service owners get real signal instead of noise, and leadership gets coverage and health as a program metric.

This role shares the team pager. Observability and incident response run one on-call roster. You take shifts with the rest of the team and act as Incident Commander when an incident needs one. It is core to the role, not an afterthought.

Engineering at MX runs hybrid infrastructure (AWS and bare metal) with services in Ruby, Go, and Java, messaging over NATS and RabbitMQ, and data on PostgreSQL and Redis. Datadog is our observability platform and incident.io is our incident response platform.

What you'll do:
  • Build and operate an observability control plane: automate baseline monitors, dashboards, and tagging standards through the Datadog API and Terraform.

  • After significant incidents, produce detection and dashboard gap packs grounded in Datadog and MX investigation patterns, with queries ready to apply.

  • Define what "good" looks like for a Ruby, Go, or Java service on Datadog (tags, golden signals, alert quality, dashboard contracts), then audit services against that standard and accept or reject readiness.

  • Validate, don't own. Service owners keep their alerts and dashboards; you confirm they are complete and correct, then move on. Escalate to engineering managers when coverage fails or an owner is missing.

  • Own the monthly observability and service-catalog health report: departed owners, stale dashboards, services with no monitors, SLO gaps, and coverage trends.

  • Run maturity assessments (baseline through SLO, launch-ready, self-serve) and track them over time.

  • Tune alerting toward zero false SEV1/2 pages and actionable SEV3/4 alerts, and coach teams on Datadog cost and cardinality.

  • Build self-serve onboarding so new services get baseline observability on day one, without a multi-week embed.

  • Share the team pager. Rotate on the shared IR & Observability on-call, triage and investigate live incidents with Datadog and MX investigation patterns, and take Incident Commander or supporting technical roles as the incident needs.

  • After incidents, close the detection loop (gap packs, new monitors, dashboards) so the pager gets quieter over time.

  • Run high-value launch and production-readiness reviews as a checkpoint, not a permanent staffing model.

Basic Requirements
  • BS in Computer Science or equivalent experience

  • 5+ years running production observability, SRE, or DevOps

  • 5+ years automation-first engineering in Python, Bash, Go, and/or Terraform, plus Kubernetes proficiency

  • AI- and workflow-literate. You've used or built scripted and AI-assisted workflows to scale reviews, audits, and docs

  • Distributed-systems debugging across microservices: latency, connection pools, queues, and cascading failure on Kubernetes and bare metal, with NATS, RabbitMQ, Postgres, and Redis

  • Shared on-call, Incident Commander-capable

Preferred Requirements
  • Fintech experience with MX-like architectures

  • Datadog preferred; strong Grafana/Prometheus, Splunk, or New Relic experience counts if you can ramp on Datadog fast

  • Google SRE practices: toil elimination, incident management, automation for self-healing

  • Cross-functional influence without authority. You've improved teams that don't report to you

  • Governance and reporting: you can produce a monthly health and compliance report leadership reads (orphans, stale entries, gaps, trends)

  • OpenTelemetry instrumentation

  • Incident response platforms (incident.io, PagerDuty, OpsGenie); prior formal Incident Commander experience

  • Golang and Ruby on Rails (the MX stack)