1

Observability Sre Jobs (NOW HIRING)

Define and own the enterprise monitoring and SRE observability strategy* Serve as the subject matter expert for Dynatrace, ServiceNow integration, and alerting architecture* Evaluate and recommend ...

We are seeking a highly skilled Site Reliability Engineer (SRE ) with strong observability expertise, proven communication skills, and the ability to drive reliability maturity across multi-team ...

Site Reliability Engineer

Charlotte, NC · On-site

$55.75 - $74/hr

Define and own the enterprise monitoring and SRE observability strategy * Serve as the subject matter expert for Dynatrace, ServiceNow integration, and alerting architecture * Evaluate and recommend ...

Site Reliability Engineer

Charlotte, NC

$55.75 - $74/hr

Define and own the enterprise monitoring and SRE observability strategy * Serve as the subject matter expert for Dynatrace, ServiceNow integration, and alerting architecture * Evaluate and recommend ...

SRE

Hartford, CT · On-site

$57.50 - $76.50/hr

The role involves utilizing various observability, monitoring, and logging tools to assist in building the MVP. Responsibilities : • Hands on experience on setting up SRE Platform • Defining SLI ...

SRE

Fremont, CA · On-site

$62.50 - $83.25/hr

The role involves working with various observability and monitoring tools to build a minimum viable product (MVP). Responsibilities : • Hands on experience on setting up SRE Platform, defining SLI ...

Site Reliability Engineer

Riverwoods, IL · On-site

$59.25 - $78.75/hr

We are looking for a Site Reliability Engineer to join our growing team in Riverwoods, IL United ... Expertise in observability tools including APM (Datadog), synthetic monitoring and log aggregation ...

You will be the first dedicated Site Reliability Engineer and own critical infrastructure end to ... Kubernetes, Docker, Terraform, Terragrunt, OpenTofu, CI/CD, GitOps, networking, and observability ...

Showing results 21-40

Observability Sre information

See salary details

$10

$63

$91

How much do observability sre jobs pay per hour?

As of Sep 2, 2026, the average hourly pay for observability sre in the United States is $63.74, according to ZipRecruiter salary data. Most workers in this role earn between $54.81 and $72.84 per hour, depending on experience, location, and employer.

What is an Observability SRE?

An Observability SRE (Site Reliability Engineer) is a specialist focused on ensuring that systems and applications are transparent, measurable, and reliable. Their main responsibility is to implement and maintain tools for monitoring, logging, and tracing, providing insights into system performance and health. Observability SREs help teams quickly detect, diagnose, and resolve issues by making system behavior visible and understandable. They play a critical role in uptime, incident response, and performance optimization, bridging the gap between software development and IT operations.

What are the key skills and qualifications needed to thrive as an Observability SRE?

To thrive as an Observability SRE, you need a solid background in systems engineering, monitoring best practices, and expertise in observability concepts, often supported by a degree in computer science or related fields. Familiarity with tools like Prometheus, Grafana, ELK stack, and cloud monitoring platforms, as well as scripting languages such as Python or Bash, is typically required. Strong problem-solving, collaboration, and communication skills help SREs respond to incidents and work across teams effectively. These skills ensure system reliability, rapid issue detection, and continuous service improvement in complex technical environments.

What are some typical challenges faced by Observability SREs when implementing monitoring solutions across diverse systems?

Observability SREs often encounter challenges when integrating monitoring tools across varied technology stacks and legacy systems. Ensuring consistent data collection, standardizing metrics, and maintaining visibility in complex, distributed environments can be difficult. Collaborating with development and operations teams to define meaningful alerts and dashboards requires strong communication and a deep understanding of both infrastructure and application behaviors. Staying up-to-date with evolving tools and best practices is also essential to address emerging observability needs.

What is the difference between Observability Sre vs Site Reliability Engineer?

AspectObservability SreSite Reliability Engineer
Primary FocusMonitoring, logging, and tracing to ensure system observabilitySystem reliability, automation, and infrastructure management
Skills & CertificationsMonitoring tools, scripting, cloud platforms, observability frameworksLinux, scripting, cloud services, automation tools
Work EnvironmentCollaborates with SRE, DevOps, and development teams on observability practicesBuilds and maintains scalable, reliable systems in production

While both roles focus on system stability, Observability Sre specializes in monitoring and diagnostics, whereas Site Reliability Engineers focus on overall system reliability and automation. They often work together to ensure robust, observable, and reliable systems.

More about Observability Sre jobs

What cities are hiring for Observability Sre jobs?

Cities with the most Observability Sre job openings:

What states have the most Observability Sre jobs?

States with the most job openings for Observability Sre jobs include:

Infographic showing various Observability Sre job openings in the United States as of August 2026, with employment types broken down into 96% Full Time, and 4% Contract. Highlights an 76% Physical, 7% Hybrid, and 17% Remote job distribution, with an average salary of $132,583 per year, or $63.7 per hour.

Site Reliability Engineer

CRC Group

Charlotte, NC • On-site

$120 - $190/hr

Other

Medical, Dental, Vision, Life, Retirement, PTO

Re-posted 18 days ago


Job description

Site Reliability EngineerSkip to main content# CRC CareersSite Reliability Engineer page is loaded## Site Reliability EngineerApplylocations: CRC - Charlotte, NC 600 S. Tryon St.time type: Full timeposted on: Posted Todayjob requisition id: R0000002909**The position is described below. If you want to apply, click the Apply button at the top or bottom of this page. You'll be required to create an account or sign in to an existing one.****If you have a disability and need assistance with the application, you can request a reasonable accommodation. Send an email to** Accessibility **(accommodation requests only; other inquiries won't receive a response).****Regular or Temporary:**Regular**Language Fluency:** English (Required)**Work Shift:**1st Shift (United States of America)**Please review the following job description:**Lead Site Reliability & Environment Monitoring Engineer (Azure / Dynatrace / ServiceNow) We are seeking a Lead Site Reliability & Environment Monitoring Engineer to establish and evolve our enterprise observability and monitoring strategy across cloud and application platforms. This is a full-time leadership role responsible for owning monitoring design, driving platform decisions, and guiding engineering teams toward modern SRE practices. This individual will act as the technical authority for monitoring and alerting, shaping how signals from Dynatrace flow into ServiceNow and enterprise messaging/paging platforms, and enabling a shift toward automated, intelligent, and self-healing operations.**Key Responsibilities****Strategic Leadership & Decision-Making*** Define and own the enterprise monitoring and SRE observability strategy* Serve as the subject matter expert for Dynatrace, ServiceNow integration, and alerting architecture* Evaluate and recommend tooling, integration patterns, and platform direction* Drive decisions on alerting philosophy, noise reduction, and signal quality improvement**Platform Ownership & Architecture*** Architect and standardize end-to-end monitoring and SRE pipelines: + Dynatrace → ServiceNow incident lifecycle + Alert correlation, deduplication, and prioritization + Integration with paging systems (PagerDuty, SMS, voice, Teams)* Establish best practices for: + Event ingestion and enrichment + Incident routing and automated assignment + Integration with CMDB and service mapping**Site Reliability Engineering (SRE) Leadership*** Lead adoption of SRE principles, including: + SLIs, SLOs, and error budgets + Reliability engineering practices across services + Proactive monitoring and resilience design* Champion a shift from reactive operations to proactive reliability engineering* Influence application and platform teams to build observable, resilient systems by design**Automation & Self-Healing Enablement*** Drive development of automated remediation and self-healing capabilities* Leverage Dynatrace workflows, Azure services, and automation frameworks to: + Reduce manual incident handling + Eliminate repeatable operational tasks + Minimize unnecessary paging**ServiceNow & Observability Integration Leadership*** Own integration between Dynatrace and ServiceNow ITSM/ITOM, including: + Incident, Event Management, and CMDB alignment + Service mapping and dependency visibility + Governance for application/service tagging* Define standards for: + Automated incident creation and resolution + Priority assignment and routing logic + Monitoring-to-ITSM data synchronization**Team Leadership & Cross-Functional Influence*** Provide technical leadership and mentorship across SRE, platform, and application teams* Act as a central point of coordination between engineering, cloud, and ITSM teams* Lead workshops and working sessions to: + Drive monitoring standardization + Align teams on reliability practices + Influence upstream architectural decisions**Operational Excellence*** Establish KPIs and drive improvement in: + Incident response and resolution times + Alert quality and paging effectiveness + Monitoring coverage across critical services* Provide leadership with clear visibility into service health and reliability trends**Required Qualifications*** 7+ years in Site Reliability Engineering, monitoring, or production engineering* Proven experience in a technical leadership or lead engineer role* Deep hands-on experience with: + Dynatrace (or equivalent observability platforms) + Microsoft Azure (IaaS, PaaS, networking, identity) + ServiceNow ITSM / ITOM (incident, event management, CMDB)* Demonstrated ability to: + Design and lead enterprise monitoring/SRE architectures + Drive platform and tooling decisions + Integrate observability, ITSM, and paging solutions**Preferred Qualifications*** Experience leading SRE or observability transformation initiatives* Strong expertise with Dynatrace–ServiceNow integrations* Experience modernizing or consolidating paging/on-call tooling* Familiarity with: + Azure-based SRE tooling or AI-assisted operations + Automation frameworks (GitHub Actions, Runbooks, etc.) + Infrastructure as Code (Terraform, ARM, Bicep)**Success Metrics*** Reduction in alert noise and unnecessary paging* Improved incident routing accuracy and MTTR* Increased adoption of self-healing and automated workflows* Strong alignment between monitoring, CMDB, and service ownership* Enterprise-wide adoption of SRE and monitoring standards**General Description of Available Benefits for Eligible Employees of CRC Group:** At CRC Group, we're committed to supporting every aspect of teammates' well-being – physical, emotional, financial, social, and professional. Our best-in-class benefits program is designed to care for the whole you, offering a wide range of coverage and support. Eligible full-time teammates enjoy access to medical, dental, vision, life, disability, and AD&D insurance; tax-advantaged savings accounts; and a 401(k) plan with company match. CRC Group also offers generous paid time off programs, including company holidays, vacation and sick days, new parent leave, and more. Eligible positions may also qualify for restricted stock units and/or a deferred compensation plan.***CRC Group supports a diverse workforce and is an Equal Opportunity Employer that does not discriminate against individuals on the basis of race, gender, color, religion, citizenship or national origin, age, sexual orientation, gender identity, disability, veteran status or other classification protected by law. CRC Group is a Drug Free Workplace.***EEO is the Law Pay Transparency Nondiscrimination Provision E-Verify #J-18808-Ljbffr