1

Site Reliability Engineer Manager Jobs (NOW HIRING)

Site Reliability Engineer

Charlotte, NC · On-site

$55.75 - $74/hr

Integration with CMDB and service mapping Site Reliability Engineering (SRE) Leadership * Lead adoption of SRE principles, including: * SLIs, SLOs, and error budgets * Reliability engineering ...

NY

$56.75 - $75.25/hr

Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity. * Develop proactive monitoring, alerting, logging, and ...

$59.75 - $79.25/hr

Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity. * Develop proactive monitoring, alerting, logging, and ...

Site Reliability Engineer (SRE)

Plano, TX · On-site

$54.50 - $72.50/hr

Site Reliability Engineer (SRE) Location: Richmond, VA or Plano, TX Work Model: Hybrid - 3 days onsite per week Duration: Long term contract Job Summary: We are seeking an experienced Site ...

Site Reliability Engineer (SRE)Skip to main contentThis website stores cookies on your computer ... Ability to manage work via ticketing system and source control (for example: Jira / DevOps / TFS ...

Showing results 21-40

Site Reliability Engineer Manager information

See salary details

$10

$63

$91

How much do site reliability engineer manager jobs pay per hour?

As of Aug 30, 2026, the average hourly pay for site reliability engineer manager in the United States is $63.74, according to ZipRecruiter salary data. Most workers in this role earn between $54.81 and $72.84 per hour, depending on experience, location, and employer.

What is a site reliability engineer manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.

How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How much do site reliability engineer managers get paid?

Site Reliability Engineer Managers typically earn between $120,000 and $180,000 annually, depending on experience, location, and company size. They often oversee teams responsible for system reliability, incident response, and infrastructure automation, requiring strong leadership and technical skills.

Is a Site Reliability Engineer Manager a stressful job?

A Site Reliability Engineer Manager role can be stressful due to the responsibility of maintaining system uptime, managing incident responses, and ensuring reliability across complex infrastructure. The job often involves working under pressure, handling outages, and coordinating teams, but it also offers opportunities for problem-solving and leadership. Stress levels vary depending on company size, team structure, and workload management skills.

What cities are hiring for Site Reliability Engineer Manager jobs?

Cities with the most Site Reliability Engineer Manager job openings:

What are the most commonly searched types of Site Reliability Engineer jobs?

The most popular types of Site Reliability Engineer jobs are:

What states have the most Site Reliability Engineer Manager jobs?

States with the most job openings for Site Reliability Engineer Manager jobs include:

Infographic showing various Site Reliability Engineer Manager job openings in the United States as of August 2026, with employment types broken down into 88% Full Time, 11% Part Time, and 1% Contract. Highlights an 80% Physical, 2% Hybrid, and 18% Remote job distribution, with an average salary of $132,583 per year, or $63.7 per hour.

Site Reliability Engineer

Charlotte, NC • On-site

CRC Group
11 - 50 employees

$55.75 - $74/hr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Re-posted 17 days ago


Job description

The position is described below. If you want to apply, click the Apply button at the top or bottom of this page. You'll be required to create an account or sign in to an existing one.
If you have a disability and need assistance with the application, you can request a reasonable accommodation. Send an email to Accessibility (accommodation requests only; other inquiries won't receive a response).
Regular or Temporary:
Regular
Language Fluency: English (Required)
Work Shift:
1st Shift (United States of America)
Please review the following job description:
Lead Site Reliability & Environment Monitoring Engineer (Azure / Dynatrace / ServiceNow)
We are seeking a Lead Site Reliability & Environment Monitoring Engineer to establish and evolve our enterprise observability and monitoring strategy across cloud and application platforms. This is a full-time leadership role responsible for owning monitoring design, driving platform decisions, and guiding engineering teams toward modern SRE practices.
This individual will act as the technical authority for monitoring and alerting, shaping how signals from Dynatrace flow into ServiceNow and enterprise messaging/paging platforms, and enabling a shift toward automated, intelligent, and self-healing operations.
Key Responsibilities
Strategic Leadership & Decision-Making
  • Define and own the enterprise monitoring and SRE observability strategy
  • Serve as the subject matter expert for Dynatrace, ServiceNow integration, and alerting architecture
  • Evaluate and recommend tooling, integration patterns, and platform direction
  • Drive decisions on alerting philosophy, noise reduction, and signal quality improvement

Platform Ownership & Architecture
  • Architect and standardize end-to-end monitoring and SRE pipelines:
    • Dynatrace → ServiceNow incident lifecycle
    • Alert correlation, deduplication, and prioritization
    • Integration with paging systems (PagerDuty, SMS, voice, Teams)
  • Establish best practices for:
    • Event ingestion and enrichment
    • Incident routing and automated assignment
    • Integration with CMDB and service mapping

Site Reliability Engineering (SRE) Leadership
  • Lead adoption of SRE principles, including:
    • SLIs, SLOs, and error budgets
    • Reliability engineering practices across services
    • Proactive monitoring and resilience design
  • Champion a shift from reactive operations to proactive reliability engineering
  • Influence application and platform teams to build observable, resilient systems by design

Automation & Self-Healing Enablement
  • Drive development of automated remediation and self-healing capabilities
  • Leverage Dynatrace workflows, Azure services, and automation frameworks to:
    • Reduce manual incident handling
    • Eliminate repeatable operational tasks
    • Minimize unnecessary paging

ServiceNow & Observability Integration Leadership
  • Own integration between Dynatrace and ServiceNow ITSM/ITOM, including:
    • Incident, Event Management, and CMDB alignment
    • Service mapping and dependency visibility
    • Governance for application/service tagging
  • Define standards for:
    • Automated incident creation and resolution
    • Priority assignment and routing logic
    • Monitoring-to-ITSM data synchronization

Team Leadership & Cross-Functional Influence
  • Provide technical leadership and mentorship across SRE, platform, and application teams
  • Act as a central point of coordination between engineering, cloud, and ITSM teams
  • Lead workshops and working sessions to:
    • Drive monitoring standardization
    • Align teams on reliability practices
    • Influence upstream architectural decisions

Operational Excellence
  • Establish KPIs and drive improvement in:
    • Incident response and resolution times
    • Alert quality and paging effectiveness
    • Monitoring coverage across critical services
  • Provide leadership with clear visibility into service health and reliability trends

Required Qualifications
  • 7+ years in Site Reliability Engineering, monitoring, or production engineering
  • Proven experience in a technical leadership or lead engineer role
  • Deep hands-on experience with:
    • Dynatrace (or equivalent observability platforms)
    • Microsoft Azure (IaaS, PaaS, networking, identity)
    • ServiceNow ITSM / ITOM (incident, event management, CMDB)
  • Demonstrated ability to:
    • Design and lead enterprise monitoring/SRE architectures
    • Drive platform and tooling decisions
    • Integrate observability, ITSM, and paging solutions

Preferred Qualifications
  • Experience leading SRE or observability transformation initiatives
  • Strong expertise with Dynatrace-ServiceNow integrations
  • Experience modernizing or consolidating paging/on-call tooling
  • Familiarity with:
    • Azure-based SRE tooling or AI-assisted operations
    • Automation frameworks (GitHub Actions, Runbooks, etc.)
    • Infrastructure as Code (Terraform, ARM, Bicep)

Success Metrics
  • Reduction in alert noise and unnecessary paging
  • Improved incident routing accuracy and MTTR
  • Increased adoption of self-healing and automated workflows
  • Strong alignment between monitoring, CMDB, and service ownership
  • Enterprise-wide adoption of SRE and monitoring standards

General Description of Available Benefits for Eligible Employees of CRC Group: At CRC Group, we're committed to supporting every aspect of teammates' well-being - physical, emotional, financial, social, and professional. Our best-in-class benefits program is designed to care for the whole you, offering a wide range of coverage and support. Eligible full-time teammates enjoy access to medical, dental, vision, life, disability, and AD&D insurance; tax-advantaged savings accounts; and a 401(k) plan with company match. CRC Group also offers generous paid time off programs, including company holidays, vacation and sick days, new parent leave, and more. Eligible positions may also qualify for restricted stock units and/or a deferred compensation plan.
CRC Group supports a diverse workforce and is an Equal Opportunity Employer that does not discriminate against individuals on the basis of race, gender, color, religion, citizenship or national origin, age, sexual orientation, gender identity, disability, veteran status or other classification protected by law. CRC Group is a Drug Free Workplace.
EEO is the Law Pay Transparency Nondiscrimination Provision E-Verify