1

Site Reliability Manager Jobs (NOW HIRING)

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineering (SRE) Manager

Frisco, TX ยท On-site

$53.50 - $71/hr

Role Summary It's an exciting time to join McAfee! We're looking for an experienced SRE Manager to lead our growing North American Site Reliability Engineering team and own reliability strategy ...

New

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer

Charlotte, NC ยท On-site

$55.75 - $74/hr

Establish best practices for: + Event ingestion and enrichment + Incident routing and automated assignment + Integration with CMDB and service mapping**Site Reliability Engineering (SRE) Leadership*

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Omaha, NE ยท On-site

$54.50 - $72.50/hr

Site Reliability Engineer (SRE) Location: Omaha, NE / Dallas, TX Job Type: Full Time Job Summary ... Highly skilled in managing production failures, conducting root cause analysis, and driving ...

Site Reliability Engineer

Charlotte, NC ยท On-site

$55.75 - $74/hr

Integration with CMDB and service mapping Site Reliability Engineering (SRE) Leadership * Lead adoption of SRE principles, including: * SLIs, SLOs, and error budgets * Reliability engineering ...

Showing results 21-40

Site Reliability Manager information

See salary details

$62K

$117.5K

$168.5K

How much do site reliability manager jobs pay per year?

As of Sep 15, 2026, the average yearly pay for site reliability manager in the United States is $117,488.00, according to ZipRecruiter salary data. Most workers in this role earn between $94,500.00 and $140,000.00 per year, depending on experience, location, and employer.

What is the difference between Site Reliability Manager vs DevOps Engineer?

AspectSite Reliability ManagerDevOps Engineer
CredentialsTypically requires a Bachelor's in Computer Science, certifications like SRE or Cloud certificationsOften holds a Bachelor's in Computer Science or related field, with certifications in cloud platforms or automation tools
Work EnvironmentLeads teams managing large-scale systems, focusing on reliability and uptimeWorks on automation, CI/CD pipelines, and infrastructure deployment
Industry UsageCommon in tech, cloud services, and large enterprisesWidely used in startups, tech companies, and organizations adopting DevOps practices

The main difference is that Site Reliability Managers focus on ensuring system reliability and managing SRE teams, while DevOps Engineers concentrate on automation, deployment, and continuous integration. Both roles require technical expertise but serve different strategic objectives within IT operations.

What is the role of a site reliability manager?

A site reliability manager (SRM) is responsible for ensuring the reliability, availability, and performance of a company's IT systems and services. They often oversee incident response, implement automation tools, and collaborate with development teams to improve system stability and scalability, typically requiring knowledge of monitoring tools and scripting skills.

What cities are hiring for Site Reliability Manager jobs?

Cities with the most Site Reliability Manager job openings:

What states have the most Site Reliability Manager jobs?

States with the most job openings for Site Reliability Manager jobs include:

What are popular job titles related to Site Reliability Manager jobs?

For Site Reliability Manager jobs, the most frequently searched job titles are:

Infographic showing various Site Reliability Manager job openings in the United States as of August 2026, with employment types broken down into 88% Full Time, 11% Part Time, and 1% Contract. Highlights an 80% Physical, 2% Hybrid, and 18% Remote job distribution, with an average salary of $117,488 per year, or $56.5 per hour.

Site Reliability Engineering (SRE) Manager

Frisco, TX โ€ข On-site

$53.50 - $71/hr

Other

Medical, Dental, Vision, Life, Retirement, PTO

Posted 3 days ago

New


Job description

Role Summary

It's an exciting time to join McAfee!

We're looking for an experienced SRE Manager to lead our growing North American Site Reliability Engineering team and own reliability strategy across our Cloud and Kubernetes platforms. You'll combine deep technical credibility with strong people leadership โ€” setting direction, growing the team, and acting as the senior escalation point during the most critical incidents โ€” while partnering closely with engineering and business leadership.

This is an onsite position located in our Frisco, TX office. We are only considering candidates within a commutable distance to the Frisco office.

Position DetailsAbout the role:
  • Lead and grow a team of SREs, setting technical direction and reliability strategy across AWS, GCP infrastructure and EKS platforms, and GKE platforms.
  • Own the organization's Incident and Problem Management processes, ensuring major incidents are handled efficiently, with timely executive communication and thorough post-incident reviews.
  • Drive the team's automation strategy, championing Python-based tooling and frameworks that reduce manual toil and improve reliability at scale.
  • Set standards for Terraform-based infrastructure-as-code, ensuring secure, scalable, and consistent provisioning practices across teams.
  • Define the organization's observability strategy, ensuring Grafana dashboards, alerting, and SQL/CloudWatch-based analysis practices scale effectively.
  • Act as a senior escalation point and incident commander for the most critical, high-severity incidents, providing calm, decisive leadership under pressure.
  • Partner with senior leadership, product, and engineering stakeholders to communicate risk, reliability posture, and remediation roadmaps clearly and confidently.
  • Own hiring, mentoring, performance management, and career development for the SRE team.
  • Design and build self-healing automation and runbooks that detect known failure patterns and trigger remediation automatically, reducing manual intervention and recovery time for recurring incidents.
  • Implement and maintain monitoring across multiple regions to ensure consistent visibility into system health, latency, and failover readiness across all deployment zones.
  • Proactively identify potential failure points and performance bottlenecks before they impact production and reduce operational workload by automating recurring manual tasks.
  • Drive ITSM process maturity across the organization, partnering with other teams to embed Incident and Problem Management best practices.
  • Manage on-call structure, escalation paths, staffing, and operational readiness for the team.
  • Report on reliability metrics, incident trends, and improvement initiatives to senior leadership.
About You:
  • 9+ years of experience in Site Reliability Engineering, DevOps, Infrastructure, or related roles, including significant experience in a leadership or management capacity.
  • Demonstrated experience building, leading, and growing high-performing technical teams.
  • Strategic reliability leadership: Balances operational excellence with long-term reliability improvements, ensuring the team addresses immediate risks while building scalable, sustainable practices.
  • Deep, hands-on background with AWS infrastructure and strong technical credibility to guide architecture and operational decisions.
  • Proven track record leading teams through complex EKS/GKE troubleshooting and operational challenges.
  • Strong technical fluency in Python for automation and Terraform for infrastructure-as-code, with the ability to guide and review the team's work.
  • Extensive experience owning ITSM processes โ€” Incident and Problem Management โ€” at an organizational level.
  • Strong command of observability practices, including Grafana dashboards, SQL, CloudWatch Logs Insights, and alerting strategy.
  • Outstanding communication skills โ€” able to clearly articulate technical risk, incident impact, and strategy to executive leadership and cross-functional stakeholders.
  • Strong stakeholder management skills, comfortable operating at the intersection of engineering, product, and business leadership.
  • AWS Certification - required (e.g., AWS Certified Solutions Architect - Professional, AWS Certified DevOps Engineer - Professional, or equivalent).
  • Certified Kubernetes Administrator (CKA) or equivalent EKS/Kubernetes certification - required.
  • Incident leadership: Provides calm, decisive direction during high-severity and high-visibility incidents, helping teams stay focused and coordinated under pressure.
  • Stakeholder trust: Builds strong working relationships with cross-functional teams, engineering leaders, product partners, and executive stakeholders.
  • Degrees are a plus, including Bachelor's degree in Computer Science, Information Technology, or a related field and/or Master's degree or MBA

#LI-Onsite

Company Overview

McAfee is a leader in personal security for consumers. Focused on protecting people, not just devices, McAfee consumer solutions adapt to usersโ€™ needs in an always online world, empowering them to live securely through integrated, intuitive solutions that protects their families and communities with the right security at the right moment.

Company Benefits and Perks
  • Bonus Program
  • 401k Retirement
  • Medical, Dental, Vision, Basic Life, Short Term Disability and Long-Term Disability Coverage
  • Paid Parental Leave
  • Support and Community Involvement
  • 14 Paid Company Holidays
  • Unlimited Paid Time Off for Exempt Employees
  • 96 Hours of Sick Time and 120 Hours of Vacation for Non-Exempt Employees Accrued Each Year

We're serious about our commitment to diversity which is why McAfee prohibits discrimination based on race, color, religion, gender, national origin, age, disability, veteran status, marital status, pregnancy, gender expression or identity, sexual orientation or any other legally protected status.

Pay Range

The anticipated compensation for this position is USD $123,650.00/Yr. - USD $229,650.00/Yr. depending on experience and qualifications.

Please click here to view and download the Job Applicant Privacy Notice, which applies to all McAfee job applicants who are residents of the state of California.

#J-18808-Ljbffr