1

Sre Manager Jobs (NOW HIRING)

Site Reliability Engineer (SRE)

Plano, TX ยท On-site

$54.50 - $72.50/hr

Site Reliability Engineer (SRE) Location: Richmond, VA or Plano, TX Work Model: Hybrid - 3 days onsite per week Duration: Long term contract Job Summary: We are seeking an experienced Site ...

Site Reliability Engineer

Charlotte, NC ยท On-site

$55.75 - $74/hr

Integration with CMDB and service mapping Site Reliability Engineering (SRE) Leadership * Lead adoption of SRE principles, including: * SLIs, SLOs, and error budgets * Reliability engineering ...

$56.75 - $75.25/hr

Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity. * Develop proactive monitoring, alerting, logging, and ...

Showing results 21-40

Sre Manager information

See salary details

$62K

$117.5K

$168.5K

How much do sre manager jobs pay per year?

As of Aug 19, 2026, the average yearly pay for sre manager in the United States is $117,488.00, according to ZipRecruiter salary data. Most workers in this role earn between $94,500.00 and $140,000.00 per year, depending on experience, location, and employer.

What is an SRE manager?

SRE Managers are leaders responsible for overseeing Site Reliability Engineering (SRE) teams. They ensure the reliability, scalability, and performance of software systems by guiding engineers in implementing best practices, automation, and monitoring processes. SRE Managers collaborate closely with development and operations teams to balance feature development with system stability. Their role also includes mentoring SREs, managing incident response, and driving improvements in system reliability and operational efficiency.

What are some common challenges SRE managers face when leading Site Reliability Engineering teams?

SRE Managers often encounter challenges balancing reliability with rapid development, ensuring their teams have the right mix of software engineering and operations skills. They must also foster a culture of continuous improvement while managing on-call rotations and incident response without causing burnout. Additionally, collaborating effectively with development and product teams to set realistic service level objectives (SLOs) and drive adoption of SRE best practices can require strong communication and negotiation skills.

What are the key skills and qualifications needed to thrive as an SRE manager, and why are they important?

To thrive as an SRE Manager, you need a deep understanding of site reliability engineering principles, strong experience with systems architecture, and a background in computer science or a related field. Familiarity with tools such as Kubernetes, Prometheus, cloud platforms, and CI/CD pipelines, as well as certifications like AWS Certified Solutions Architect, are commonly expected. Leadership, effective communication, and problem-solving skills are crucial for driving team performance and collaborating across departments. These skills ensure high system reliability, efficient incident management, and a culture of continuous improvement within technical organizations.

What is the difference between Sre Manager vs DevOps Engineer?

AspectSre ManagerDevOps Engineer
CredentialsTypically requires a Bachelor's/Master's in CS or related field, with certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentLeads teams, manages incident response, and oversees reliability strategiesFocuses on automation, CI/CD pipelines, and infrastructure deployment
Industry UsageCommon in large tech companies, financial services, and cloud providersWidely used across startups, tech firms, and enterprises adopting DevOps practices

The Sre Manager and DevOps Engineer roles share overlapping skills in cloud computing, automation, and infrastructure management. While the Sre Manager oversees reliability and team coordination, the DevOps Engineer focuses on implementing automation tools and deployment pipelines. Both roles are crucial for modern IT operations, but the Sre Manager typically has a broader leadership responsibility, whereas the DevOps Engineer is more hands-on with technical implementation.

What does an SRE manager do?

An SRE manager oversees the reliability and performance of software systems, leading teams that implement automation, monitoring, and incident response processes. They coordinate efforts to ensure system availability, scalability, and efficiency, often using tools like monitoring dashboards and incident management platforms. Strong leadership, technical expertise, and understanding of service level objectives are essential in this role.
More about Sre Manager jobs

What cities are hiring for Sre Manager jobs?

Cities with the most Sre Manager job openings:

What are the most commonly searched types of Sre jobs?

The most popular types of Sre jobs are:

What states have the most Sre Manager jobs?

States with the most job openings for Sre Manager jobs include:

Infographic showing various Sre Manager job openings in the United States as of August 2026, with employment types broken down into 33% Full Time, and 67% Contract. Highlights an 67% In-person, and 33% Remote job distribution, with an average salary of $117,488 per year, or $56.5 per hour.

Site Reliability Engineer (SRE)

IT America Inc

Plano, TX โ€ข On-site

$54.50 - $72.50/hr

Contractor

Re-posted 16 days ago


Job description

Position: Site Reliability Engineer (SRE)

Location: Richmond, VA or Plano, TX

Work Model: Hybrid – 3 days onsite per week

Duration: Long term contract

Job Summary:

We are seeking an experienced Site Reliability Engineer (SRE) to support cloud-native platforms and production systems for a large enterprise environment. This role will focus on ensuring high availability, reliability, performance, and scalability of mission-critical applications running on AWS.

Strong Preference: Former Capital One engineers. Candidates must be able to provide verifiable Capital One credentials and be eligible for rehire.

Key Responsibilities:

  • Design, build, and maintain highly reliable, scalable, and resilient systems in AWS
  • Monitor system health, performance, and availability using SRE best practices
  • Implement automation to reduce manual operational work
  • Troubleshoot production incidents and perform root cause analysis (RCA)
  • Develop and maintain scripts and tools to improve system reliability and efficiency
  • Partner with application development, platform, and infrastructure teams
  • Support on-call rotations and incident response as required
  • Enforce operational excellence, security, and compliance standards

Required Skills & Qualifications:

  • Former Capital One experience – HIGHLY preferred
  • Must provide credentials for rehire eligibility verification
  • Strong hands-on experience with AWS (EC2, EKS, Lambda, CloudWatch, IAM, etc.)
  • Python scripting experience strongly preferred
  • Bash or Shell scripting experience will also be considered
  • Experience with Linux-based systems and troubleshooting
  • Understanding of SRE concepts: SLIs, SLOs, error budgets, monitoring, and alerting
  • Experience supporting production environments at scale

Preferred Qualifications:

  • Experience with CI/CD pipelines
  • Infrastructure as Code (Terraform, CloudFormation)
  • Containerization and orchestration (Docker, Kubernetes)
  • Observability tools (Prometheus, Grafana, Datadog, CloudWatch)
  • Experience working in highly regulated enterprise environments