1

Site Reliability Engineer Manager Jobs in Washington

SRE ENGINEER/ MANAGER

Reston, VA · On-site

$59.25 - $78.75/hr

Job Summary (Sr. Manager SRE): - Design, implement, and manage scalable, secure, and fault-tolerant cloud infrastructure using AWS, Azure, or GCP. - Automate infrastructure provisioning and ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

... manage Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets for ... Site reliability engineering, monitoring, automation, incident response, performance optimization ...

Everforth ECS is seeking a Senior Site Reliability Engineer to work remotely . Everforth ECS is seeking talented professionals to join our successful and growing team in building the next-generation ...

SRE Engineer

Washington, DC · On-site

$64.50 - $85.75/hr

As a core member of the reliability team, you will apply formal SRE principles--such as defining SLIs/SLOs and managing error budgets--to optimize capacity and resiliency, ensure strict security ...

The SRE will work closely with software developers and operations teams to improve system ... Manage cloud and database system maintenance, debugging production issues as they arise. * Improve ...

Everforth ECS is seeking a Senior Site Reliability Engineer to work remotely . Everforth ECS is seeking talented professionals to join our successful and growing team in building the next-generation ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

... manage Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets for ... Site reliability engineering, monitoring, automation, incident response, performance optimization ...

Site Reliability Engineer (SRE)

Vienna, VA · On-site

$57.25 - $76/hr

Up to 2 years in duration MUST HAVES: • Minimum of 8 years of experience as a Site Reliability Engineerwith a strong understanding of SRE principles for highly scalable and reliable systems • ...

Strong understanding of JVM fundamentals (heap/memory management, garbage collection, OOM issues, thread analysis) * Proven experience with SRE practices, including: * Incident response and on-call ...

Site Reliability Engineer - Hybrid

Reston, VA · On-site

$59.25 - $78.75/hr

Second round would be an in-person interview Manager's call notes * This is an SRE role. SRE is under a shared services team within Fannie Mae who works with different application teams. So, multi ...

Showing results 21-40

Site Reliability Engineer Manager information

See Washington salary details

$12

$72

$104

How much do site reliability engineer manager jobs pay per hour?

As of Aug 17, 2026, the average hourly pay for site reliability engineer manager in Washington is $72.19, according to ZipRecruiter salary data. Most workers in this role earn between $62.07 and $82.50 per hour, depending on experience, location, and employer.

What is a site reliability engineer manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.

How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How much do site reliability engineer managers get paid?

Site Reliability Engineer Managers typically earn between $120,000 and $180,000 annually, depending on experience, location, and company size. They often oversee teams responsible for system reliability, incident response, and infrastructure automation, requiring strong leadership and technical skills.

Is a Site Reliability Engineer Manager a stressful job?

A Site Reliability Engineer Manager role can be stressful due to the responsibility of maintaining system uptime, managing incident responses, and ensuring reliability across complex infrastructure. The job often involves working under pressure, handling outages, and coordinating teams, but it also offers opportunities for problem-solving and leadership. Stress levels vary depending on company size, team structure, and workload management skills.

What are the most commonly searched types of Site Reliability Engineer jobs in Washington?

The most popular types of Site Reliability Engineer jobs in Washington are:

What cities in Washington are hiring for Site Reliability Engineer Manager jobs?

Cities in Washington with the most Site Reliability Engineer Manager job openings:

Infographic showing various Site Reliability Engineer Manager job openings in Washington as of August 2026, with employment types broken down into 84% Full Time, 9% Part Time, 2% Temporary, and 5% Contract. Highlights an 77% Physical, 2% Hybrid, and 21% Remote job distribution, with an average salary of $150,163 per year, or $72.2 per hour.

SRE ENGINEER/ MANAGER

Accenuate IT Solutions

Reston, VA • On-site

$59.25 - $78.75/hr

Contractor

Re-posted yesterday


Job description

Job Summary (Sr. Manager SRE):
- Design, implement, and manage scalable, secure, and fault-tolerant cloud infrastructure using AWS, Azure, or GCP.
- Automate infrastructure provisioning and configuration with IaC tools such as Terraform, CloudFormation, and Ansible.
- Develop automation for anomaly detection, recovery, toil reduction, self-healing, and cloud cost optimization.
- Lead implementation of DevSecOps best practices and maintain secure CI/CD pipelines (e.g., GitLab, Jenkins, Docker).
- Enforce security through IAM, RBAC, vulnerability remediation, and security scanning tools (SAST/DAST/SCA).
- Architect and manage microservices, serverless solutions, and APIs with a focus on fault tolerance and resilience.
- Implement and maintain monitoring, logging, and observability using CloudWatch, Splunk/SignalFX, Dynatrace, and OpenTelemetry.
- Drive incident management, lead root cause analysis, postmortems, and minimize MTTR/MTTD.
- Define and monitor system reliability metrics (SLOs, SLIs, error budgets).
- Conduct chaos engineering, resiliency assessments, and implement self-healing architectures.
- Manage and optimize databases (PostgreSQL, MongoDB, DynamoDB, Oracle, Redshift) and provide production support.
- Participate in on-call rotations and support incident/problem management.
- Collaborate with development, QA, and operations teams to implement shift-left testing practices (BDD, TDD, Unit, Regression).
- Maintain architecture diagrams, knowledge documentation, and disaster recovery plans.
- Communicate effectively with stakeholders and demonstrate strong relationship management across teams.