1

Director Site Reliability Engineering Jobs in Washington

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

Responsibilities : • Define, implement, and maintain site reliability engineering practices for mission-critical applications and shared services, with emphasis on uptime, resiliency ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

Responsibilities : • Define, implement, and maintain site reliability engineering practices for mission-critical applications and shared services, with emphasis on uptime, resiliency ...

About the Role ServiceNow is seeking a Director of SREReliability Engineeringto lead a strategic ... Embedded within the Site Reliability & Database Engineering organization, this leader willdrive ...

Site Reliability Engineer (SRE)

Washington, DC

$64.50 - $85.75/hr

Bachelor's degree in Computer Science, Engineering, or a related technical field. * 3+ years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, or infrastructure-focused roles.

New

Bangalore (Ecospace) Candescent Site Reliability Engineering (SRE) mission is to proactively ensure the reliability, availability and performance of our Digital First banking applications. As a ...

... engineering and systems administration practices to ensure the reliability, availability, and ... The SRE will help build resilient systems that scale, automate manual processes, manage fleetwide ...

SRE Engineer

Washington, DC · On-site

$64.50 - $85.75/hr

Reliability Engineering: Champion SRE metrics including Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets; design and execute resiliency test plans and support ...

next page

Showing results 1-20

Director Site Reliability Engineering information

How much do director site reliability engineering get paid?

Director of Site Reliability Engineering typically earns a salary ranging from $150,000 to $250,000 annually, depending on experience, company size, and location. They often oversee teams using tools like Kubernetes and Prometheus and require strong leadership and technical skills.

What is a director site reliability engineering?

A Director of Site Reliability Engineering (SRE) leads teams responsible for ensuring the availability, performance, and scalability of software systems. They define reliability best practices, drive automation, and collaborate with engineering and product teams to improve system resilience. This role requires strong leadership, technical expertise, and a focus on balancing innovation with operational stability.

What are the main challenges faced by a director site reliability engineering, and how can I prepare for them?

A Director of Site Reliability Engineering often encounters challenges such as balancing rapid feature delivery with system stability, managing complex incident responses, and fostering a culture of continuous improvement. Additionally, aligning reliability goals with business objectives and securing cross-functional buy-in can be demanding. To prepare, it is helpful to gain experience in high-scale system management, develop strong leadership and communication abilities, and cultivate a proactive approach to risk management and automation. Staying up to date with the latest SRE practices and building relationships with both engineering and business teams will also support your success in this pivotal role.

What are the key skills and qualifications needed to thrive as a director site reliability engineering?

To thrive as a Director Site Reliability Engineering, you need extensive experience in software engineering, infrastructure management, incident response, and people leadership, often supported by a degree in computer science or a related field. Familiarity with cloud platforms (such as AWS, GCP, or Azure), automation tools (Terraform, Ansible), monitoring systems (Prometheus, Datadog), and relevant certifications like CKA or AWS Solutions Architect is valued. Outstanding communication, stakeholder management, and strategic vision are key soft skills that set leaders apart in this role. These abilities ensure the reliability, scalability, and efficiency of critical systems while effectively guiding and motivating technical teams.

What job categories do people searching Director Site Reliability Engineering jobs in Washington look for? The top searched job categories for Director Site Reliability Engineering jobs in Washington are:
What cities in Washington are hiring for Director Site Reliability Engineering jobs? Cities in Washington with the most Director Site Reliability Engineering job openings:
Infographic showing various Director Site Reliability Engineering job openings in Washington as of August 2026, with employment types broken down into 90% Full Time, and 10% Part Time. Highlights an 70% In-person, 5% Hybrid, and 25% Remote job distribution.

Site Reliability Engineering (SRE) Director

AARP

Washington, DC • On-site

$158K/yr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Posted 5 days ago


AARP rating

7.9

Company rating: 7.9 out of 10

Based on 6 frontline employees who took The Breakroom Quiz

138th of 771 rated non-profit organizations


Job description

Overview
AARP is the nation's largest nonprofit, nonpartisan organization dedicated to empowering people 50 and older to choose how they live as they age. With a nationwide presence, AARP strengthens communities and advocates for what matters most to the more than 100 million Americans 50-plus and their families: health and financial security, and personal fulfillment. AARP also works for individuals in the marketplace by sparking new solutions and allowing carefully chosen, high-quality products and services to carry the AARP name. As a trusted source for news and information, AARP produces the nation's largest-circulation publications, AARP The Magazine and the AARP Bulletin.
The Site Reliability Engineering (SRE) Director is responsible for leading the strategy, operational discipline, and engineering practices that ensure the reliability, availability, scalability, and performance of technical and digital products and platforms. This role drives automation, observability, incident response maturity, and continuous improvement while building a SRE function that supports both enterprise continuity and innovation.
Responsibilities
  • Oversees the reliability and operational readiness of production systems, ensuring platforms can scale effectively and perform consistently under changing demand.
  • Implements and continuously improves automated testing (e.g., QA, regression, performance, and reliability testing) within CI/CD pipelines to proactively identify defects, reduce production risk, and ensure consistent, high-quality releases.
  • Develops and implements reliability goals, service level objectives, and performance expectations for critical digital products and platforms.
  • Establishes and strengthens monitoring, logging, alerting, and observability practices to provide clear visibility into service health and performance.
  • Owns the operational framework for incident management, change management, escalation, service restoration, root cause analysis, and post-incident review.
  • Partners with digital teams to embed reliability, automation, and operational excellence across the software lifecycle and platforms.

Qualifications
  • Bachelor's degree or equivalent experience in Computer Science, Information Technology, Software Engineering, or a related field; advanced degree preferred.
  • Minimum 8 years of progressive experience in Site Reliability Engineering, DevOps, cloud operations, infrastructure engineering, or a related field, including leadership of enterprise-scale reliability and operations functions.
  • Demonstrated experience developing and executing SRE strategies, operational frameworks, and reliability programs that improve system availability, scalability, performance, and organizational resiliency.
  • 5+ years in cloud-native architectures, distributed systems, infrastructure automation, CI/CD pipelines, and modern software delivery practices supporting mission-critical platforms, with experience establishing and governing Secure Software Development Lifecycle (SSDLC) practices that embed security, compliance, reliability, and risk management throughout the software delivery lifecycle.

AARP will not sponsor an employment visa for this position at this time.
Additional Requirements
  • Regular and reliable job attendance
  • Effective verbal and written communication skills
  • Exhibit respect and understanding of others to maintain professional relationships
  • Independent judgement in evaluation options to make sound decisions
  • In office/open office environment with the ability to work effectively surrounded by moderate noise

Hybrid Work Environment
AARP observes Mondays and Fridays as remote workdays, except for essential functions. Remote work can only be done within the United States and its territories.
Compensation and Benefits
AARP offers a competitive compensation and benefits package including a 401(k); 100% company-funded pension plan; health, dental, and vision plans; life insurance; paid time off to include company and individual holidays, vacation, sick, caregiving, and parental leave; performance-based and peer-based recognition and tuition reimbursement.
Equal Employment Opportunity
AARP is an equal opportunity employer committed to hiring a diverse workforce and sustaining an inclusive culture. AARP does not discriminate on the basis of race, ethnicity, religion, sex, color, national origin, age, sexual orientation, gender identity or expression, mental or physical disability, genetic information, veteran status, or on any other basis prohibited by applicable law.

What AARP employees say

Pay

Hours and flexibility

Workplace

Get the full story on Breakroom