1

Director Site Reliability Engineering Jobs in Washington

Site Reliability Engineer

Sterling, VA ยท On-site

$56.50 - $75/hr

Global Skill Development Council (GSDC) Site Reliability Engineering (SRE) Foundation Certification (CSREF). * AWS Certified SysOps Administrator - Associate. * Google Cloud Certified Professional ...

Site Reliability Engineer

Sterling, VA ยท On-site

$56.50 - $75/hr

Global Skill Development Council (GSDC) Site Reliability Engineering (SRE) Foundation Certification (CSREF). * AWS Certified SysOps Administrator - Associate. * Google Cloud Certified Professional ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to ...

Showing results 21-40

Director Site Reliability Engineering information

How much do director site reliability engineering get paid?

Director of Site Reliability Engineering typically earns a salary ranging from $150,000 to $250,000 annually, depending on experience, company size, and location. They often oversee teams using tools like Kubernetes and Prometheus and require strong leadership and technical skills.

What is a director site reliability engineering?

A Director of Site Reliability Engineering (SRE) leads teams responsible for ensuring the availability, performance, and scalability of software systems. They define reliability best practices, drive automation, and collaborate with engineering and product teams to improve system resilience. This role requires strong leadership, technical expertise, and a focus on balancing innovation with operational stability.

What are the main challenges faced by a director site reliability engineering, and how can I prepare for them?

A Director of Site Reliability Engineering often encounters challenges such as balancing rapid feature delivery with system stability, managing complex incident responses, and fostering a culture of continuous improvement. Additionally, aligning reliability goals with business objectives and securing cross-functional buy-in can be demanding. To prepare, it is helpful to gain experience in high-scale system management, develop strong leadership and communication abilities, and cultivate a proactive approach to risk management and automation. Staying up to date with the latest SRE practices and building relationships with both engineering and business teams will also support your success in this pivotal role.

What are the key skills and qualifications needed to thrive as a director site reliability engineering?

To thrive as a Director Site Reliability Engineering, you need extensive experience in software engineering, infrastructure management, incident response, and people leadership, often supported by a degree in computer science or a related field. Familiarity with cloud platforms (such as AWS, GCP, or Azure), automation tools (Terraform, Ansible), monitoring systems (Prometheus, Datadog), and relevant certifications like CKA or AWS Solutions Architect is valued. Outstanding communication, stakeholder management, and strategic vision are key soft skills that set leaders apart in this role. These abilities ensure the reliability, scalability, and efficiency of critical systems while effectively guiding and motivating technical teams.

What job categories do people searching Director Site Reliability Engineering jobs in Washington look for?

The top searched job categories for Director Site Reliability Engineering jobs in Washington are:

What cities in Washington are hiring for Director Site Reliability Engineering jobs?

Cities in Washington with the most Director Site Reliability Engineering job openings:

Infographic showing various Director Site Reliability Engineering job openings in Washington as of August 2026, with employment types broken down into 88% Full Time, and 12% Part Time. Highlights an 79% In-person, 3% Hybrid, and 18% Remote job distribution.

Site Reliability Engineer

Nightwing

Sterling, VA โ€ข On-site

$56.50 - $75/hr

Full-time

Medical, Dental, Vision, Retirement, PTO

This job post hasย expired today.ย Applications are no longer accepted.


Job description

Nightwing provides technically advanced full-spectrum cyber, data operations, systems integration and intelligence mission support services to meet our customers' most demanding challenges. Our capabilities include cyber space operations, cyber defense and resiliency, vulnerability research, ubiquitous technical surveillance, data intelligence, lifecycle mission enablement, and software modernization. Nightwing brings disruptive technologies, agility, and competitive offerings to customers in the intelligence community, defense, civil, and commercial markets.
Job Title: Site Reliability Engineer
Location: Sterling, VA
Clearance: TS/SCI Poly
**This position is CONTINGENT upon contract award**
The Site Reliability Engineer (SRE) collaboratively works closely with the contract leadership, Platform teams, and Sponsor to refine the operational and technical strategy to automate key portions of IT operations and enable the Product team (Platform) to bring new software or new features to production as quickly as possible. The SRE executes and analyzes manual IT operations/admin tasks (log analysis, performance tuning, patch management, testing, and incident response) and converts them to automated tasks. The SRE works with the Platform, Network and Data Operations teams to assist in deployment planning and onboard systems. They assist with monitoring, system analysis, and IT operations support. Daily tasks include, but are not limited to:
  • Work with Sponsor, Mission partners, and technical personnel to deliver robust scalable operations architecture that meets the customer goals for the enterprise.
  • Analyze, define, and document requirements for data, workflow, logical processes, hardware and operating system environment, and network connectivity, other system interfaces, internal and external checks and controls, and outputs.
  • Monitor and track metrics, logs and traces across all services in the system/network and provide context for identifying root causes in the event of an incident, performance degradation, or availability issue.
  • Perform Network/Cloud optimization and resilience planning
  • Develop capabilities to automate hardware/software provisioning, monitoring, patching, and troubleshooting.
  • Collaborate with and assist Platform team and leadership in network and security health, intrusions or inappropriate activities.
  • Optimize business processes, workflows, and service operations by building efficient on-call processes and streamlining alerting workflows.
  • Leverage operational data to automate systems administration, operations and incident response processes to improve enterprise reliability to manage IT environment complexity.
  • Works with LSA, Lab Manager, and CM to compose technical documents including Design, Deployment, System specifications and Host Nation baselines, updates, user's manuals, training materials, installation guides, proposals, and reports.
  • Work with the OM to implement ITSM best practices for ICA/Service discrepancy and reporting, issue resolution and operations support to include Tier 2/3 escalation.

Required Skills:
  • Programming: Proficiency in at least one programming language (e.g., Python, Go, Java, or JavaScript) is essential for automating tasks and developing tools.
  • Linux/Unix Systems Administration: Strong knowledge of Linux/Unix operating systems, including command-line tools and system administration tasks.
  • Networking: Understanding of network protocols, infrastructure, and troubleshooting techniques.
  • Database Management: Familiarity with database technologies and principles.
  • Automation: Experience with automation tools and techniques, such as configuration management (e.g., Ansible, Puppet, Chef) and orchestration (e.g., Kubernetes).
  • Monitoring and Logging: Experience with monitoring tools and logging systems.
  • Problem-Solving: Strong analytical and problem-solving skills to diagnose and resolve system issues.
  • Communication: Ability to communicate technical information clearly and concisely to both technical and non-technical audiences.
  • Collaboration: Ability to work effectively with cross-functional teams, including software developers and operations personnel.

Desired Skills:
  • Cloud Technologies: Experience with cloud platforms (e.g., AWS, Google Cloud, Azure).
  • Containerization: Knowledge of containerization technologies (e.g., Docker, Kubernetes).
  • DevOps Principles: Understanding DevOps principles and practices.
  • Service Level Objectives (SLOs) and Service Level Agreements (SLAs): Experience with defining, tracking, and managing SLOs and SLAs.
  • Data Analysis: Experience with data analysis and visualization tools.

Desired Certs:
  • Global Skill Development Council (GSDC) Site Reliability Engineering (SRE) Foundation Certification (CSREF).
  • AWS Certified SysOps Administrator - Associate.
  • Google Cloud Certified Professional Cloud Architect.
  • Azure Certified Solutions Architect Expert.

Salary & Benefits
The salary associated with this position ($122,000-$253,000) is commensurate with the selected candidate's qualifications, years of relevant experience, and demonstrated level of expertise. Compensation will be determined based on these factors to ensure alignment with skills, responsibilities, and market standards.
Nightwing offers medical, vision and dental insurance coverage in addition to a 401k plan, PTO, Holidays, and additional insurances.
Nightwing is An Equal Opportunity/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or veteran status, age or any other federally protected class.