1

Reliability Engineer Manager Jobs in Ashburn, VA

SRE ENGINEER/ MANAGER

Reston, VA ยท On-site

$59.25 - $78.75/hr

Job Summary (Sr. Manager SRE): - Design, implement, and manage scalable, secure, and fault-tolerant cloud infrastructure using AWS, Azure, or GCP. - Automate infrastructure provisioning and ...

Reliability Engineer

Rockville, MD ยท On-site

$104K - $131K/yr

Help develop and continuously improve Quantum Space's reliability engineering processes, procedures, analysis templates, and requirements within our AS9100 Rev D Quality Management System. What It ...

Reliability Engineer

Rockville, MD ยท On-site

$120 - $180/hr

Help develop and continuously improve Quantum Space's reliability engineering processes, procedures, analysis templates, and requirements within our AS9100 Rev D Quality Management System. What It ...

New

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

Site Reliability Engineer (SRE) Remote No sponsorship available. Must be able to obtain a Public ... Support incident management activities, including troubleshooting, root-cause analysis, mitigation ...

Koniag Management Solutions, LLC (KMS), a Koniag Government Services (KGS) company, is hiring a Site Reliability Engineer (SRE). Position requires an active Top Secret/SCI clearance with ability to ...

Site Reliability Engineer (SRE)

Washington, DC ยท On-site

$64.50 - $85.75/hr

Design and manage CI/CD pipelines using GitHub Actions, Jenkins, or AWS CodePipeline. * Automate ... Define and monitor SRE metrics including SLIs, SLOs, and error budgets. * Perform performance ...

Site Reliability Engineer (SRE)

Vienna, VA ยท On-site

$57.25 - $76/hr

The AWS Site Reliability Engineer (SRE) is responsible for the operational health, availability ... CloudWatch, performance tuning in cloud environments, IaC tools, Databricks management and ...

next page

Showing results 1-20

Reliability Engineer Manager information

See Ashburn, VA salary details

$62.4K

$120.6K

$144.2K

How much do reliability engineer manager jobs pay per year?

As of Aug 13, 2026, the average yearly pay for reliability engineer manager in Ashburn, VA is $120,639.00, according to ZipRecruiter salary data. Most workers in this role earn between $104,800.00 and $131,900.00 per year, depending on experience, location, and employer.

What does a reliability engineer manager do?

A Reliability Engineer Manager oversees teams responsible for improving the reliability and performance of systems, machinery, or processes within an organization. They develop maintenance strategies, lead root cause analyses of failures, and implement best practices to minimize downtime and costs. Additionally, they collaborate with other departments to ensure that reliability goals align with business objectives and compliance standards. Their role is crucial in industries such as manufacturing, energy, and technology, where system uptime and safety are critical.

What are some common challenges reliability engineer managers face when balancing long-term reliability improvements with immediate operational demands?

Reliability Engineer Managers often need to prioritize urgent maintenance issues while also driving long-term reliability initiatives. Balancing these competing demands can be challenging, as immediate equipment failures may require quick fixes that temporarily interrupt ongoing improvement projects. Effective managers work closely with operations, maintenance, and engineering teams to communicate priorities, allocate resources, and implement sustainable solutions that address root causes rather than just symptoms. This role typically involves using data-driven decision-making and fostering a culture of proactive maintenance and continuous improvement.

What are the key skills and qualifications needed to thrive as a reliability engineer manager?

To thrive as a Reliability Engineer Manager, you need a strong background in engineering principles, reliability analysis, and maintenance strategies, typically supported by a degree in engineering and experience in reliability roles. Familiarity with reliability-centered maintenance (RCM), failure mode and effects analysis (FMEA), and asset management software such as SAP or Maximo is common, along with certifications like Certified Reliability Engineer (CRE). Leadership, problem-solving, and effective communication are vital soft skills for managing teams and driving cross-functional initiatives. These competencies are crucial for minimizing downtime, optimizing equipment performance, and ensuring long-term operational efficiency.

What is the difference between Reliability Engineer Manager vs Reliability Engineer?

AspectReliability EngineerReliability Engineer Manager
Required CredentialsBachelor's in Engineering or related field; certifications like CRC, CRESame as Reliability Engineer, plus leadership experience
Work EnvironmentDesign, analyze, and improve system reliability; often in teamsOversees Reliability Engineers; manages projects and teams
Employer & Industry UsageManufacturing, aerospace, energy, automotiveSame industries, with added managerial responsibilities
Common Search & ComparisonFocuses on technical skills and hands-on reliability tasksFocuses on leadership, team management, and strategic planning

The main difference between a Reliability Engineer and a Reliability Engineer Manager lies in their responsibilities. The Reliability Engineer focuses on technical analysis and system improvements, while the Reliability Engineer Manager oversees teams, manages projects, and develops strategies to enhance reliability across the organization.

What are the most commonly searched types of Reliability Engineer jobs in Ashburn, VA?

The most popular types of Reliability Engineer jobs in Ashburn, VA are:

What are popular job titles related to Reliability Engineer Manager jobs in Ashburn, VA?

For Reliability Engineer Manager jobs in Ashburn, VA, the most frequently searched job titles are:

What job categories do people searching Reliability Engineer Manager jobs in Ashburn, VA look for?

The top searched job categories for Reliability Engineer Manager jobs in Ashburn, VA are:

What cities near Ashburn, VA are hiring for Reliability Engineer Manager jobs?

Cities near Ashburn, VA with the most Reliability Engineer Manager job openings:

SRE ENGINEER/ MANAGER

Accenuate IT Solutions

Reston, VA โ€ข On-site

$59.25 - $78.75/hr

Contractor

Re-posted 27 days ago


Job description

Job Summary (Sr. Manager SRE):
- Design, implement, and manage scalable, secure, and fault-tolerant cloud infrastructure using AWS, Azure, or GCP.
- Automate infrastructure provisioning and configuration with IaC tools such as Terraform, CloudFormation, and Ansible.
- Develop automation for anomaly detection, recovery, toil reduction, self-healing, and cloud cost optimization.
- Lead implementation of DevSecOps best practices and maintain secure CI/CD pipelines (e.g., GitLab, Jenkins, Docker).
- Enforce security through IAM, RBAC, vulnerability remediation, and security scanning tools (SAST/DAST/SCA).
- Architect and manage microservices, serverless solutions, and APIs with a focus on fault tolerance and resilience.
- Implement and maintain monitoring, logging, and observability using CloudWatch, Splunk/SignalFX, Dynatrace, and OpenTelemetry.
- Drive incident management, lead root cause analysis, postmortems, and minimize MTTR/MTTD.
- Define and monitor system reliability metrics (SLOs, SLIs, error budgets).
- Conduct chaos engineering, resiliency assessments, and implement self-healing architectures.
- Manage and optimize databases (PostgreSQL, MongoDB, DynamoDB, Oracle, Redshift) and provide production support.
- Participate in on-call rotations and support incident/problem management.
- Collaborate with development, QA, and operations teams to implement shift-left testing practices (BDD, TDD, Unit, Regression).
- Maintain architecture diagrams, knowledge documentation, and disaster recovery plans.
- Communicate effectively with stakeholders and demonstrate strong relationship management across teams.