1

Reliability Engineer Manager Jobs in Washington, DC

The SRE will work closely with software developers and operations teams to improve system ... Manage cloud and database system maintenance, debugging production issues as they arise. * Improve ...

Reliability Engineer

Ashburn, VA ยท On-site

$120 - $160/hr

The ideal candidate will lead monitoring solutions, manage ITIL engineers, automate processes, and collaborate across IT and business teams to improve service reliability. Expertise in AWS ...

Site Reliability Engineer - Hybrid

Reston, VA ยท On-site

$59.25 - $78.75/hr

Second round would be an in-person interview Manager's call notes * This is an SRE role. SRE is under a shared services team within Fannie Mae who works with different application teams. So, multi ...

Reliability Engineer

Washington, DC ยท On-site

$116K - $146K/yr

The ideal candidate will lead monitoring solutions, manage ITIL engineers, automate processes, and collaborate across IT and business teams to improve service reliability. Expertise in AWS ...

Site Reliability Engineer IV

Sterling, VA ยท On-site

$120 - $160/hr

Strong understanding of JVM fundamentals (heap/memory management, garbage collection, OOM issues, thread analysis) * Proven experience with SRE practices, including: * Incident response and on-call ...

Site Reliability Engineer

Washington, DC ยท On-site

$112K - $179K/yr

The SRE will drive automation initiatives, observability improvements, and incident response ... Define and manage Service Level Objectives (SLOs) and Service Level Indicators (SLIs). * Support ...

Reliability Engineer

Mclean, VA ยท On-site

$103K - $130K/yr

The Reliability Engineer will act generally as a member of a design, analysis or review team on ... Coordinate and work closely with other engineering, logistics, financial, and program management ...

Site Reliability Engineer

Washington, DC ยท On-site

$112K - $179K/yr

The SRE will drive automation initiatives, observability improvements, and incident response ... Define and manage Service Level Objectives (SLOs) and Service Level Indicators (SLIs). * Support ...

Showing results 41-60

Reliability Engineer Manager information

See Washington, DC salary details

$69.1K

$133.6K

$159.7K

How much do reliability engineer manager jobs pay per year?

As of Aug 17, 2026, the average yearly pay for reliability engineer manager in Washington, DC is $133,616.00, according to ZipRecruiter salary data. Most workers in this role earn between $116,100.00 and $146,100.00 per year, depending on experience, location, and employer.

What does a reliability engineer manager do?

A Reliability Engineer Manager oversees teams responsible for improving the reliability and performance of systems, machinery, or processes within an organization. They develop maintenance strategies, lead root cause analyses of failures, and implement best practices to minimize downtime and costs. Additionally, they collaborate with other departments to ensure that reliability goals align with business objectives and compliance standards. Their role is crucial in industries such as manufacturing, energy, and technology, where system uptime and safety are critical.

What are the key skills and qualifications needed to thrive as a reliability engineer manager?

To thrive as a Reliability Engineer Manager, you need a strong background in engineering principles, reliability analysis, and maintenance strategies, typically supported by a degree in engineering and experience in reliability roles. Familiarity with reliability-centered maintenance (RCM), failure mode and effects analysis (FMEA), and asset management software such as SAP or Maximo is common, along with certifications like Certified Reliability Engineer (CRE). Leadership, problem-solving, and effective communication are vital soft skills for managing teams and driving cross-functional initiatives. These competencies are crucial for minimizing downtime, optimizing equipment performance, and ensuring long-term operational efficiency.

What are some common challenges reliability engineer managers face when balancing long-term reliability improvements with immediate operational demands?

Reliability Engineer Managers often need to prioritize urgent maintenance issues while also driving long-term reliability initiatives. Balancing these competing demands can be challenging, as immediate equipment failures may require quick fixes that temporarily interrupt ongoing improvement projects. Effective managers work closely with operations, maintenance, and engineering teams to communicate priorities, allocate resources, and implement sustainable solutions that address root causes rather than just symptoms. This role typically involves using data-driven decision-making and fostering a culture of proactive maintenance and continuous improvement.

What is the difference between Reliability Engineer Manager vs Reliability Engineer?

AspectReliability EngineerReliability Engineer Manager
Required CredentialsBachelor's in Engineering or related field; certifications like CRC, CRESame as Reliability Engineer, plus leadership experience
Work EnvironmentDesign, analyze, and improve system reliability; often in teamsOversees Reliability Engineers; manages projects and teams
Employer & Industry UsageManufacturing, aerospace, energy, automotiveSame industries, with added managerial responsibilities
Common Search & ComparisonFocuses on technical skills and hands-on reliability tasksFocuses on leadership, team management, and strategic planning

The main difference between a Reliability Engineer and a Reliability Engineer Manager lies in their responsibilities. The Reliability Engineer focuses on technical analysis and system improvements, while the Reliability Engineer Manager oversees teams, manages projects, and develops strategies to enhance reliability across the organization.

What are the most commonly searched types of Reliability Engineer jobs in Washington, DC?

The most popular types of Reliability Engineer jobs in Washington, DC are:

What job categories do people searching Reliability Engineer Manager jobs in Washington, DC look for?

The top searched job categories for Reliability Engineer Manager jobs in Washington, DC are:

DevOps Site Reliability Engineer (SRE)

IT Veterans LLC

Washington, DC โ€ข On-site

$64.50 - $85.75/hr

Other

Posted 23 days ago


Job description

DevOps Site Reliability Engineer (SRE)
Location: Washington, D.C.
Clearance Required: TS/SCI
Position Overview
IT Veterans is seeking a DevOps Site Reliability Engineer (SRE) to support the reliability, performance, and operational stability of a mission-critical enterprise platform. This position plays a vital role in ensuring continuous availability across multi-cloud environments while supporting software deployments, infrastructure monitoring, incident response, and system automation.
The ideal candidate will help bridge development and operations by implementing reliable deployment practices, building robust monitoring capabilities, and rapidly responding to production issues to maintain the required 99.9% platform availability for critical Department of Defense (DoD) systems.
Key Responsibilities:
  • Monitor the health and performance of enterprise infrastructure through continuous system monitoring and automated telemetry to support the required 99.9% platform uptime.
  • Participate in an on-call rotation and respond to major incidents or platform outages within one hour of notification, executing rapid troubleshooting and system stabilization activities.
  • Develop, maintain, and enhance automation scripts and internal tools that streamline diagnostics, health checks, and routine operational tasks.
  • Design and maintain dashboards that provide real-time visibility into platform health, including uptime, API performance, incident status, and other key operational metrics.
  • Coordinate directly with Cloud Service Providers (CSPs) during infrastructure outages or service disruptions to expedite issue resolution.
  • Continuously assess system reliability, logging, monitoring, and overall architecture, providing recommendations that improve scalability, resiliency, and operational efficiency.
Required Qualifications:
  • TS/SCI security clearance.
  • Strong understanding of Site Reliability Engineering (SRE) principles and best practices.
  • Hands-on experience deploying and managing containerized applications using Kubernetes.
  • Experience administering and troubleshooting multi-cloud environments, including Google Cloud Platform (Google Cloud Platform), Microsoft Azure, and Amazon Web Services (AWS).
  • Experience implementing and maintaining enterprise monitoring, logging, and automated alerting solutions.
  • Proficiency with scripting and automation using languages such as Python, Bash, or similar technologies.
Preferred Qualifications:
  • Passion for building and maintaining highly reliable, mission-critical systems with demanding uptime requirements.
  • Ability to remain composed and methodical while responding to high-priority production incidents.
  • Strong troubleshooting, root cause analysis, and diagnostic skills with a focus on rapid issue resolution.
  • A continuous improvement mindset with an emphasis on automation and eliminating repetitive operational tasks.
  • Experience supporting secure, cloud-native environments within government or defense organizations is a plus.

At IT Veterans LLC, we are committed to providing an environment of mutual respect where equal employment opportunities are available to all applicants and teammates without regard to race, color, religion, sex, pregnancy, national origin, age, physical and mental disability, marital status, sexual orientation, gender identity, gender expression, genetic information, military and veteran status, and any other characteristic protected by applicable law. We believe that diversity and inclusion among our teammates is critical to our success.