1

Incident Problem Manager Jobs in Washington, DC (NOW HIRING)

Manager, Incident Response

Mclean, VA ยท On-site +1

$150K - $175K/yr

About the Role As the Manager, Incident Response at Pondurance, you will help manage our Incident ... Proven track record of complex problem-solving and decision-making ability * Expert level of ...

The manager leads a team responsible for incident, request, problem, and knowledge management, while partnering closely with Infrastructure, Networking, Security, Applications, and business ...

Manager, Incident Response

Mclean, VA ยท Remote

$150K - $175K/yr

About the Role As the Manager, Incident Response at Pondurance, you will help manage our Incident ... Proven track record of complex problem-solving and decision-making ability * Expert level of ...

Showing results 41-60

Incident Problem Manager information

See Washington, DC salary details

$41.3K

$185.1K

$219.2K

How much do incident problem manager jobs pay per year?

As of Aug 22, 2026, the average yearly pay for incident problem manager in Washington, DC is $185,071.00, according to ZipRecruiter salary data. Most workers in this role earn between $146,100.00 and $218,600.00 per year, depending on experience, location, and employer.

What does an Incident Problem Manager do?

An Incident Problem Manager is responsible for overseeing the process of identifying, investigating, and resolving incidents and underlying problems within an organization's IT systems. They work to minimize the impact of disruptions, coordinate responses to incidents, and analyze root causes to prevent future occurrences. This role often involves collaborating with technical teams, managing communication with stakeholders, and ensuring that procedures are followed according to IT service management frameworks like ITIL. Their goal is to improve IT service reliability and reduce downtime for the business.

How does an Incident Problem Manager typically collaborate with technical and non-technical teams during major incidents?

Incident Problem Managers play a critical role in bridging communication between technical teams (like IT support, network engineers, or developers) and non-technical stakeholders (such as business unit leaders or customer service). During major incidents, they coordinate response efforts, facilitate status updates, and ensure all parties are aligned on next steps and remediation plans. Effective collaboration involves translating complex technical issues into clear, actionable information for non-technical audiences, managing expectations, and driving post-incident reviews to prevent recurrence. This cross-functional coordination is essential for minimizing business impact and ensuring swift resolution.

What are the key skills and qualifications needed to thrive as an Incident Problem Manager, and why are they important?

To thrive as an Incident Problem Manager, you need strong analytical skills, IT service management knowledge, and experience with incident and problem resolution processes, often supported by ITIL certification. Familiarity with ITSM tools like ServiceNow, Jira Service Management, or BMC Remedy is typically required. Exceptional communication, leadership, and critical thinking abilities enable effective coordination and root-cause analysis across teams. These skills are crucial to minimizing downtime, ensuring service continuity, and driving long-term improvements in IT operations.

What is the difference between Incident Problem Manager vs Incident Coordinator?

AspectIncident Problem ManagerIncident Coordinator
Primary RoleManages the lifecycle of incidents and problems to minimize impact and prevent recurrenceCoordinates incident response activities, ensuring timely resolution and communication
CertificationsITIL Foundation, Problem Management certificationsITIL Foundation, Incident Management certifications
Work EnvironmentTypically in IT service management teams, focusing on problem analysisOperational teams, focusing on incident handling and communication

While both roles are involved in incident management, the Incident Problem Manager focuses on identifying root causes and preventing future issues, whereas the Incident Coordinator handles day-to-day incident response and communication. Both roles are essential for effective IT service delivery but differ in scope and responsibilities.

What are popular job titles related to Incident Problem Manager jobs in Washington, DC?

For Incident Problem Manager jobs in Washington, DC, the most frequently searched job titles are:

What job categories do people searching Incident Problem Manager jobs in Washington, DC look for?

The top searched job categories for Incident Problem Manager jobs in Washington, DC are:

Infographic showing various Incident Problem Manager job openings in Washington, DC as of August 2026, with employment types broken down into 87% Full Time, 11% Part Time, and 2% Contract. Highlights an 83% Physical, 2% Hybrid, and 15% Remote job distribution, with an average salary of $185,071 per year, or $89 per hour.

INCIDENT MANAGEMENT ENGINEER

BTree Solutions Inc

Reston, VA โ€ข On-site

Contractor

Re-posted 11 days ago


Job description

Job Title: INCIDENT MANAGEMENT ENGINEER
Location: Reston, VA of Plano,TX
Duration: 12+ Months
Visa: USC, GC, H1B and EAD
Contract Type: W2

Description:

In this incident management function, manage incidents to resolution in a 24/7/365 environment using the incident management processes, effectively guide incident and triage calls from a technical perspective, share technical details obtained from monitoring tools and dashboards to aid troubleshooting, outline details of resolution activities, recommend and implement improved processes, provide timely status updates to stakeholders, assist with postmortem related activities and support various efforts related to operational improvements. Manage efforts to maintain application in production, including troubleshooting stoppages, repairing bugs, documenting application performance, and coordinating with technology infrastructure management.

KEY JOB FUNCTIONS

  • Excellent communicator who can manage IT incidents to resolution in a 24/7/365 environment using the Fannie Mae incident management processes and communicate management of incident status, impact and resolution actions.
  • Hands on experience managing and monitoring applications deployed on Amazon Web Services (AWS).
  • Troubleshooting and resolving incidents on the AWS cloud infrastructure.
  • Experience with building tools for monitoring and troubleshooting of system resources in an AWS environment. Ability to triage AWS related incidents using monitoring tools on AWS Cloud.
  • Experience with performance engineering of AWS Cloud applications.
  • Hands on experience working with AWS tools like EC2, ELB, RDS, Redshift, DynamoDB, Aurora, Route53, ECS, Lambada, S3, Batch, CloudWatch, CloudTrail, WAF etc.
  • Hands on experience with transaction level monitoring using Dynatrace and Splunk.
  • Ability to perform transaction level monitoring and troubleshooting in AWS cloud platform.
  • Eyes on glass monitoring of the health of applications as well as the underlying infrastructure.
  • Monitoring experience with tools like Extrahop, SolarWinds, Netcool suite, Catchpoint, MoogSoft.
  • Ability to analyze dashboards and reporting/monitoring tools to look at trends and patterns in application health and performance.
  • Proactively looking for hardware, software, and environmental alerts or malfunctions.
  • Effectively lead and guide Incident triage calls from a technical perspective analyzing different components of the infrastructure and application environment via the use of a variety of monitoring tools and processes.
  • Troubleshoot the incidents and identify root cause quickly using operations, wire data analytics, application performance management and event correlation monitoring tools.
  • Perform analysis of data, evaluating multiple application protocols including web, database, storage, and supporting infrastructure such as AWS, UNIX, DNS, LDAP, SSL, SMTP, and FTP.
  • Influence other technical teams on the calls and articulate troubleshooting steps effectively.
  • Lead required technical follow-up calls for critical incidents.
  • Assist with documentation of Root Cause Analysis (RCA) or Correction of Errors (COE) and data quality for all ECC communicated incidents.
  • Ensure appropriate functional and management escalation takes place as per the standards and procedures.
  • Follow up on items that could potentially negatively impact production operations, assist with postmortem related activities and support various efforts related to operational improvements.
  • Based on recommendations from management, implement new and improved processes, change processes, perform new tasks, create reports and address ad-hoc requests.
  • Participate in on-call rotation. Ability to work on any shifts as needed including weekends and night shifts.
  • Ability to report incident details and metrics to senior leadership.
EDUCATION:
  • Bachelor's Degree or equivalent required.
MINIMUM EXPERIENCE:
  • 10+ years of related experience
SPECIALIZED KNOWLEDGE & SKILLS:
  • 10+ years of working experience with different IT Infrastructure components such as Unix/ Linux Servers, Wintel Servers, AWS, networks, firewalls, routers, load balancers, VPN, Apache, web logic, LDAP, Active Directory, Exchange, Oracle/MS SQL databases, SAN, Virtualization, Email systems, Enterprise monitoring and access management solutions for single sign on. Subject matter expertise is not required and experience with at least eight of the above is preferred.
  • Senior level hands-on working experience with Amazon Web Services (AWS).
  • Understanding different layers of the AWS Infrastructure e.g. WAF, R53, CloudFront, Load Balancing, HA features.
  • Proven methodical approach to problem identification, monitoring, problem solving and resolution.
  • Ability to analyze different components of the infrastructure and application environments during Incident triage calls.
  • Ability to trace transaction failures and debug the root cause in various layers of the AWS infrastructure and services.
  • Aptitude to influence other technical teams on incident calls & articulate troubleshooting steps effectively.
  • Experience & confidence working with all levels of management; excellent written and verbal skills.
  • Able to quickly and concisely communicate with senior management on technical issues in non-technical terms and to run large conference calls during Incident calls with a wide range of personnel and management levels.
  • Strong relationship management skills and aptitude to multi-task and work well in a high stress environment, both within teams and independently.
  • AWS Solution Architect Associate or higher certification
  • Monitoring and observability experience.
  • Experience with monitoring dashboards for incident detection and alerting.
  • Perform end-to-end analysis of transactions under an observability environment.
  • Troubleshoot incidents and identify root cause quickly using wire data analytics, application performance management and event correlation monitoring tools.
  • Diagnose and resolve incidents by providing factual data from the various monitoring and instrumentation systems.
  • Monitor applications and infrastructure using tools like Splunk, DynaTrace, OpenTel, Catchpoint, xMatters, SignalFx, xMatters, SolarWinds, Extrahop etc.
Preferred Qualifications:
  • Management and troubleshooting of Middleware products on UNIX and Linux environments. Knowledge of Service Oriented Architecture (SOA), Java etc.
  • Prior Financial industry experience.
  • Experience with OpenTel