1

Incident Problem Management Engineer Jobs (NOW HIRING)

Data Science and Data Engineering Job Qualifications: Skills: Data Management Analysis (Inactive ... No Data Incident & Problem Management Analyst Seize your opportunity to make a personal impact ...

... Management bridge during Major Incidents Ability to determine the Incident priority based on ... with Problem Management team during Major Incidents for RCA Additional Information All your ...

Problem Management Engineer

Boston, MA · On-site

$112K - $146K/yr

REQUIRED SKILLS Problem Management Leadership * Lead and oversee the end‑to‑end Problem ... Partner closely with Incident Management, Change Management, Resiliency/Reliability Engineering ...

INCIDENT/PROBLEM MANAGER

Atlanta, GA · On-site

$68K - $122K/yr

... of incident management and problem management experience; or any equivalent combination of ... Good understanding of service development, management and operations industry standard. Strong ...

INCIDENT/PROBLEM MANAGER

Atlanta, GA · On-site

$68K - $122K/yr

... of incident management and problem management experience; or any equivalent combination of ... Good understanding of service development, management and operations industry standard. Strong ...

next page

Showing results 1-20

Incident Problem Management Engineer information

See salary details

$36.5K

$163.4K

$193.5K

How much do incident problem management engineer jobs pay per year?

As of Sep 9, 2026, the average yearly pay for incident problem management engineer in the United States is $163,404.00, according to ZipRecruiter salary data. Most workers in this role earn between $129,000.00 and $193,000.00 per year, depending on experience, location, and employer.

What is an incident problem management engineer?

An Incident Problem Management Engineer is an IT professional responsible for identifying, analyzing, and resolving incidents and problems within an organization's technology environment. They manage the process of restoring normal service operation as quickly as possible after incidents, and work to identify root causes to prevent future issues. Their role often involves close collaboration with other IT teams, conducting post-incident reviews, and implementing solutions to improve system reliability and performance. They play a crucial role in maintaining business continuity and minimizing the impact of IT disruptions.

How does an incident problem management engineer typically interact with other IT teams during major incidents?

An Incident Problem Management Engineer works closely with cross-functional IT teams, such as network operations, application support, and cybersecurity, to coordinate rapid response during major incidents. They facilitate communication between stakeholders, ensure accurate documentation of issues, and lead root cause analysis sessions post-incident. Strong collaboration and clear communication skills are essential, as the engineer often acts as the central point of contact to drive incident resolution and implement long-term preventative measures.

What key skills and qualifications are needed to thrive as an incident problem management engineer, and why are they important?

To excel as an Incident Problem Management Engineer, you need strong analytical skills, a solid understanding of ITIL processes, and experience with incident and problem resolution in complex IT environments. Familiarity with tools like ServiceNow, Jira, and monitoring systems, as well as ITIL or relevant certifications, is highly valued. Excellent communication, critical thinking, and the ability to remain calm under pressure are standout soft skills for this role. These abilities are crucial for minimizing service disruptions, addressing root causes efficiently, and ensuring continuous improvement in IT service management.

What is the difference between Incident Problem Management Engineer vs Network Operations Center (NOC) Engineer?

AspectIncident Problem Management EngineerNetwork Operations Center (NOC) Engineer
CertificationsITIL, Network+, CCNANetwork+, CCNA, CompTIA Security+
Work EnvironmentIT service management, incident analysis, problem resolutionMonitoring network infrastructure, troubleshooting connectivity issues
Employer & IndustryIT service providers, large enterprisesTelecom, internet service providers, data centers
Common Search & ComparisonFocus on incident and problem resolution processesFocus on network monitoring and troubleshooting

The Incident Problem Management Engineer primarily handles incident analysis and problem resolution within IT service management frameworks, often working on root cause analysis and process improvement. In contrast, the Network Operations Center (NOC) Engineer monitors network infrastructure, troubleshoots connectivity issues, and ensures network uptime. While both roles require technical certifications and involve troubleshooting, their focus areas and work environments differ significantly.

What are popular job titles related to Incident Problem Management Engineer jobs?

For Incident Problem Management Engineer jobs, the most frequently searched job titles are:

Director, Incident & Problem Management

Westlake, TX • On-site

Full-time

Re-posted 2 days ago


Job description

Director, Incident & Problem ManagementWho We Are

Solera is a global leader in data and software services, operating at scale across multiple regions, platforms, and products. Our Global IT organization underpins the reliability, availability, and performance of over 400+ products worldwide.

The Role

We are seeking a Director of Incident & Problem Management to lead our global NOC and Problem Management functions. This role is accountable for the end-to-end management of major incidents, operational stability, and root cause remediation across all regions and products.

You will drive a step-change in operational maturity-moving from reactive incident handling to proactive service resilience. This includes building a high-performing global NOC, embedding robust problem management practices, and leveraging data and automation to reduce incident frequency and impact.

This is a high-visibility leadership role requiring strong execution, stakeholder management, and the ability to operate in a complex, fast-paced environment.

What You'll Do

Operational Leadership

  • Lead the global NOC and Problem Management teams, operating 24/7 across multiple regions
  • Own the end-to-end major incident management process, including escalation, coordination, and communication
  • Establish clear command-and-control structures during incidents to drive rapid resolution

Problem Management & Continuous Improvement

  • Implement and mature a structured problem management framework
  • Ensure high-quality root cause analysis (RCA) and action tracking
  • Drive permanent fixes and reduce repeat incidents and known errors

Service Stability & Risk Reduction

  • Identify systemic risks, single points of failure, and operational weaknesses
  • Define and implement mitigation strategies across teams and platforms
  • Improve service reliability, uptime, and customer impact metrics

Data, Automation & AI

  • Leverage data and analytics to identify incident trends and operational insights
  • Introduce automation and AI-driven capabilities to improve detection, triage, and response
  • Drive a shift from reactive support to predictive operations

Stakeholder & Executive Engagement

  • Act as the central point of accountability for major incidents at an executive level
  • Communicate clearly and directly with senior leadership during critical events
  • Partner with Product, Engineering, and Support teams to improve service outcomes

Team & Capability Building

  • Build and scale a high-performing, globally distributed team
  • Define clear roles, responsibilities, and performance expectations
  • Develop leadership capability within the NOC and Problem Management functions
What You'll Bring

Experience

  • 10+ years in IT operations, incident management, or service delivery leadership roles
  • Proven experience leading global, distributed teams in a 24/7 environment
  • Strong track record of improving operational maturity and service stability

Technical & Operational Expertise

  • Deep understanding of incident and problem management frameworks (e.g. ITIL)
  • Experience managing large-scale, complex production environments
  • Strong data-driven mindset with experience using dashboards and reporting tools
  • Exposure to automation, AI, or modern operational tooling

Leadership & Behaviour

  • Strong leadership presence with the ability to operate under pressure
  • Clear, direct communicator, especially in high-stakes situations
  • Highly accountable, outcome-focused, and execution-driven
  • Ability to challenge constructively and drive change across functions

It is impossible to list every requirement for, or responsibility of, any position. Similarly, we cannot identify all the skills a position may require since job responsibilities and the Company's needs may change over time. Therefore, the above job description is not comprehensive or exhaustive. The Company reserves the right to adjust, add to or

eliminate any aspect of the above description. The Company also retains the right to require all employees to undertake additional or different job responsibilities when necessary to meet business needs.


EQUAL OPPORTUNITY EMPLOYER
The Solera group is an equal opportunity employer and complies with all applicable federal, state, and local fair employment practices laws.