Incident & Problem Management: Triage, investigate, and resolve production incidents; participate in root cause analysis and drive long-term fixes. * Release & Deployment Operations: Coordinate and ...
Incident & Problem Management: Triage, investigate, and resolve production incidents; participate in root cause analysis and drive long-term fixes. * Release & Deployment Operations: Coordinate and ...
Incident & Problem Management: Triage, investigate, and resolve production incidents; participate in root cause analysis and drive long-term fixes. * Release & Deployment Operations: Coordinate and ...
Incident & Problem Management: Triage, investigate, and resolve production incidents; participate in root cause analysis and drive long-term fixes. * Release & Deployment Operations: Coordinate and ...
Experience with incident, problem, and service-request management. * Experience performing root cause analysis and troubleshooting complex enterprise issues. * Experience working in an on-call ...
New
Quick apply
Experience with incident, problem, and service-request management. * Experience performing root cause analysis and troubleshooting complex enterprise issues. * Experience working in an on-call ...
New
Network Architect
$64 - $85.75/hr
Advanced Cisco certifications highly desired (CCIE, CCNP, CCDP, CCSP, etc) Familiarity with ITIL v3 based service delivery model using change/incident/problem management highly desired. Additional ...
Network Architect
$64 - $85.75/hr
Advanced Cisco certifications highly desired (CCIE, CCNP, CCDP, CCSP, etc) Familiarity with ITIL v3 based service delivery model using change/incident/problem management highly desired. Additional ...
Sr Director, Engineering - MP&A and Finance
Philadelphia, PA · On-site
$220K - $240K/yr
... release management, incident/problem management, and operational excellence. • Experience operating within governance, risk, compliance, audit, and security frameworks, including SOX-aware ...
Sr Director, Engineering - MP&A and Finance
Philadelphia, PA · On-site
$220K - $240K/yr
... release management, incident/problem management, and operational excellence. • Experience operating within governance, risk, compliance, audit, and security frameworks, including SOX-aware ...
SAP Supply Chain Solution Architect
Chester, VA · On-site
$135K - $175K/yr
Drive incident and problem management with a strong root cause and prevention mindset, not just ticket resolution. * Establish and maintain clear operating procedures, functional playbooks, and ...
SAP Supply Chain Solution Architect
Chester, VA · On-site
$135K - $175K/yr
Drive incident and problem management with a strong root cause and prevention mindset, not just ticket resolution. * Establish and maintain clear operating procedures, functional playbooks, and ...
Client Computing Engineer
Lancaster, PA · On-site
Working experience with, but not limited to, the following process areas: incident management, problem management, change management, patching, identity and access management (IAM), lifecycle ...
Client Computing Engineer
Lancaster, PA · On-site
Working experience with, but not limited to, the following process areas: incident management, problem management, change management, patching, identity and access management (IAM), lifecycle ...
Working experience with, but not limited to, the following process areas: incident management, problem management, change management, patching, identity and access management (IAM), lifecycle ...
Working experience with, but not limited to, the following process areas: incident management, problem management, change management, patching, identity and access management (IAM), lifecycle ...
ServiceNow CMDB Architect
Philadelphia, PA · On-site
$51.50 - $71/hr
Incident, Problem, Change, and Asset Mgt.) • Experience working with stakeholders to understand ... ITSM, CMDB, Discovery, experience with ServiceNow enterprise architecture, security, database ...
ServiceNow CMDB Architect
Philadelphia, PA · On-site
$51.50 - $71/hr
Incident, Problem, Change, and Asset Mgt.) • Experience working with stakeholders to understand ... ITSM, CMDB, Discovery, experience with ServiceNow enterprise architecture, security, database ...
IT Service Desk Lead
Lancaster, PA · On-site
... management processes (incident, problem and change management) and platforms (service desk and IT asset management). The Service Desk Lead operates as part of the Client Computing team under the ...
IT Service Desk Lead
Lancaster, PA · On-site
... management processes (incident, problem and change management) and platforms (service desk and IT asset management). The Service Desk Lead operates as part of the Client Computing team under the ...
IT Service Desk Lead
Lancaster, PA · On-site
... management processes (incident, problem and change management) and platforms (service desk and IT asset management). The Service Desk Lead operates as part of the Client Computing team under the ...
IT Service Desk Lead
Lancaster, PA · On-site
... management processes (incident, problem and change management) and platforms (service desk and IT asset management). The Service Desk Lead operates as part of the Client Computing team under the ...
Familiarity with ITIL or ITSM practices, including incident, problem, and request management. * Experience with hybrid and cloud networking environments. * Exposure to automation, network monitoring ...
Familiarity with ITIL or ITSM practices, including incident, problem, and request management. * Experience with hybrid and cloud networking environments. * Exposure to automation, network monitoring ...
Core responsibilities include leading major incident and problem management activities, as well as overseeing Event, Request, Change, and Knowledge management processes. This is an opportunity to ...
Core responsibilities include leading major incident and problem management activities, as well as overseeing Event, Request, Change, and Knowledge management processes. This is an opportunity to ...
Core responsibilities include leading major incident and problem management activities, as well as overseeing Event, Request, Change, and Knowledge management processes. This is an opportunity to ...
Core responsibilities include leading major incident and problem management activities, as well as overseeing Event, Request, Change, and Knowledge management processes. This is an opportunity to ...
Participate in after-hours major incident problem resolution on an on-call basis Required Skills ... Operating System Management * Database Technology * Disaster Recovery and Continuity of Operations
Participate in after-hours major incident problem resolution on an on-call basis Required Skills ... Operating System Management * Database Technology * Disaster Recovery and Continuity of Operations
ITIL (Incident/Problem Management, CMDB) (ServiceNow) The ideal candidate should have working knowledge of: * DBA best practices and concepts: database backup & recovery, database performance tuning ...
ITIL (Incident/Problem Management, CMDB) (ServiceNow) The ideal candidate should have working knowledge of: * DBA best practices and concepts: database backup & recovery, database performance tuning ...
ITIL (Incident/Problem Management, CMDB) (ServiceNow) The ideal candidate should have working knowledge of: * DBA best practices and concepts: database backup & recovery, database performance tuning ...
ITIL (Incident/Problem Management, CMDB) (ServiceNow) The ideal candidate should have working knowledge of: * DBA best practices and concepts: database backup & recovery, database performance tuning ...
Senior Site EHS Manager
Carlisle, PA · On-site
$80K - $109K/yr
Lead, track, and review all site level incident problem solving for complete root causes, trends, corrective action completion. * Manage execution of Carlisle wide and/or Business Unit read-across ...
Senior Site EHS Manager
Carlisle, PA · On-site
$80K - $109K/yr
Lead, track, and review all site level incident problem solving for complete root causes, trends, corrective action completion. * Manage execution of Carlisle wide and/or Business Unit read-across ...
Senior Site EHS Manager
Carlisle, PA · On-site
$80K - $109K/yr
Lead, track, and review all site level incident problem solving for complete root causes, trends, corrective action completion. * Manage execution of Carlisle wide and/or Business Unit read-across ...
Senior Site EHS Manager
Carlisle, PA · On-site
$80K - $109K/yr
Lead, track, and review all site level incident problem solving for complete root causes, trends, corrective action completion. * Manage execution of Carlisle wide and/or Business Unit read-across ...
Site Reliability Engineer (SRE
Pittsburgh, PA · On-site
$53.25 - $70.75/hr
Participate in incident, problem, and change management processes. * Perform Root Cause Analysis (RCA) and implement permanent solutions to recurring production issues. * Develop and maintain ...
New
Site Reliability Engineer (SRE
Pittsburgh, PA · On-site
$53.25 - $70.75/hr
Participate in incident, problem, and change management processes. * Perform Root Cause Analysis (RCA) and implement permanent solutions to recurring production issues. * Develop and maintain ...
New
Incident Problem Manager information
What does an Incident Problem Manager do?
How does an Incident Problem Manager typically collaborate with technical and non-technical teams during major incidents?
What are the key skills and qualifications needed to thrive as an Incident Problem Manager, and why are they important?
What is the difference between Incident Problem Manager vs Incident Coordinator?
| Aspect | Incident Problem Manager | Incident Coordinator |
|---|---|---|
| Primary Role | Manages the lifecycle of incidents and problems to minimize impact and prevent recurrence | Coordinates incident response activities, ensuring timely resolution and communication |
| Certifications | ITIL Foundation, Problem Management certifications | ITIL Foundation, Incident Management certifications |
| Work Environment | Typically in IT service management teams, focusing on problem analysis | Operational teams, focusing on incident handling and communication |
While both roles are involved in incident management, the Incident Problem Manager focuses on identifying root causes and preventing future issues, whereas the Incident Coordinator handles day-to-day incident response and communication. Both roles are essential for effective IT service delivery but differ in scope and responsibilities.
What are popular job titles related to Incident Problem Manager jobs in Pennsylvania?
For Incident Problem Manager jobs in Pennsylvania, the most frequently searched job titles are:
What job categories do people searching Incident Problem Manager jobs in Pennsylvania look for?
The top searched job categories for Incident Problem Manager jobs in Pennsylvania are:
What cities in Pennsylvania are hiring for Incident Problem Manager jobs?
Cities in Pennsylvania with the most Incident Problem Manager job openings:

Lead Enterprise Software Engineer - Application & IT Operations
Philadelphia, PA • On-site
Full-time
Medical, Dental, Vision, Retirement, PTO
Re-posted 12 days ago
Wolters Kluwer rating
9.1
Based on 29 frontline employees who took The Breakroom Quiz
25th of 247 rated software companies
Job description
Wolters Kluwer is seeking a motivated and talented Lead AppOps Engineer to join our dynamic team. This role is ideal for an individual with hands-on experience in application operations, production support, and cloud-platform application management within a mature engineering environment. As a Lead AppOps Engineer, you will play a crucial role in ensuring the stability, reliability, availability, and operational excellence of our enterprise applications.
In this role, your primary responsibility will be managing the end-to-end operational lifecycle of critical applications. You will work closely with engineering teams, Cloud Operations, Security, Compliance, and other key partners to ensure smooth application deployments, reliable production operation, robust monitoring, and proactive issue prevention. While DevSecOps engineering and CI/CD pipeline development remain important components of the ecosystem, this role places greater emphasis on application runtime health, operational workflows, incident management, release readiness, and environment reliability.
You will be involved in a wide range of application-centric operational tasks, including monitoring application health, performance optimization, managing incidents and escalations, coordinating releases, and ensuring proper alerting, logging, and observability. Your foundational understanding of application behavior, platform dependencies, and production operations will be critical as you help implement operational best practices and optimize runbook-driven workflows. You will also collaborate with senior engineers and architects to maintain strong application performance and platform resiliency.
As a Lead AppOps Engineer, you will be expected to actively participate in operational reviews, root cause analysis, change management, and continuous improvement of application support processes. You will work on initiatives that require attention to detail, strong analytical thinking, and a commitment to operational quality. Your ability to collaborate effectively, communicate clearly with cross-functional teams, and partner with stakeholders will be essential to delivering reliable and high-quality application services.
This position offers a fantastic opportunity for growth and career development within a supportive, modern, and innovative environment. You will have the chance to work with enterprise-scale applications, modern observability tools, and cloud platforms while supporting a high-performing team. If you are passionate about application operations, reliability, and service excellence, we encourage you to apply and join our team at Wolters Kluwer.
ESSENTIAL DUTIES AND RESPONSIBILITIES
- Own production reliability for critical applications; define and track SLOs, error budgets, and capacity/performance baselines.
- Lead major incident response, drive clear business/technical communications, and ensure data-driven root cause analysis with preventative actions.
- Direct release and change operations: assess risk, enforce readiness gates, validate post-deployment health, and improve change success rate.
- Architect operational observability: design dashboards, alert strategies, log/trace pipelines, and runbook automation for rapid diagnosis and recovery.
- Establish and continuously improve operational standards, guardrails, and runbooks; automate repetitive tasks to reduce toil.
- Partner with engineering on resiliency patterns (circuit breakers, bulkheads, graceful degradation, retries) and performance tuning.
- Plan and execute capacity management, scaling strategies, and DR/BCP readiness, including failover testing and scenario exercises.
- Champion security-by-default in operations: secrets hygiene, patch/vulnerability remediation, certificate/DNS management, least-privilege access.
- Mentor AppOps engineers; provide technical guidance, code/review for automation, and develop on-call excellence.
- Drive service reviews with stakeholders; publish operational KPIs (MTTR, change success rate, incident rate) and lead continuous improvement roadmaps.
- Application Runtime Management: Monitor application health, availability, and performance across environments; proactively identify issues and optimize application behavior.
- Incident & Problem Management: Triage, investigate, and resolve production incidents; participate in root cause analysis and drive long-term fixes.
- Release & Deployment Operations: Coordinate and execute application deployments, ensure release readiness, validate post-deployment health, and collaborate with engineering teams for smooth rollouts.
- Environment & Configuration Management: Maintain application environments, configuration baselines, secrets, access controls, and platform dependencies, ensuring consistency and compliance.
- Monitoring, Logging & Observability: Implement and maintain dashboards, alerts, and log pipelines using enterprise observability tools to ensure system transparency and rapid diagnosis.
- Operational Automation: Develop and enhance runbooks, automate repeatable workflows, reduce manual toil, and improve operational efficiency.
- SLA, SLO, and Reliability Improvements: Track key reliability metrics, enforce operational standards, and drive continuous optimization to meet or exceed service commitments.
- Change Management: Support change reviews, evaluate operational risks, ensure compliance with WK change processes, and validate operational readiness for all changes.
- Security & Compliance Alignment: Ensure adherence to security standards, support vulnerability remediation efforts, and maintain compliance with organizational policies.
- CrossFunctional Collaboration: Partner with Engineering, CloudOps, Security, Compliance, and other teams to resolve issues, improve service quality, and enhance application resilience.
OTHER DUTIES
- Performs other duties as assigned by management.
- On call rotation responsibilities with the Service Delivery and Operations Team
Requirements
Technical Requirements
- Advanced expertise in operating applications on Azure and/or AWS, including networking, load balancers, DNS, certificates, storage, and messaging services. Strong knowledge of application operations in cloud environments (Azure/AWS).
- Hands-on with observability stacks (Datadog, Grafana/Prometheus, ELK/OpenSearch, Open Telemetry) and alert engineering.
- Experience with incident management, RCA, and operational troubleshooting.
- Strong practical understanding of CI/CD concepts and collaboration with release teams; experience validating releases in lower/production environments.
- Familiarity with infrastructure components: load balancers, storage networking, DNS and certificates.
- Proficiency in automation and scripting (PowerShell, Bash, Python) to build runbooks, health checks, and remediation workflows.
- Experience with deployment strategies (blue/green, rolling, canary) and traffic management.
- Security and compliance in operations: vulnerability remediation, secrets and key management, audit readiness.
- Ability to interpret logs, metrics, traces, and performance data.
- Experience managing multi-environment application lifecycles (Dev, QA, UAT, Prod).
- Infrastructure as Code (IaC): Terraform (modules, workspaces), Azure ARM/Bicep or AWS CloudFormation; policy-as-code and environment drift detection.
Functional Requirements
- Ability to ensure 24x7 application reliability and operational excellence.
- Manage end-to-end application lifecycle including deployments, configurations, and environment health.
- Collaborate with engineering, CloudOps, and Security teams to ensure smooth operations.
- Own operational KPIs such as uptime, MTTR, change success rate, and SLA/SLO adherence.
- Perform release coordination, deployment validation, and post-release monitoring.
- Lead incident response, communication, and escalation handling.
- Participate in change management and risk assessments for all application changes.
- Maintain runbooks, SOPs, and operational documentation.
- Drive continuous improvement for operational workflows and process maturity.
- Support audit, compliance, and security requirements for applications.
Job Experience
8-10 Years
Qualifications
- Bachelor's degree in computer science, Information Systems, or a related field.
- Vendor certifications preferred: Azure Administrator/Architect or AWS SysOps/DevOps Professional; ITIL Foundation (or higher).
- Terraform Associate/Professional (or equivalent IaC certification) preferred; SRE Foundation a plus.
- Proven experience leading incident response, conducting RCAs, and implementing preventative controls.
- Excellent communication, stakeholder management, and mentoring skills in global, fast-paced environments.
- Strong understanding of Software Engineering Principals
- Industry recognized Kubernetes Certification.
Thought Leadership and Soft skills:
- Strong ownership mindset with a bias for automation, measurement, and continuous improvement.
- Ability to translate technical risks and trade-offs into business language for decision-makers.
#LI-Hybrid
Our Interview PracticesTo maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process.
Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process.
Compensation:
$118,300.00 - $207,400.00 USDThis role is eligible for Bonus.Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process.
Additional Information:Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
What Wolters Kluwer employees say
Pay
Benefits
Hours and flexibility
Workplace
Get the full story on Breakroom
About Wolters Kluwer
Sourced by ZipRecruiter
Wolters Kluwer Global Business Services is designed to provide services to the business units in the areas of technology, sourcing, procurement, legal, finance, and accounting which includes our North American-Accounting Center. These global centers promote team collaboration using best practices around a specific focus area to drive results and enhance operational efficiencies. There is a constant endeavor to benchmark against best-in-class industry standards to improve the quality of deliverables, increase cost savings, enhance productivity, and reduce time to market for products and applications
Industry
Accounting services, library and information services and it services
Company size
10,000+ Employees
Headquarters location
Philadelphia, PA, US