Architect operational observability: design dashboards, alert strategies, log/trace pipelines, and runbook automation for rapid diagnosis and recovery. * Establish and continuously improve ...
Architect operational observability: design dashboards, alert strategies, log/trace pipelines, and runbook automation for rapid diagnosis and recovery. * Establish and continuously improve ...
Architect operational observability: design dashboards, alert strategies, log/trace pipelines, and runbook automation for rapid diagnosis and recovery. * Establish and continuously improve ...
Architect operational observability: design dashboards, alert strategies, log/trace pipelines, and runbook automation for rapid diagnosis and recovery. * Establish and continuously improve ...
Architect operational observability: design dashboards, alert strategies, log/trace pipelines, and runbook automation for rapid diagnosis and recovery. * Establish and continuously improve ...
Architect operational observability: design dashboards, alert strategies, log/trace pipelines, and runbook automation for rapid diagnosis and recovery. * Establish and continuously improve ...
The engineer will help improve platform availability, reliability, observability, and automation while supporting large-scale virtualization services and infrastructure operations. The role will also ...
Quick apply
The engineer will help improve platform availability, reliability, observability, and automation while supporting large-scale virtualization services and infrastructure operations. The role will also ...
The engineer will help improve platform availability, reliability, observability, and automation while supporting large-scale virtualization services and infrastructure operations. The role will also ...
The engineer will help improve platform availability, reliability, observability, and automation while supporting large-scale virtualization services and infrastructure operations. The role will also ...
Improve system observability through metrics, structured logging, dashboards, and alerting * Participate in code and design reviews with a strong emphasis on security, correctness, and failure modes
Improve system observability through metrics, structured logging, dashboards, and alerting * Participate in code and design reviews with a strong emphasis on security, correctness, and failure modes
Solution Test lead
Indianapolis, IN · On-site
Scope spans four major capabilities and platform components Core Capabilities & Responsibilities * SRE & DevOps integration, Monitoring and observability * Quality Engineering & Test Engineering ...
Solution Test lead
Indianapolis, IN · On-site
Scope spans four major capabilities and platform components Core Capabilities & Responsibilities * SRE & DevOps integration, Monitoring and observability * Quality Engineering & Test Engineering ...
Gateway observability, real-time usage monitoring, cost attribution by team and use case, and anomaly detection. Enterprise tool and MCP registry, the governed catalog of tools, APIs, and data ...
Gateway observability, real-time usage monitoring, cost attribution by team and use case, and anomaly detection. Enterprise tool and MCP registry, the governed catalog of tools, APIs, and data ...
New Relic in the United States is seeking a technology leader to guide and grow our global engineering team and advance the next generation of Application Performance Monitoring. You will lead with ...
New Relic in the United States is seeking a technology leader to guide and grow our global engineering team and advance the next generation of Application Performance Monitoring. You will lead with ...
IT Infrastructure Architect
Carmel, IN · On-site
$78/hr
Observability. * Server and network topology. Founded in 2010 and headquartered in the Washington, DC metro area, Cynet Systems Inc. is a leading staffing and recruiting powerhouse. Proudly ...
Quick apply
IT Infrastructure Architect
Carmel, IN · On-site
$78/hr
Observability. * Server and network topology. Founded in 2010 and headquartered in the Washington, DC metro area, Cynet Systems Inc. is a leading staffing and recruiting powerhouse. Proudly ...
AWS Data Engineer Databricks
Indianapolis, IN · On-site
$109K - $131K/yr
Exposure to data observability and lineage tools * Experience with monitoring, logging, and observability frameworks Highly Preferred Skills * Experience in Healthcare / Pharma domain * Exposure to ...
AWS Data Engineer Databricks
Indianapolis, IN · On-site
$109K - $131K/yr
Exposure to data observability and lineage tools * Experience with monitoring, logging, and observability frameworks Highly Preferred Skills * Experience in Healthcare / Pharma domain * Exposure to ...
Drive Artificial Intelligence operations, event driven automation, and observability initiatives. * Improve infrastructure availability and reliability through repeatable operational patterns.
Quick apply
Drive Artificial Intelligence operations, event driven automation, and observability initiatives. * Improve infrastructure availability and reliability through repeatable operational patterns.
Drive Artificial Intelligence operations, event driven automation, and observability initiatives. * Improve infrastructure availability and reliability through repeatable operational patterns.
Drive Artificial Intelligence operations, event driven automation, and observability initiatives. * Improve infrastructure availability and reliability through repeatable operational patterns.
Leads adoption of DevSecOps and SRE practices; improves CI/CD pipelines, observability, secure-by-default guardrails, and internal tooling to accelerate delivery and reduce operational load.
Leads adoption of DevSecOps and SRE practices; improves CI/CD pipelines, observability, secure-by-default guardrails, and internal tooling to accelerate delivery and reduce operational load.
IT - Lead Platform Engineer
Jasper, IN · On-site
$91K - $120K/yr
Establish platform-wide observability using: * Azure Monitor, Application Insights, Grafana * Define and track platform KPIs: * Availability * Deployment success rate * Recovery time * Engineering ...
IT - Lead Platform Engineer
Jasper, IN · On-site
$91K - $120K/yr
Establish platform-wide observability using: * Azure Monitor, Application Insights, Grafana * Define and track platform KPIs: * Availability * Deployment success rate * Recovery time * Engineering ...
Sourcing - Manager, Engineering - (Guidewire PolicyCenter Platform)
Indianapolis, IN · On-site
$159K - $165K/yr
Leads adoption of DevSecOps and SRE practices; improves CI/CD pipelines, observability, secure-by-default guardrails, and internal tooling to accelerate delivery and reduce operational load.
Sourcing - Manager, Engineering - (Guidewire PolicyCenter Platform)
Indianapolis, IN · On-site
$159K - $165K/yr
Leads adoption of DevSecOps and SRE practices; improves CI/CD pipelines, observability, secure-by-default guardrails, and internal tooling to accelerate delivery and reduce operational load.
Senior DevOps Engineer
$124K - $159K/yr
Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. * Ensure infrastructure and automation ...
Senior DevOps Engineer
$124K - $159K/yr
Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. * Ensure infrastructure and automation ...
Senior DevOps Engineer
$124K - $159K/yr
Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. * Ensure infrastructure and automation ...
Senior DevOps Engineer
$124K - $159K/yr
Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. * Ensure infrastructure and automation ...
Senior DevOps Engineer
$124K - $159K/yr
Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. * Ensure infrastructure and automation ...
Senior DevOps Engineer
$124K - $159K/yr
Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. * Ensure infrastructure and automation ...
Linux Platform Operations Engineer
Indianapolis, IN · On-site
$66K - $89K/yr
You will have groundbreaking chances to transform our operations processes using proactive, predictive, and automated AI & Observability capabilities. * Be Your Best: You will bring a high learning ...
Linux Platform Operations Engineer
Indianapolis, IN · On-site
$66K - $89K/yr
You will have groundbreaking chances to transform our operations processes using proactive, predictive, and automated AI & Observability capabilities. * Be Your Best: You will bring a high learning ...
Observability information
See Indiana salary details
$15.55 - $21.61
0% of jobs
$21.61 - $27.66
0% of jobs
$27.66 - $33.71
2% of jobs
$33.71 - $39.76
5% of jobs
$39.76 - $45.81
10% of jobs
$48.65 is the 25th percentile. Wages below this are outliers.
$45.81 - $51.86
17% of jobs
The median wage is $56.64 / hr.
$51.86 - $57.91
20% of jobs
$57.91 - $63.96
18% of jobs
$65.05 is the 75th percentile. Wages above this are outliers.
$63.96 - $70.02
15% of jobs
$70.02 - $76.07
9% of jobs
$76.07 - $82.12
4% of jobs
$15
$57
$82
How much do observability jobs pay per hour?
What are the key skills and qualifications needed to thrive in an observability role?
To thrive in an Observability role, you need a strong background in monitoring, alerting, logging, and analyzing system performance, often supported by a degree in computer science or related field. Familiarity with tools such as Prometheus, Grafana, Datadog, Splunk, and experience with cloud platforms and scripting languages is crucial. Excellent problem-solving, communication, and collaboration skills help you work effectively with cross-functional engineering and operations teams. These capabilities are essential to ensure system reliability, quickly detect issues, and maintain seamless digital experiences.
What is an observability?
An Observability job focuses on ensuring the performance, reliability, and health of software systems by collecting, analyzing, and visualizing telemetry data such as logs, metrics, and traces. Professionals in this field work with monitoring tools, distributed tracing, and alerting systems to detect and troubleshoot issues proactively. They collaborate with engineering and operations teams to improve system visibility, reduce downtime, and enhance overall system performance.
Is observability a good career?
What does an observability do?
In an Observability role, your daily tasks often include designing and maintaining monitoring dashboards, configuring alerts, analyzing system logs, and working closely with development and operations teams to troubleshoot issues. You'll proactively identify areas of improvement to increase system reliability, document monitoring strategies, and support incident response efforts. Collaboration is key, as you may participate in post-incident reviews and help drive architectural improvements based on the data you collect. The role is dynamic and requires a proactive approach to ensure systems stay healthy and downtime is minimized.

Full-time
Medical, Dental, Vision, Retirement, PTO
Posted 14 days ago
Wolters Kluwer rating
9.0
Based on 28 frontline employees who took The Breakroom Quiz
33rd of 241 rated software companies
Job description
Wolters Kluwer is seeking a motivated and talented Lead AppOps Engineer to join our dynamic team. This role is ideal for an individual with hands-on experience in application operations, production support, and cloud-platform application management within a mature engineering environment. As a Lead AppOps Engineer, you will play a crucial role in ensuring the stability, reliability, availability, and operational excellence of our enterprise applications.
In this role, your primary responsibility will be managing the end-to-end operational lifecycle of critical applications. You will work closely with engineering teams, Cloud Operations, Security, Compliance, and other key partners to ensure smooth application deployments, reliable production operation, robust monitoring, and proactive issue prevention. While DevSecOps engineering and CI/CD pipeline development remain important components of the ecosystem, this role places greater emphasis on application runtime health, operational workflows, incident management, release readiness, and environment reliability.
You will be involved in a wide range of application-centric operational tasks, including monitoring application health, performance optimization, managing incidents and escalations, coordinating releases, and ensuring proper alerting, logging, and observability. Your foundational understanding of application behavior, platform dependencies, and production operations will be critical as you help implement operational best practices and optimize runbook-driven workflows. You will also collaborate with senior engineers and architects to maintain strong application performance and platform resiliency.
As a Lead AppOps Engineer, you will be expected to actively participate in operational reviews, root cause analysis, change management, and continuous improvement of application support processes. You will work on initiatives that require attention to detail, strong analytical thinking, and a commitment to operational quality. Your ability to collaborate effectively, communicate clearly with cross-functional teams, and partner with stakeholders will be essential to delivering reliable and high-quality application services.
This position offers a fantastic opportunity for growth and career development within a supportive, modern, and innovative environment. You will have the chance to work with enterprise-scale applications, modern observability tools, and cloud platforms while supporting a high-performing team. If you are passionate about application operations, reliability, and service excellence, we encourage you to apply and join our team at Wolters Kluwer.
ESSENTIAL DUTIES AND RESPONSIBILITIES
- Own production reliability for critical applications; define and track SLOs, error budgets, and capacity/performance baselines.
- Lead major incident response, drive clear business/technical communications, and ensure data-driven root cause analysis with preventative actions.
- Direct release and change operations: assess risk, enforce readiness gates, validate post-deployment health, and improve change success rate.
- Architect operational observability: design dashboards, alert strategies, log/trace pipelines, and runbook automation for rapid diagnosis and recovery.
- Establish and continuously improve operational standards, guardrails, and runbooks; automate repetitive tasks to reduce toil.
- Partner with engineering on resiliency patterns (circuit breakers, bulkheads, graceful degradation, retries) and performance tuning.
- Plan and execute capacity management, scaling strategies, and DR/BCP readiness, including failover testing and scenario exercises.
- Champion security-by-default in operations: secrets hygiene, patch/vulnerability remediation, certificate/DNS management, least-privilege access.
- Mentor AppOps engineers; provide technical guidance, code/review for automation, and develop on-call excellence.
- Drive service reviews with stakeholders; publish operational KPIs (MTTR, change success rate, incident rate) and lead continuous improvement roadmaps.
- Application Runtime Management: Monitor application health, availability, and performance across environments; proactively identify issues and optimize application behavior.
- Incident & Problem Management: Triage, investigate, and resolve production incidents; participate in root cause analysis and drive long-term fixes.
- Release & Deployment Operations: Coordinate and execute application deployments, ensure release readiness, validate post-deployment health, and collaborate with engineering teams for smooth rollouts.
- Environment & Configuration Management: Maintain application environments, configuration baselines, secrets, access controls, and platform dependencies, ensuring consistency and compliance.
- Monitoring, Logging & Observability: Implement and maintain dashboards, alerts, and log pipelines using enterprise observability tools to ensure system transparency and rapid diagnosis.
- Operational Automation: Develop and enhance runbooks, automate repeatable workflows, reduce manual toil, and improve operational efficiency.
- SLA, SLO, and Reliability Improvements: Track key reliability metrics, enforce operational standards, and drive continuous optimization to meet or exceed service commitments.
- Change Management: Support change reviews, evaluate operational risks, ensure compliance with WK change processes, and validate operational readiness for all changes.
- Security & Compliance Alignment: Ensure adherence to security standards, support vulnerability remediation efforts, and maintain compliance with organizational policies.
- CrossFunctional Collaboration: Partner with Engineering, CloudOps, Security, Compliance, and other teams to resolve issues, improve service quality, and enhance application resilience.
OTHER DUTIES
- Performs other duties as assigned by management.
- On call rotation responsibilities with the Service Delivery and Operations Team
Requirements
Technical Requirements
- Advanced expertise in operating applications on Azure and/or AWS, including networking, load balancers, DNS, certificates, storage, and messaging services. Strong knowledge of application operations in cloud environments (Azure/AWS).
- Hands-on with observability stacks (Datadog, Grafana/Prometheus, ELK/OpenSearch, Open Telemetry) and alert engineering.
- Experience with incident management, RCA, and operational troubleshooting.
- Strong practical understanding of CI/CD concepts and collaboration with release teams; experience validating releases in lower/production environments.
- Familiarity with infrastructure components: load balancers, storage networking, DNS and certificates.
- Proficiency in automation and scripting (PowerShell, Bash, Python) to build runbooks, health checks, and remediation workflows.
- Experience with deployment strategies (blue/green, rolling, canary) and traffic management.
- Security and compliance in operations: vulnerability remediation, secrets and key management, audit readiness.
- Ability to interpret logs, metrics, traces, and performance data.
- Experience managing multi-environment application lifecycles (Dev, QA, UAT, Prod).
- Infrastructure as Code (IaC): Terraform (modules, workspaces), Azure ARM/Bicep or AWS CloudFormation; policy-as-code and environment drift detection.
Functional Requirements
- Ability to ensure 24x7 application reliability and operational excellence.
- Manage end-to-end application lifecycle including deployments, configurations, and environment health.
- Collaborate with engineering, CloudOps, and Security teams to ensure smooth operations.
- Own operational KPIs such as uptime, MTTR, change success rate, and SLA/SLO adherence.
- Perform release coordination, deployment validation, and post-release monitoring.
- Lead incident response, communication, and escalation handling.
- Participate in change management and risk assessments for all application changes.
- Maintain runbooks, SOPs, and operational documentation.
- Drive continuous improvement for operational workflows and process maturity.
- Support audit, compliance, and security requirements for applications.
Job Experience
8-10 Years
Qualifications
- Bachelor's degree in computer science, Information Systems, or a related field.
- Vendor certifications preferred: Azure Administrator/Architect or AWS SysOps/DevOps Professional; ITIL Foundation (or higher).
- Terraform Associate/Professional (or equivalent IaC certification) preferred; SRE Foundation a plus.
- Proven experience leading incident response, conducting RCAs, and implementing preventative controls.
- Excellent communication, stakeholder management, and mentoring skills in global, fast-paced environments.
- Strong understanding of Software Engineering Principals
- Industry recognized Kubernetes Certification.
Thought Leadership and Soft skills:
- Strong ownership mindset with a bias for automation, measurement, and continuous improvement.
- Ability to translate technical risks and trade-offs into business language for decision-makers.
#LI-Hybrid
Our Interview PracticesTo maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process.
Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process.
Compensation:
$118,300.00 - $207,400.00 USDThis role is eligible for Bonus.Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process.
Additional Information:Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
What Wolters Kluwer employees say
Pay
Benefits
Hours and flexibility
Workplace
Get the full story on Breakroom