Key Responsibilities Event Correlation & Observability Engineering * Design, configure, and continuously improve event correlation rules and alerting strategies across platforms such as Splunk ITSI ...
Key Responsibilities Event Correlation & Observability Engineering * Design, configure, and continuously improve event correlation rules and alerting strategies across platforms such as Splunk ITSI ...
Key Responsibilities Event Correlation & Observability Engineering * Design, configure, and continuously improve event correlation rules and alerting strategies across platforms such as Splunk ITSI ...
Key Responsibilities Event Correlation & Observability Engineering * Design, configure, and continuously improve event correlation rules and alerting strategies across platforms such as Splunk ITSI ...
Key Responsibilities Event Correlation & Observability Engineering * Design, configure, and continuously improve event correlation rules and alerting strategies across platforms such as Splunk ITSI ...
Key Responsibilities Event Correlation & Observability Engineering * Design, configure, and continuously improve event correlation rules and alerting strategies across platforms such as Splunk ITSI ...
Champions observability and production excellence, ensuring today's features are built with ... Partner with DevOps and SRE to improve Kubernetes-based deployments and cloud infrastructure ...
Champions observability and production excellence, ensuring today's features are built with ... Partner with DevOps and SRE to improve Kubernetes-based deployments and cloud infrastructure ...
Champions observability and production excellence, ensuring today's features are built with ... Partner with DevOps and SRE to improve Kubernetes-based deployments and cloud infrastructure ...
Champions observability and production excellence, ensuring today's features are built with ... Partner with DevOps and SRE to improve Kubernetes-based deployments and cloud infrastructure ...
Senior Platform Engineer
Chattanooga, TN · On-site
$95K - $130K/yr
... observability platforms. • Champion Site Reliability Engineering (SRE) practices including incident response, root cause analysis, runbook development, service reliability, and operational ...
Senior Platform Engineer
Chattanooga, TN · On-site
$95K - $130K/yr
... observability platforms. • Champion Site Reliability Engineering (SRE) practices including incident response, root cause analysis, runbook development, service reliability, and operational ...
Systems Engineer - Cloud Ops
Memphis, TN · On-site
$54.25 - $72.50/hr
As a Systems Engineer on the Cloud Operations team, you will be responsible for deploying, managing ... Monitor system performance using observability tools (Dynatrace, Cloud Monitoring, Prometheus ...
Systems Engineer - Cloud Ops
Memphis, TN · On-site
$54.25 - $72.50/hr
As a Systems Engineer on the Cloud Operations team, you will be responsible for deploying, managing ... Monitor system performance using observability tools (Dynatrace, Cloud Monitoring, Prometheus ...
Software Developer 5
Nashville, TN · On-site
Join the team building Oracle Cloud Infrastructure's state of the art observability platform ... OCI Monitoring and Logging serve as foundational platforms used by OCI engineering teams to operate ...
Software Developer 5
Nashville, TN · On-site
Join the team building Oracle Cloud Infrastructure's state of the art observability platform ... OCI Monitoring and Logging serve as foundational platforms used by OCI engineering teams to operate ...
Senior Director, Software Engineering - Crypto
Nashville, TN · On-site
$244K/yr
Improve operational efficiency through automation, observability, engineering excellence, and cloud service efficiency initiatives. * Develop leaders and engineering talent across multiple ...
Senior Director, Software Engineering - Crypto
Nashville, TN · On-site
$244K/yr
Improve operational efficiency through automation, observability, engineering excellence, and cloud service efficiency initiatives. * Develop leaders and engineering talent across multiple ...
Principal Site Reliability Engineer
Nashville, TN · On-site
$55 - $73.25/hr
Observability & Operational Excellence * Design comprehensive monitoring, logging, tracing, and ... Drive engineering excellence through code reviews, design reviews, technical guidance, and ...
Principal Site Reliability Engineer
Nashville, TN · On-site
$55 - $73.25/hr
Observability & Operational Excellence * Design comprehensive monitoring, logging, tracing, and ... Drive engineering excellence through code reviews, design reviews, technical guidance, and ...
Senior Cloud DevOps Engineer
Franklin, TN · Remote
$125K - $161K/yr
EC2) while optimizing for performance, cost, and observability. * Cloud Migrations: Execute ... Engineer Professionalcertification, oractively pursuing. Preferred Qualifications * Working ...
Senior Cloud DevOps Engineer
Franklin, TN · Remote
$125K - $161K/yr
EC2) while optimizing for performance, cost, and observability. * Cloud Migrations: Execute ... Engineer Professionalcertification, oractively pursuing. Preferred Qualifications * Working ...
Uphold production standards, lead the design of observability, performance and resilience testing ... Engineering Craftsmanship: Maintain accountability for the operational integrity of production and ...
Uphold production standards, lead the design of observability, performance and resilience testing ... Engineering Craftsmanship: Maintain accountability for the operational integrity of production and ...
Applied AI SRE III - PxE GPS
$50 - $66.50/hr
Uphold production standards, lead the design of observability, performance and resilience testing ... Engineering Craftsmanship: Maintain accountability for the operational integrity of production and ...
New
Applied AI SRE III - PxE GPS
$50 - $66.50/hr
Uphold production standards, lead the design of observability, performance and resilience testing ... Engineering Craftsmanship: Maintain accountability for the operational integrity of production and ...
New
Lead Engineer AI Agent Platform - Nashville, TN - 100% Onsite (5 Days Onsite/week) - Contract - CTS
Nashville, TN · On-site
$99K - $130K/yr
Lead Engineer AI Agent Platform Job Location: Nashville, TN - 100% Onsite (5 Days Onsite/week) Job ... Evaluation, Observability & Testing (Must-Have):- Hands-on experience with LLM evaluation ...
Lead Engineer AI Agent Platform - Nashville, TN - 100% Onsite (5 Days Onsite/week) - Contract - CTS
Nashville, TN · On-site
$99K - $130K/yr
Lead Engineer AI Agent Platform Job Location: Nashville, TN - 100% Onsite (5 Days Onsite/week) Job ... Evaluation, Observability & Testing (Must-Have):- Hands-on experience with LLM evaluation ...
Set production standards, lead the design of observability, performance and resilience testing, and ... Engineering Craftsmanship: Maintain accountability for the operational integrity of production and ...
New
Set production standards, lead the design of observability, performance and resilience testing, and ... Engineering Craftsmanship: Maintain accountability for the operational integrity of production and ...
New
Set production standards, lead the design of observability, performance and resilience testing, and ... Engineering Craftsmanship: Maintain accountability for the operational integrity of production and ...
New
Set production standards, lead the design of observability, performance and resilience testing, and ... Engineering Craftsmanship: Maintain accountability for the operational integrity of production and ...
New
Software Engineer, Backend
Knoxville, TN · On-site
$155K - $222K/yr
Programming experience in Rust, Scala, Python ... Experience with observability tooling such asOpenTelemetry, Prometheus, Grafana. * Experience with ...
Software Engineer, Backend
Knoxville, TN · On-site
$155K - $222K/yr
Programming experience in Rust, Scala, Python ... Experience with observability tooling such asOpenTelemetry, Prometheus, Grafana. * Experience with ...
Set production standards, lead the design of observability, performance and resilience testing, and ... Engineering Craftsmanship: Maintain accountability for the operational integrity of production and ...
New
Set production standards, lead the design of observability, performance and resilience testing, and ... Engineering Craftsmanship: Maintain accountability for the operational integrity of production and ...
New
Software Developer 5
Nashville, TN · On-site
Join the team building Oracle Cloud Infrastructure's state of the art observability platform ... OCI Monitoring and Logging serve as foundational platforms used by OCI engineering teams to operate ...
Software Developer 5
Nashville, TN · On-site
Join the team building Oracle Cloud Infrastructure's state of the art observability platform ... OCI Monitoring and Logging serve as foundational platforms used by OCI engineering teams to operate ...
Software Engineer, Backend
$155K - $222K/yr
Programming experience in Rust, Scala, Python ... Experience with observability tooling such asOpenTelemetry, Prometheus, Grafana. * Experience with ...
Software Engineer, Backend
$155K - $222K/yr
Programming experience in Rust, Scala, Python ... Experience with observability tooling such asOpenTelemetry, Prometheus, Grafana. * Experience with ...
Observability Engineer information
See Tennessee salary details
$22.56 - $30.33
2% of jobs
$30.33 - $38.09
4% of jobs
$38.09 - $45.86
6% of jobs
$45.86 - $53.62
8% of jobs
$54.95 is the 25th percentile. Wages below this are outliers.
$53.62 - $61.39
23% of jobs
The median wage is $65.27 / hr.
$61.39 - $69.15
12% of jobs
$73.87 is the 75th percentile. Wages above this are outliers.
$69.15 - $76.91
32% of jobs
$76.91 - $84.68
11% of jobs
$84.68 - $92.44
1% of jobs
$92.44 - $100.21
0% of jobs
$100.21 - $107.97
1% of jobs
$22
$65
$107
How much do observability engineer jobs pay per hour?
What are the typical responsibilities of an Observability Engineer on a daily basis?
As an Observability Engineer, your daily responsibilities usually include designing and maintaining monitoring and alerting systems, analyzing system logs, and collaborating with development teams to improve service reliability. You'll work regularly with a suite of observability tools to ensure that infrastructure and applications are performing optimally, quickly responding to any incidents or anomalies detected. Additionally, you help establish best practices for instrumentation and metrics collection, often leading initiatives to enhance visibility into complex, distributed systems. The work is a blend of proactive system design and reactive problem-solving, requiring both technical expertise and strong teamwork skills.
What engineers make $500,000 a year?
What is the salary of observability engineer?
What does an observability engineer do?
What does an Observability Engineer do?
An Observability Engineer is responsible for designing, implementing, and maintaining monitoring, logging, and tracing systems to ensure the health, performance, and reliability of applications and infrastructure. They work with tools like Prometheus, Grafana, OpenTelemetry, and ELK to collect and analyze telemetry data. Their goal is to provide visibility into system behavior, detect and diagnose issues quickly, and improve overall system observability. They collaborate with developers, SREs, and operational teams to create automated and scalable observability solutions.
What engineers make 300,000 a year?
What are the key skills and qualifications needed to thrive in the Observability Engineer position, and why are they important?
To thrive as an Observability Engineer, you need a solid understanding of monitoring, logging, and tracing systems, as well as expertise in programming, cloud infrastructure, and incident response. Proficiency in tools like Prometheus, Grafana, ELK stack, and familiarity with cloud platforms such as AWS, Azure, or GCP are commonly required, and certifications like AWS Certified DevOps Engineer can be advantageous. Strong analytical thinking, collaborative skills, and effective communication are essential soft skills for diagnosing issues and working across development and operations teams. These competencies are vital for proactively maintaining system reliability, ensuring performance, and resolving complications before they impact business operations.

Full-time
Medical, Dental, Vision, Life, Retirement, PTO
Re-posted 19 days ago
Humana rating
8.0
Based on 263 frontline employees who took The Breakroom Quiz
161st of 300 rated insurance
Job description
The Senior Problem, Incident, and Event Management Engineer is responsible for advancing enterprise event correlation, observability, and incident detection capabilities to proactively identify and mitigate service disruptions before user impact. This role specializes in leveraging platforms such as Splunk, Dynatrace, and ServiceNow to aggregate, correlate, and operationalize data across multiple systems to drive timely escalation and resolution of critical incidents.
This position plays a key role in evolving from reactive incident response to data-driven, predictive operations, utilizing CMDB-driven context, criticality tiering, and advanced analytics to prioritize and escalate issues with significant business impact.
- Design, configure, and continuously improve event correlation rules and alerting strategies across platforms such as Splunk ITSI and Dynatrace
- Integrate data from multiple monitoring, application, and infrastructure sources to create meaningful, actionable events
- Normalize and enrich event data using standardized fields and metadata to improve correlation accuracy and reduce noise
- Drive reduction of false positives and duplicate alerts through correlation, aggregation, and suppression strategies
- Develop and maintain operational and executive dashboards in Splunk and other reporting tools
- Translate technical telemetry into clear, business-aligned insights, highlighting service health, degradation, and emerging risks
- Partner with command center, TOC, and incident teams to ensure dashboards support real-time decision making and escalation
- Leverage correlated event data and observability insights to trigger proactive incident identification prior to user-reported impact
- Apply criticality tiering and CMDB data to assess business impact and drive proper prioritization and escalation paths
- Partner with ServiceNow stakeholders to improve workflows, reporting, and automation capabilities
- Leverage CMDB relationships and service mapping where available to enrich event data with application, infrastructure, and business context
- Utilize service ownership, business criticality, and operational hours data to inform prioritization decisions
- Partner with CMDB and service mapping teams to improve data quality and completeness
- Analyze patterns across incidents, alerts, and events to identify systemic issues and opportunities for improvement
- Partner with Problem Management to eliminate recurring issues through structural fixes
- Drive improvements in monitoring coverage, alert quality, and detection speed
- Contribute to a shift toward predictive, AIOps-driven operations
Use your skills to make an impact
Required Qualifications
- 3-5+ years of experience in Incident, Event, or Problem Management
- Hands-on experience with Splunk (preferably ITSI) and Dynatrace or similar observability platforms
- Experience building dashboards, reports, and analytics to support operational decision-making
- Experience with ServiceNow ITSM, including incident lifecycle management and reporting
- Strong analytical skills with the ability to correlate data across multiple systems and platforms
- Experience working with event correlation, alerting strategies, or AIOps concepts
- Ability to assess business impact using priority models, criticality tiers, and service context
- Strong communication skills with the ability to translate technical findings into actionable insights
Preferred Qualifications
- Experience with CMDB, service mapping, or application dependency mapping
- Exposure to enterprise monitoring ecosystems (e.g., APM, synthetic monitoring, infrastructure monitoring)
- Experience supporting command center, TOC, or major incident management environments
- Knowledge of ITIL frameworks and service management best practices
- Experience with automation or scripting (Python, PowerShell, or similar)
- Bachelor's Degree in Business, Computer Science, or a related field or equal experience
- ITIL v5 certification
- Previous experience in the health care industry
Additional Information:
Limited Geography Remote - This is a remote position but located within a specific geography.
To ensure Home or Hybrid Home/Office employees' ability to work effectively, the self-provided internet service of Home or Hybrid Home/Office employees must meet the following criteria:
At minimum, a download speed of 25 Mbps and an upload speed of 10 Mbps is required; wireless, wired cable or DSL connection is suggested.
Satellite, cellular and microwave connection can be used only if approved by leadership.
Employees who live and work from Home in the state of California, Illinois, Montana, or South Dakota will be provided a bi-weekly payment for their internet expense.
Humana will provide Home or Hybrid Home/Office employees with telephone equipment appropriate to meet the business requirements for their position/job.
Work from a dedicated space lacking ongoing interruptions to protect member PHI / HIPAA information.
Scheduled Weekly Hours
40Pay Range
The compensation range below reflects a good faith estimate of starting base pay for full time (40 hours per week) employment at the time of posting. The pay range may be higher or lower based on geographic location and individual pay will vary based on demonstrated job related skills, knowledge, experience, education, certifications, etc.Description of Benefits
Humana, Inc. and its affiliated subsidiaries (collectively, "Humana") offers competitive benefits that support whole-person well-being. Associate benefits are designed to encourage personal wellness and smart healthcare decisions for you and your family while also knowing your life extends outside of work. Among our benefits, Humana provides medical, dental and vision benefits, 401(k) retirement savings plan, time off (including paid time off, company and personal holidays, paid parental and caregiver leave), short-term and long-term disability, life insurance and many other opportunities.About us
About Humana: Humana Inc. (NYSE: HUM) is a leading U.S. healthcare company. Through our Humana insurance services and our CenterWell healthcare services, we make it easier for the millions of people we serve to achieve their best health - delivering the care and service they need, when they need it. These efforts are leading to a better quality of life for people with Medicare and Medicaid, families, individuals, military service personnel, and communities at large. Learn more about what we offer atHumana.comand atCenterWell.com.
Equal Opportunity Employer
It is the policy of Humana not to discriminate against any employee or applicant for employment because of race, color, religion, sex, sexual orientation, gender identity, national origin, age, marital status, genetic information, disability or protected veteran status. It is also the policy of Humana to take affirmative action, in compliance with Section 503 of the Rehabilitation Act and VEVRAA, to employ and to advance in employment individuals with disability or protected veteran status, and to base all employment decisions only on valid job requirements. This policy shall apply to all employment actions, including but not limited to recruitment, hiring, upgrading, promotion, transfer, demotion, layoff, recall, termination, rates of pay or other forms of compensation and selection for training, including apprenticeship, at all levels of employment.
About Humana
Sourced by ZipRecruiter
Humana Inc., headquartered in Louisville, KY., is a leading health care company that offers a wide range of insurance products and health and wellness services that incorporate an integrated approach to lifelong well-being. By leveraging the strengths of its core businesses, Humana believes it can better explore opportunities for existing and emerging adjacencies in health care that can further enhance wellness opportunities for the millions of people across the nation with whom the company has relationships.
Industry
Health care and social assistance
Company size
10,000+ Employees
Headquarters location
Louisville, KY, US
Year founded
1961