1

Observability Jobs in Pennsylvania (NOW HIRING)

Build and operate observability infrastructure across the three pillars - metrics, logging, and tracing - including dashboards, actionable alerting, and health/performance monitoring * Work to ...

Product Management - Dynatrace, AVP

Berwyn, PA · On-site

$223K - $234K/yr

This role will define the vision, roadmap, and success metrics for observability across infrastructure and applications, ensuring alignment with business priorities while collaborating with ...

Contribute to the design, implementation, and support of enterprise monitoring and observability solutions across cloud and on prem environments. * Assist with CI/CD pipelines for automated ...

In Cloud Observability, we care about making Google Cloud Platform the most observable cloud -- from ensuring that all GCP services have complete telemetry (logs, metrics, and traces), to building ...

DataOps & Reliability Engineer

Erie, PA · On-site

$99K - $125K/yr

The role focuses on pipeline automation, monitoring and observability, incident management, release management, and platform reliability to ensure stable and efficient operations. Key ...

Implement AI security, observability, and human oversight controls * Ensure compliance with NIST AI RMF, ISO 42001, GDPR, HIPAA, SOC2, and related regulations * Support enterprise AI transformation ...

next page

Showing results 1-20

Observability information

See Pennsylvania salary details

$16

$60

$86

How much do observability jobs pay per hour?

As of Aug 21, 2026, the average hourly pay for observability in Pennsylvania is $60.68, according to ZipRecruiter salary data. Most workers in this role earn between $50.87 and $69.62 per hour, depending on experience, location, and employer.

What is an observability?

An Observability job focuses on ensuring the performance, reliability, and health of software systems by collecting, analyzing, and visualizing telemetry data such as logs, metrics, and traces. Professionals in this field work with monitoring tools, distributed tracing, and alerting systems to detect and troubleshoot issues proactively. They collaborate with engineering and operations teams to improve system visibility, reduce downtime, and enhance overall system performance.

What does an observability do?

In an Observability role, your daily tasks often include designing and maintaining monitoring dashboards, configuring alerts, analyzing system logs, and working closely with development and operations teams to troubleshoot issues. You'll proactively identify areas of improvement to increase system reliability, document monitoring strategies, and support incident response efforts. Collaboration is key, as you may participate in post-incident reviews and help drive architectural improvements based on the data you collect. The role is dynamic and requires a proactive approach to ensure systems stay healthy and downtime is minimized.

What are the key skills and qualifications needed to thrive in an observability role?

To thrive in an Observability role, you need a strong background in monitoring, alerting, logging, and analyzing system performance, often supported by a degree in computer science or related field. Familiarity with tools such as Prometheus, Grafana, Datadog, Splunk, and experience with cloud platforms and scripting languages is crucial. Excellent problem-solving, communication, and collaboration skills help you work effectively with cross-functional engineering and operations teams. These capabilities are essential to ensure system reliability, quickly detect issues, and maintain seamless digital experiences.

Is observability a good career?

Observability is a growing field within IT and software engineering that involves monitoring, logging, and analyzing system performance using tools like Prometheus, Grafana, and Elasticsearch. It offers opportunities for specialization, high demand for skills, and roles in DevOps and site reliability engineering, making it a viable career choice for those interested in system reliability and automation.

What are the most commonly searched types of Observability jobs in Pennsylvania?

The most popular types of Observability jobs in Pennsylvania are:

What are popular job titles related to Observability jobs in Pennsylvania?

For Observability jobs in Pennsylvania, the most frequently searched job titles are:

What job categories do people searching Observability jobs in Pennsylvania look for?

The top searched job categories for Observability jobs in Pennsylvania are:

Infographic showing various Observability job openings in Pennsylvania as of August 2026, with employment types broken down into 9% Internship, 73% Full Time, and 18% Contract. Highlights an 82% In-person, and 18% Remote job distribution, with an average salary of $126,211 per year, or $60.7 per hour.

Monitoring and Observability Engineer

Veterans Sourcing Group, LLC

Pittsburgh, PA • On-site

Full-time

Re-posted 17 hours ago


Job description

Job Summary:
Veterans Sourcing Group, LLC is seeking a skilled Cloud Monitoring and Observability Engineer to design, implement, and optimize end-to-end monitoring and observability solutions for mission-critical applications in the Azure environment. The role involves collaborating with various teams to ensure comprehensive monitoring coverage and maintaining enterprise monitoring platforms to deliver actionable insights.
Responsibilities:
• Seeking a skilled Cloud Monitoring and Observability Engineer (Azure) engineer to design, implement, and optimize end-to-end monitoring and observability solutions for a mission-critical application deployed in the Azure environment.
• The ideal candidate has hands-on experience with enterprise monitoring tools—such as AppDynamics, Thousand Eyes, NetScout, and SolarWinds (or equivalent alternatives)—and a strong background in building scalable, secure, and compliant observability stacks for cloud deployments.
• Will collaborate closely with application engineering, cloud platform, network, and security teams to ensure comprehensive coverage across application, infrastructure, and network layers
• Design and implement end-to-end monitoring, alerting, and observability for an Azure-hosted application across application, infrastructure, network, and user experience layers.
• Configure, integrate, and maintain enterprise monitoring platforms to deliver actionable telemetry, performance baselines, and SLA/SLO tracking.
• Build dashboards, health checks, synthetic tests, and alerting workflows; optimize alert fidelity to minimize noise and improve signal-to-noise ratio.
• Establish and document telemetry standards (metrics, logs, traces), data collection strategies, and service-level indicators (SLIs) aligned to reliability objectives (SLOs).
• Integrate Azure-native services (Azure Monitor, Log Analytics, Application Insights) with enterprise tools to provide unified visibility and correlation.
• Implement network performance monitoring, path visibility, and internet/extranet testing using NPM tools (e.g., ThousandEyes, NetScout); leverage infrastructure monitoring platforms (e.g., SolarWinds) for device and service health.
• Instrument applications with APM tools (e.g., AppDynamics, Dynatrace, New Relic) for business transaction monitoring, dependency mapping, and root-cause analysis; tune anomaly detection and policy thresholds.
• Collaborate with DevOps/SRE teams to embed monitoring into CI/CD and infrastructure-as-code patterns; ensure new services adhere to observability standards.
• Define runbooks and escalation paths; support incident response and post-incident reviews with data-driven insights and remediation recommendations.
• Ensure monitoring solutions meet applicable security and compliance requirements; support audit requests with clear documentation and evidence.
• Conduct capacity and performance trend analysis; recommend optimization, right-sizing, and resilience improvements.
• Provide knowledge transfer, documentation, and training on monitoring tools, best practices, and operational workflows.
Qualifications:
Required:
• 5+ years implementing enterprise monitoring/observability for cloud or hybrid environments, including mission-critical applications.
• Demonstrable expertise with at least one tool in each category (or equivalent), including production deployments, advanced configuration, and operational use: Application Performance Monitoring (APM): AppDynamics, Dynatrace, or New Relic.
• Experience instrumenting services for business transaction tracing, code-level diagnostics, service maps, and anomaly detection.
• Ability to design APM dashboards and create alert policies with appropriate thresholds and baselines.
• Network Performance Monitoring (NPM) / Digital Experience Monitoring (DEM): Thousand Eyes, NetScout, or Kentik.
• Experience with synthetic tests, path visualization, packet-level analysis, and internet/WAN performance monitoring.
• Ability to configure endpoint agents, BGP/DNS tests, and multi-hop path monitoring for user experience correlation.
• Infrastructure Monitoring and Event Management: SolarWinds, Microsoft SCOM, Datadog, or Prometheus/Grafan.
• Experience monitoring servers, containers, network devices, and cloud services; creating availability and capacity dashboards.
• Proficiency with alert routing, de-duplication, and event correlation.
• Strong Azure monitoring experience: Azure Monitor, Log Analytics (KQL), Application Insights, and integration with third-party tools.
• Solid understanding of distributed tracing, metrics, and log aggregation; familiarity with Open Telemetry concepts and data pipelines.
• Scripting/automation skills (PowerShell, Python, or Bash) to automate monitoring configuration, agent deployment, test creation, and reporting.
• Networking fundamentals (DNS, BGP, HTTP, TLS, TCP/IP), CDN concepts, and WAN performance monitoring; ability to correlate app and network telemetry.
• Experience supporting incident response and performance troubleshooting across applications, infrastructure, and network layers.
• Excellent documentation and communication skills; collaborative mindset with engineering, operations, and security stakeholders.
Preferred:
• Background in regulated environments (financial services, government, healthcare) with compliance-aware monitoring design.
• Experience with log aggregation and SIEM/SOAR platforms (e.g., Splunk, Elastic) and integration with APM/NPM tools.
• Integration experience with ITSM platforms (e.g., ServiceNow) for incident, change, and problem management workflows.
• Familiarity with infrastructure-as-code (ARM/Bicep/Terraform) and embedding observability into IaC patterns; experience with CI/CD integration.
• Exposure to SRE practices (SLIs/SLOs, error budgets, reliability reviews) and capacity/performance planning.
• Ability to code in one or more of the following languages for instrumentation, custom telemetry, SDK integration, and tooling automation: Java: Implementing Open Telemetry SDKs/agents, custom instrumentation, and APM tagging; building synthetic test harnesses. .NET (C#): Instrumenting ASP.NET services, configuring APM auto-instrumentation, writing custom exporters and health probes. Python: Building automation scripts, collectors/exporters, synthetic tests, and integrating with monitoring APIs and SDKs.
Company:
Welcome to the Veterans Souring Group company profile. Veterans Sourcing Group (VSG) is a “Service Disabled Veteran Owned Small Business – SDVOSB”. Founded in , the company is headquartered in New City, USA, with a team of 11-50 employees. The company is currently Early Stage.

Veterans Sourcing Group logo

About Veterans Sourcing Group

Sourced by ZipRecruiter

Veterans Sourcing Group is a renowned company headquartered in New York, NY, US and is committed to providing high-quality, reliable staffing solutions and advisory services. The company operates in the human resources and staffing industry, specializing in veteran hiring. They offer various solutions to meet client needs, including strategic consultancy, professional search, and contract staffing. The company was founded by Beth Vines and Bruce Madnick, respected professionals who recognized a gap in the market for veteran-focused staffing services which prompted them to establish Veterans Sourcing Group in 2011.

Industry

Recruiting and staffing services

Company size

51 - 200 Employees

Headquarters location

New York, NY, US

Social media