Responsibilities : • Seeking a skilled Cloud Monitoring and Observability Engineer (Azure) engineer to design, implement, and optimize end-to-end monitoring and observability solutions for a ...
Responsibilities : • Seeking a skilled Cloud Monitoring and Observability Engineer (Azure) engineer to design, implement, and optimize end-to-end monitoring and observability solutions for a ...
Senior Elastic Observability Engineer - Plex
Indiana, PA · Remote
$114K - $172K/yr
The Elastic Observability Engineer is an important member of our Cloud Operations team, building a world-class Application Performance Monitoring solution to support observability-driven development.
Senior Elastic Observability Engineer - Plex
Indiana, PA · Remote
$114K - $172K/yr
The Elastic Observability Engineer is an important member of our Cloud Operations team, building a world-class Application Performance Monitoring solution to support observability-driven development.
Engineer - Observability & Monitoring
Pittsburgh, PA · On-site
$51.25 - $70.25/hr
Engineer - Observability & Monitoring Reports To: Associate Manager - Engineering Location: Pittsburgh, PA American Eagle is a youth culture brand grounded in denim. Our purpose extends beyond making ...
Engineer - Observability & Monitoring
Pittsburgh, PA · On-site
$51.25 - $70.25/hr
Engineer - Observability & Monitoring Reports To: Associate Manager - Engineering Location: Pittsburgh, PA American Eagle is a youth culture brand grounded in denim. Our purpose extends beyond making ...
Engineer - Observability & Monitoring
Pittsburgh, PA · On-site
$51.25 - $70.25/hr
Engineer - Observability & Monitoring Reports To: Associate Manager - Engineering Location: Pittsburgh, PA American Eagle is a youth culture brand grounded in denim. Our purpose extends beyond making ...
Engineer - Observability & Monitoring
Pittsburgh, PA · On-site
$51.25 - $70.25/hr
Engineer - Observability & Monitoring Reports To: Associate Manager - Engineering Location: Pittsburgh, PA American Eagle is a youth culture brand grounded in denim. Our purpose extends beyond making ...
Engineer - Observability & Monitoring
$51.25 - $70.25/hr
Engineer - Observability & Monitoring Reports To: Associate Manager - Engineering Location: Pittsburgh, PA American Eagle is a youth culture brand grounded in denim. Our purpose extends beyond making ...
Engineer - Observability & Monitoring
$51.25 - $70.25/hr
Engineer - Observability & Monitoring Reports To: Associate Manager - Engineering Location: Pittsburgh, PA American Eagle is a youth culture brand grounded in denim. Our purpose extends beyond making ...
Engineer - Observability & Monitoring
Pittsburgh, PA · On-site
$51.25 - $70.25/hr
Engineer - Observability & Monitoring Reports To: Associate Manager - Engineering Location: Pittsburgh, PA American Eagle is a youth culture brand grounded in denim. Our purpose extends beyond making ...
Engineer - Observability & Monitoring
Pittsburgh, PA · On-site
$51.25 - $70.25/hr
Engineer - Observability & Monitoring Reports To: Associate Manager - Engineering Location: Pittsburgh, PA American Eagle is a youth culture brand grounded in denim. Our purpose extends beyond making ...
Partner with engineering and operations teams to use observability data for continuous improvement and outage prevention. Reliability and Incident Enablement * Partner with technology teams to ...
Partner with engineering and operations teams to use observability data for continuous improvement and outage prevention. Reliability and Incident Enablement * Partner with technology teams to ...
Partner with engineering and operations teams to use observability data for continuous improvement and outage prevention. Reliability and Incident Enablement * Partner with technology teams to ...
Partner with engineering and operations teams to use observability data for continuous improvement and outage prevention. Reliability and Incident Enablement * Partner with technology teams to ...
Platform Engineer
Philadelphia, PA · On-site
Job Title Platform Engineer/DevOps Engineer Client Confidential (Financial Services) Location ... Contribute to the design, implementation, and support of enterprise monitoring and observability ...
Quick apply
Platform Engineer
Philadelphia, PA · On-site
Job Title Platform Engineer/DevOps Engineer Client Confidential (Financial Services) Location ... Contribute to the design, implementation, and support of enterprise monitoring and observability ...
Senior / Staff Software Engineer (Observability / SRE)
Pittsburgh, PA · On-site +1
$148K - $249K/yr
Qualifications: - 5+ years software engineering or systems/performance engineering experience (BS ... Bonus: - Experience deploying and managing observability platforms (OpenTelemetry, Grafana OSS ...
Quick apply
Senior / Staff Software Engineer (Observability / SRE)
Pittsburgh, PA · On-site +1
$148K - $249K/yr
Qualifications: - 5+ years software engineering or systems/performance engineering experience (BS ... Bonus: - Experience deploying and managing observability platforms (OpenTelemetry, Grafana OSS ...
Build and operate observability infrastructure across the three pillars - metrics, logging, and ... engineering, and technology problems in support of the Navy, the Intel Community (IC), and other ...
Posted today
Build and operate observability infrastructure across the three pillars - metrics, logging, and ... engineering, and technology problems in support of the Navy, the Intel Community (IC), and other ...
Posted today
Senior / Staff Software Engineer (Observability / SRE)
Pittsburgh, PA · On-site +1
$148K - $249K/yr
Qualifications: - 5+ years software engineering or systems/performance engineering experience (BS ... Bonus: - Experience deploying and managing observability platforms (OpenTelemetry, Grafana OSS ...
Senior / Staff Software Engineer (Observability / SRE)
Pittsburgh, PA · On-site +1
$148K - $249K/yr
Qualifications: - 5+ years software engineering or systems/performance engineering experience (BS ... Bonus: - Experience deploying and managing observability platforms (OpenTelemetry, Grafana OSS ...
Senior / Staff Software Engineer (Observability / SRE)
Pittsburgh, PA · On-site +1
$148K - $249K/yr
Qualifications: - 5+ years software engineering or systems/performance engineering experience (BS ... Bonus: - Experience deploying and managing observability platforms (OpenTelemetry, Grafana OSS ...
Senior / Staff Software Engineer (Observability / SRE)
Pittsburgh, PA · On-site +1
$148K - $249K/yr
Qualifications: - 5+ years software engineering or systems/performance engineering experience (BS ... Bonus: - Experience deploying and managing observability platforms (OpenTelemetry, Grafana OSS ...
AI Orchestration Engineer
$120K - $202K/yr
Build AI observability, evaluation, telemetry, and performance measurement solutions. Education & Qualifications Minimum Qualifications * Bachelor's degree in Computer Science, Data Engineering ...
AI Orchestration Engineer
$120K - $202K/yr
Build AI observability, evaluation, telemetry, and performance measurement solutions. Education & Qualifications Minimum Qualifications * Bachelor's degree in Computer Science, Data Engineering ...
AI Orchestration Engineer
Berwyn, PA · On-site
$90K - $157K/yr
Build AI observability, evaluation, telemetry, and performance measurement solutions. Education & Qualifications Minimum Qualifications * Bachelor's degree in Computer Science, Data Engineering ...
AI Orchestration Engineer
Berwyn, PA · On-site
$90K - $157K/yr
Build AI observability, evaluation, telemetry, and performance measurement solutions. Education & Qualifications Minimum Qualifications * Bachelor's degree in Computer Science, Data Engineering ...
DataOps & Reliability Engineer
Erie, PA · On-site
$99K - $125K/yr
The role focuses on pipeline automation, monitoring and observability, incident management, release ... engineering. • Experience with pipeline automation, orchestration, scheduling, and CI/CD. • ...
DataOps & Reliability Engineer
Erie, PA · On-site
$99K - $125K/yr
The role focuses on pipeline automation, monitoring and observability, incident management, release ... engineering. • Experience with pipeline automation, orchestration, scheduling, and CI/CD. • ...
Knowledge of Kubernetes observability, monitoring, logging, and tracing solutions * Experience ... Engineer and maintain upstream Kubernetes environments supporting scalable software delivery
Knowledge of Kubernetes observability, monitoring, logging, and tracing solutions * Experience ... Engineer and maintain upstream Kubernetes environments supporting scalable software delivery
Senior Platform Engineer
$101K - $139K/yr
Lead cloud costoptimization efforts, monitor consumption trends, identify savings opportunities, and support observability engineering including logging, monitoring, and performance tuning. Drive ...
Senior Platform Engineer
$101K - $139K/yr
Lead cloud costoptimization efforts, monitor consumption trends, identify savings opportunities, and support observability engineering including logging, monitoring, and performance tuning. Drive ...
URBN Senior DevOps Engineer
Philadelphia, PA · On-site
$124K - $159K/yr
... Engineer to manage and evolve their platform layer for ecommerce and AI initiatives. The role ... Build for reliability, observability, and graceful failure. • CI/CD & Deployment Pipelines ...
URBN Senior DevOps Engineer
Philadelphia, PA · On-site
$124K - $159K/yr
... Engineer to manage and evolve their platform layer for ecommerce and AI initiatives. The role ... Build for reliability, observability, and graceful failure. • CI/CD & Deployment Pipelines ...
Senior AWS Agent core Platform Engineer-Hybrid
Reading, PA · On-site
$60 - $63/hr
Job Responsibilities Observability * Assess CloudWatch, X-Ray, Bedrock logging, AgentCore traces vs. agentic workflow requirements; produce gap analysis, Setup observability in Dynatrace * Design ...
Senior AWS Agent core Platform Engineer-Hybrid
Reading, PA · On-site
$60 - $63/hr
Job Responsibilities Observability * Assess CloudWatch, X-Ray, Bedrock logging, AgentCore traces vs. agentic workflow requirements; produce gap analysis, Setup observability in Dynatrace * Design ...
Observability Engineer information
See Pennsylvania salary details
$20.11 - $27.04
2% of jobs
$27.04 - $33.96
4% of jobs
$33.96 - $40.88
6% of jobs
$40.88 - $47.80
8% of jobs
$48.98 is the 25th percentile. Wages below this are outliers.
$47.80 - $54.72
23% of jobs
The median wage is $58.18 / hr.
$54.72 - $61.64
12% of jobs
$65.85 is the 75th percentile. Wages above this are outliers.
$61.64 - $68.56
32% of jobs
$68.56 - $75.48
11% of jobs
$75.48 - $82.41
1% of jobs
$82.41 - $89.33
0% of jobs
$89.33 - $96.25
1% of jobs
$20
$58
$96
How much do observability engineer jobs pay per hour?
What does an observability engineer do?
An Observability Engineer is responsible for designing, implementing, and maintaining monitoring, logging, and tracing systems to ensure the health, performance, and reliability of applications and infrastructure. They work with tools like Prometheus, Grafana, OpenTelemetry, and ELK to collect and analyze telemetry data. Their goal is to provide visibility into system behavior, detect and diagnose issues quickly, and improve overall system observability. They collaborate with developers, SREs, and operational teams to create automated and scalable observability solutions.
What are the key skills and qualifications needed to thrive as an observability engineer?
To thrive as an Observability Engineer, you need a solid understanding of monitoring, logging, and tracing systems, as well as expertise in programming, cloud infrastructure, and incident response. Proficiency in tools like Prometheus, Grafana, ELK stack, and familiarity with cloud platforms such as AWS, Azure, or GCP are commonly required, and certifications like AWS Certified DevOps Engineer can be advantageous. Strong analytical thinking, collaborative skills, and effective communication are essential soft skills for diagnosing issues and working across development and operations teams. These competencies are vital for proactively maintaining system reliability, ensuring performance, and resolving complications before they impact business operations.

Full-time
Re-posted 21 days ago
Job description
Veterans Sourcing Group, LLC is seeking a skilled Cloud Monitoring and Observability Engineer to design, implement, and optimize end-to-end monitoring and observability solutions for mission-critical applications in the Azure environment. The role involves collaborating with various teams to ensure comprehensive monitoring coverage and maintaining enterprise monitoring platforms to deliver actionable insights.
Responsibilities:
• Seeking a skilled Cloud Monitoring and Observability Engineer (Azure) engineer to design, implement, and optimize end-to-end monitoring and observability solutions for a mission-critical application deployed in the Azure environment.
• The ideal candidate has hands-on experience with enterprise monitoring tools—such as AppDynamics, Thousand Eyes, NetScout, and SolarWinds (or equivalent alternatives)—and a strong background in building scalable, secure, and compliant observability stacks for cloud deployments.
• Will collaborate closely with application engineering, cloud platform, network, and security teams to ensure comprehensive coverage across application, infrastructure, and network layers
• Design and implement end-to-end monitoring, alerting, and observability for an Azure-hosted application across application, infrastructure, network, and user experience layers.
• Configure, integrate, and maintain enterprise monitoring platforms to deliver actionable telemetry, performance baselines, and SLA/SLO tracking.
• Build dashboards, health checks, synthetic tests, and alerting workflows; optimize alert fidelity to minimize noise and improve signal-to-noise ratio.
• Establish and document telemetry standards (metrics, logs, traces), data collection strategies, and service-level indicators (SLIs) aligned to reliability objectives (SLOs).
• Integrate Azure-native services (Azure Monitor, Log Analytics, Application Insights) with enterprise tools to provide unified visibility and correlation.
• Implement network performance monitoring, path visibility, and internet/extranet testing using NPM tools (e.g., ThousandEyes, NetScout); leverage infrastructure monitoring platforms (e.g., SolarWinds) for device and service health.
• Instrument applications with APM tools (e.g., AppDynamics, Dynatrace, New Relic) for business transaction monitoring, dependency mapping, and root-cause analysis; tune anomaly detection and policy thresholds.
• Collaborate with DevOps/SRE teams to embed monitoring into CI/CD and infrastructure-as-code patterns; ensure new services adhere to observability standards.
• Define runbooks and escalation paths; support incident response and post-incident reviews with data-driven insights and remediation recommendations.
• Ensure monitoring solutions meet applicable security and compliance requirements; support audit requests with clear documentation and evidence.
• Conduct capacity and performance trend analysis; recommend optimization, right-sizing, and resilience improvements.
• Provide knowledge transfer, documentation, and training on monitoring tools, best practices, and operational workflows.
Qualifications:
Required:
• 5+ years implementing enterprise monitoring/observability for cloud or hybrid environments, including mission-critical applications.
• Demonstrable expertise with at least one tool in each category (or equivalent), including production deployments, advanced configuration, and operational use: Application Performance Monitoring (APM): AppDynamics, Dynatrace, or New Relic.
• Experience instrumenting services for business transaction tracing, code-level diagnostics, service maps, and anomaly detection.
• Ability to design APM dashboards and create alert policies with appropriate thresholds and baselines.
• Network Performance Monitoring (NPM) / Digital Experience Monitoring (DEM): Thousand Eyes, NetScout, or Kentik.
• Experience with synthetic tests, path visualization, packet-level analysis, and internet/WAN performance monitoring.
• Ability to configure endpoint agents, BGP/DNS tests, and multi-hop path monitoring for user experience correlation.
• Infrastructure Monitoring and Event Management: SolarWinds, Microsoft SCOM, Datadog, or Prometheus/Grafan.
• Experience monitoring servers, containers, network devices, and cloud services; creating availability and capacity dashboards.
• Proficiency with alert routing, de-duplication, and event correlation.
• Strong Azure monitoring experience: Azure Monitor, Log Analytics (KQL), Application Insights, and integration with third-party tools.
• Solid understanding of distributed tracing, metrics, and log aggregation; familiarity with Open Telemetry concepts and data pipelines.
• Scripting/automation skills (PowerShell, Python, or Bash) to automate monitoring configuration, agent deployment, test creation, and reporting.
• Networking fundamentals (DNS, BGP, HTTP, TLS, TCP/IP), CDN concepts, and WAN performance monitoring; ability to correlate app and network telemetry.
• Experience supporting incident response and performance troubleshooting across applications, infrastructure, and network layers.
• Excellent documentation and communication skills; collaborative mindset with engineering, operations, and security stakeholders.
Preferred:
• Background in regulated environments (financial services, government, healthcare) with compliance-aware monitoring design.
• Experience with log aggregation and SIEM/SOAR platforms (e.g., Splunk, Elastic) and integration with APM/NPM tools.
• Integration experience with ITSM platforms (e.g., ServiceNow) for incident, change, and problem management workflows.
• Familiarity with infrastructure-as-code (ARM/Bicep/Terraform) and embedding observability into IaC patterns; experience with CI/CD integration.
• Exposure to SRE practices (SLIs/SLOs, error budgets, reliability reviews) and capacity/performance planning.
• Ability to code in one or more of the following languages for instrumentation, custom telemetry, SDK integration, and tooling automation: Java: Implementing Open Telemetry SDKs/agents, custom instrumentation, and APM tagging; building synthetic test harnesses. .NET (C#): Instrumenting ASP.NET services, configuring APM auto-instrumentation, writing custom exporters and health probes. Python: Building automation scripts, collectors/exporters, synthetic tests, and integrating with monitoring APIs and SDKs.
Company:
Welcome to the Veterans Souring Group company profile. Veterans Sourcing Group (VSG) is a “Service Disabled Veteran Owned Small Business – SDVOSB”. Founded in , the company is headquartered in New City, USA, with a team of 11-50 employees. The company is currently Early Stage.
About Veterans Sourcing Group
Sourced by ZipRecruiter
Veterans Sourcing Group is a renowned company headquartered in New York, NY, US and is committed to providing high-quality, reliable staffing solutions and advisory services. The company operates in the human resources and staffing industry, specializing in veteran hiring. They offer various solutions to meet client needs, including strategic consultancy, professional search, and contract staffing. The company was founded by Beth Vines and Bruce Madnick, respected professionals who recognized a gap in the market for veteran-focused staffing services which prompted them to establish Veterans Sourcing Group in 2011.
Industry
Recruiting and staffing services
Company size
51 - 200 Employees
Headquarters location
New York, NY, US