1

Monitoring Alerting Jobs (NOW HIRING)

Senior Observability Engineer

Chelsea, MA ยท On-site

$141K - $181K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

Senior Observability Engineer

Boston, MA ยท On-site

$141K - $181K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

Senior Observability Engineer

Cambridge, MA ยท On-site

$142K - $182K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

Senior Observability Engineer

Hingham, MA

$136K - $175K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

Senior Observability Engineer

Hyde Park, MA ยท On-site

$132K - $170K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

Senior Observability Engineer

Salem, MA ยท On-site

$142K - $182K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

Senior Observability Engineer

Wayland, MA

$149K - $192K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

AI with SRE

Irving, TX ยท On-site

$54.75 - $72.75/hr

Incorporate GenAI tooling and agentic capabilities to strengthen reliability outcomes across monitoring/alerting, rapid incident response, change management/testing, and DevOps/deployment processes.

Senior Observability Engineer

Quincy, MA ยท On-site

$136K - $175K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

Senior Observability Engineer

Reading, MA ยท On-site

$137K - $176K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

Senior Observability Engineer

Randolph, MA

$132K - $170K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

Senior Observability Engineer

Peabody, MA ยท On-site

$144K - $185K/yr

The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy ...

Showing results 41-60

Monitoring Alerting information

What is monitoring alerting?

Monitoring alerting refers to the process of continuously observing systems, applications, or networks to detect issues, anomalies, or failures in real-time, and automatically notifying the appropriate teams or individuals when predefined thresholds or events occur. This practice helps organizations quickly respond to incidents, minimize downtime, and maintain system reliability. Monitoring tools collect performance and health data, while alerting mechanisms ensure that critical problems are promptly addressed before they impact users or business operations.

What are the key skills and qualifications needed to thrive as a monitoring alerting specialist?

To thrive as a Monitoring and Alerting Specialist, you need a solid understanding of IT systems, incident management, and network troubleshooting, often supported by a degree in computer science or related fields. Familiarity with monitoring tools such as Nagios, Zabbix, Splunk, or Prometheus, and certifications like CompTIA Network+ or ITIL, are typically required. Strong attention to detail, analytical thinking, and clear communication are important soft skills for identifying issues and coordinating responses. These skills ensure timely detection and resolution of system problems, minimizing downtime and maintaining operational reliability.

How does a monitoring alerting professional typically collaborate with IT and development teams during incident response?

Monitoring Alerting professionals play a critical role in incident response by acting as the first line of defense when anomalies or outages are detected. They work closely with IT and development teams by promptly communicating relevant alerts, providing context, and escalating issues according to established protocols. During an incident, they may participate in real-time troubleshooting sessions, share logs and metrics, and help coordinate a swift resolution. This collaborative approach ensures minimal downtime and effective root-cause analysis for future prevention.

What is the difference between Monitoring Alerting vs Network Monitoring?

AspectMonitoring AlertingNetwork Monitoring
Primary FocusDetects and notifies about system or application issuesMonitors network traffic, devices, and performance
Tools & CertificationsAlert management tools, monitoring platforms, certifications like CompTIA Linux+Network analyzers, SNMP tools, certifications like Cisco CCNA
Work EnvironmentIT operations, DevOps, system administratorsNetwork engineers, infrastructure teams

Monitoring Alerting focuses on identifying and notifying about issues within systems or applications, while Network Monitoring specifically tracks network performance and device health. Both roles require technical skills and often overlap in IT environments, but they serve different operational needs.

What other helpful pages are available for Monitoring Alerting?

Other pages related to Monitoring Alerting:

Infographic showing various Monitoring Alerting job openings in the United States as of September 2026, with employment types broken down into 1% As Needed, 85% Full Time, 10% Part Time, and 4% Contract. Highlights an 91% Physical, 2% Hybrid, and 7% Remote job distribution.

Senior Enterprise Monitoring and Event Management Engineer

Columbus, OH โ€ข On-site

Manifest Solutions
IT Servicesย โ€ขย 51 - 200 employees

$101K - $138K/yr

Full-time

Re-posted yesterday


Job description

Manifest Solutions is currently seeking a Infrastructure Engineer Sr / Principal - Enterprise Observability & Monitoring Engineer (Dynatrace) in Columbus, OH.
Key Responsibilities
Enterprise Monitoring & Observability
  • Administer and support Dynatrace enterprise monitoring environments.
  • Design and implement monitoring solutions for business-critical applications and infrastructure.
  • Develop standards for monitoring, alerting, dashboards, and operational observability.
  • Continuously improve enterprise monitoring coverage and accuracy.
  • Tune alerting thresholds and event correlation to reduce false positives while ensuring timely incident detection.
  • Support enterprise observability initiatives across on-premises, cloud, and hybrid platforms.

Application Performance Monitoring (APM)
  • Deploy and administer Dynatrace OneAgent technologies.
  • Onboarding large scale applications
  • Monitor application health, service dependencies, user experience, and transaction performance.
  • Support Real User Monitoring (RUM) and Synthetic Monitoring implementations.
  • Analyze application performance issues and identify performance bottlenecks.
  • Provide end-to-end visibility across application ecosystems.

Infrastructure Monitoring
  • Monitor Windows, Linux, VMware, OpenShift, and cloud-hosted environments.
  • Support monitoring for databases, middleware, network infrastructure, storage systems, domain services, load balancers, and enterprise applications.
  • Assist infrastructure teams with capacity planning and performance optimization.
  • Proactively identify monitoring gaps and service risks.

Monitoring Engineering & Automation
  • Develop custom monitoring solutions and Dynatrace Extensions 2.0.
  • Create and maintain custom SQL, Oracle, and enterprise application monitoring extensions.
  • Build dashboards, metrics, health checks, and alerting policies.
  • Automate monitoring deployment and configuration processes.
  • Support infrastructure-as-code and configuration management initiatives where applicable.

Incident Response & Root Cause Analysis
  • Participate in major incident response and war room activities.
  • Perform root cause analysis for application, infrastructure, and monitoring-related incidents.
  • Provide monitoring expertise during outage investigations.
  • Develop corrective actions and preventative monitoring improvements.

ServiceNow & Event Management Integration
  • Support integration between Dynatrace, ServiceNow, and enterprise event management platforms.
  • Define monitoring requirements needed for automated incident creation.
  • Assist with event correlation, CI alignment, and ticket automation workflows.
  • Partner with operations teams to improve incident response processes.

Cloud & Modern Application Monitoring
  • Support monitoring solutions for AWS, Azure, SaaS, containerized, and microservices-based applications.
  • Deploy and support ActiveGate infrastructure.
  • Assist application teams with onboarding modern cloud workloads into enterprise monitoring platforms.
  • Support SSO, authentication, synthetic transaction monitoring, and end-user experience monitoring.

Customer Engagement & Consulting
  • Consult with application owners and infrastructure teams regarding monitoring requirements.
  • Lead onboarding activities for new applications and services.
  • Conduct monitoring requirement workshops and technical discovery sessions.
  • Provide guidance and best practices across the enterprise.

Governance, Compliance & Documentation
  • Maintain monitoring architecture documentation, procedures, and standards.
  • Support NERC/CIP and other regulatory monitoring requirements.
  • Participate in change management, CAB reviews, and operational governance processes.
  • Develop and maintain operational runbooks and support documentation.

Required Qualifications
Education
Bachelor's Degree in Computer Science, Information Technology, Information Systems, Engineering, or related technical field; or equivalent work experience.
Experience
  • 5+ years supporting enterprise monitoring or observability platforms.
  • 5+ years supporting large-scale enterprise infrastructure environments.
  • Experience administering Dynatrace or comparable monitoring solutions.
  • Experience working in regulated utility, critical infrastructure, energy, or large enterprise environments preferred.

Required Technical Skills
  • Dynatrace Administration and Engineering
  • Application Performance Monitoring (APM)
  • Observability Engineering
  • Synthetic Monitoring
  • Real User Monitoring (RUM)
  • Windows Server Administration
  • Linux Administration
  • VMware Monitoring
  • AWS and Azure Monitoring
  • ActiveGate Deployment and Administration
  • Monitoring Dashboard Development
  • Alert Management and Event Correlation
  • ServiceNow Integration
  • Root Cause Analysis
  • Enterprise Incident Management
  • TCP/IP, DNS, SSL, Load Balancers, and Networking Fundamentals
  • Database Monitoring (Oracle, SQL Server, PostgreSQL, etc.)
  • ITIL Incident, Problem, and Change Management

Preferred Qualifications
  • Dynatrace Professional or Associate Certification
  • Splunk administration experience
  • SolarWinds experience
  • OpenShift or Kubernetes monitoring experience
  • Experience developing Dynatrace Extensions 2.0
  • PowerShell, Python, Bash, or automation scripting
  • ServiceNow Event Management integration experience
  • Experience supporting enterprise authentication solutions (Entra ID, SSO, LDAP, SAML, OAuth)
  • Experience supporting NERC/CIP regulated environments

Leadership & Soft Skills
  • Strong customer service and stakeholder management skills
  • Exceptional troubleshooting and analytical abilities
  • Ability to communicate effectively with technical and non-technical audiences
  • Ability to prioritize multiple projects and operational demands
  • Self-directed and highly accountable
  • Strong collaboration skills across infrastructure, security, cloud, and application teams
  • Ability to lead technical initiatives and mentor junior engineers
  • Strategic mindset with the ability to align monitoring capabilities with business outcomes