1

Observability Jobs in Ohio (NOW HIRING)

Infrastructure Engineer Staff - Dynatrace SME

Gahanna, OH · On-site

$101K - $132K/yr

AEP is seeking an experienced Enterprise Observability and Monitoring Engineer to serve as a subject matter expert for enterprise monitoring technologies including Dynatrace and related monitoring ...

SDET

Columbus, OH · On-site

$48.25 - $62.25/hr

Observability tools (datadog, splunk, Grafana) * Kubernetes Infrastructure SDET - Enterprise Ordering Platform Overview Client is hiring an Infrastructure SDET to support the rollout of next ...

Site Reliability Engineer

Strongsville, OH · On-site

$52.50 - $70/hr

Develop and maintain observability solutions that provide visibility into application health, performance, and reliability. Track and manage dashboard-related projects and initiatives from request ...

Site Reliability Engineer

Columbus, OH · On-site

$53.25 - $70.75/hr

Resources must carry SRE-aligned mindset to support our Databricks/Genie/Attacama/Sigma platform and observability operations, with a strong preference for Databricks and/or Sigma expertise while ...

SDET

Columbus, OH · On-site

$50/hr

Observability tools (datadog, splunk, Grafana) * Kubernetes Infrastructure SDET - Enterprise Ordering Platform Overview Client is hiring an Infrastructure SDET to support the rollout of next ...

$54.75 - $72.75/hr

Ensure full-stack observability across infrastructure and services. * Design and maintain CI/CD pipelines that are fast, reproducible, and safe. * Improve deployment strategies (rollouts, canaries ...

Improve system observability through metrics, structured logging, dashboards, and alerting * Participate in code and design reviews with a strong emphasis on security, correctness, and failure modes

Senior Cloud Engineer (Azure)

Maumee, OH

$52.50 - $70.25/hr

Strengthen observability, incident response, and cloud cost optimization. Job Responsibilities: * Contribute to the design and implementation process for Azure Architecture, ensuring alignment with ...

Showing results 21-40

Observability information

See Ohio salary details

$15

$57

$82

How much do observability jobs pay per hour?

As of Aug 10, 2026, the average hourly pay for observability in Ohio is $57.55, according to ZipRecruiter salary data. Most workers in this role earn between $48.22 and $66.06 per hour, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive in an observability role?

To thrive in an Observability role, you need a strong background in monitoring, alerting, logging, and analyzing system performance, often supported by a degree in computer science or related field. Familiarity with tools such as Prometheus, Grafana, Datadog, Splunk, and experience with cloud platforms and scripting languages is crucial. Excellent problem-solving, communication, and collaboration skills help you work effectively with cross-functional engineering and operations teams. These capabilities are essential to ensure system reliability, quickly detect issues, and maintain seamless digital experiences.

What is an observability?

An Observability job focuses on ensuring the performance, reliability, and health of software systems by collecting, analyzing, and visualizing telemetry data such as logs, metrics, and traces. Professionals in this field work with monitoring tools, distributed tracing, and alerting systems to detect and troubleshoot issues proactively. They collaborate with engineering and operations teams to improve system visibility, reduce downtime, and enhance overall system performance.

Is observability a good career?

Observability is a growing field within IT and software engineering that involves monitoring, logging, and analyzing system performance to ensure reliability and efficiency. It often requires skills in tools like Prometheus, Grafana, and data analysis, making it a valuable and in-demand career path with opportunities for advancement. The role typically involves collaboration across teams and continuous learning to keep up with evolving technologies.

What does an observability do?

In an Observability role, your daily tasks often include designing and maintaining monitoring dashboards, configuring alerts, analyzing system logs, and working closely with development and operations teams to troubleshoot issues. You'll proactively identify areas of improvement to increase system reliability, document monitoring strategies, and support incident response efforts. Collaboration is key, as you may participate in post-incident reviews and help drive architectural improvements based on the data you collect. The role is dynamic and requires a proactive approach to ensure systems stay healthy and downtime is minimized.

What are the most commonly searched types of Observability jobs in Ohio? The most popular types of Observability jobs in Ohio are:
What are popular job titles related to Observability jobs in Ohio? For Observability jobs in Ohio, the most frequently searched job titles are:
What job categories do people searching Observability jobs in Ohio look for? The top searched job categories for Observability jobs in Ohio are:
What cities in Ohio are hiring for Observability jobs? Cities in Ohio with the most Observability job openings:
Infographic showing various Observability job openings in Ohio as of August 2026, with employment types broken down into 85% Full Time, 10% Part Time, and 5% Contract. Highlights an 76% Physical, 6% Hybrid, and 18% Remote job distribution, with an average salary of $119,701 per year, or $57.5 per hour.

Platform Operations Engineer (Site Reliability Engineer)

Vertiv Co

Westerville, OH • On-site

$100 - $130/hr

Other

Re-posted 29 days ago


Vertiv rating

6.7

Company rating: 6.7 out of 10

Based on 64 frontline employees who took The Breakroom Quiz

377th of 487 rated machine equipment manufacturers


Job description

Job Summary

Vertiv is seeking a skilled Platform Operations Engineer (Site Reliability Engineer) to serve as the owner of cross‑platform observability, incident management, and operational reliability within Vertiv’s Digital organization. This individual contributor role is responsible for designing, implementing, and continuously improving monitoring and alerting solutions across Vertiv’s digital platform ecosystem — including Compass AI, Writer AI, Site Scope, UiPath, Workato, Cursor, and other approved enterprise tools — while owning incident response processes, SLA management, and operational governance. The Platform Operations / SRE will operate within the Digital organization and play a central role in advancing Vertiv’s Operational Excellence strategic priority by ensuring the availability, performance, and resilience of platforms that power critical digital workflows and business functions.

As an individual contributor in a lead capacity, this role includes proactive reliability engineering — applying SRE principles such as SLOs, error budgets, and blameless post‑mortems — and embedding secure coding and operational governance practices across the Digital organization. The Platform Operations / SRE Engineer will define and enforce observability standards, lead incident response and root cause analysis, manage platform‑level SLAs, and partner with engineering, security, and business stakeholders to ensure that all digital platforms meet agreed availability and performance targets.

This position partners closely with IT Security, NPDI, Digital delivery teams, and business operations, and is based on site at Vertiv’s Westerville, OH headquarters.

Responsibilities
  • Own Cross‑Platform Monitoring & Observability: Design, implement, and maintain end‑to‑end monitoring, alerting, and observability solutions across Vertiv’s digital platform ecosystem — including AI platforms, automation tools, and internal applications — ensuring real‑time visibility into system health, performance, and availability.
  • Lead Incident Response & Management: Serve as the primary escalation point and incident commander for P1/P2 incidents across Digital platforms; lead root cause analysis (RCA), blameless post‑mortems, and corrective action tracking to prevent recurrence and reduce mean time to resolution (MTTR).
  • Manage Platform SLAs & Reliability Targets: Define, instrument, and enforce service level objectives (SLOs), service level indicators (SLIs), and error budgets across Digital platforms; produce regular SLA performance reports for leadership and drive platform improvements to meet or exceed agreed availability and performance targets.
  • Drive Secure Coding & Operational Governance: Champion secure coding practices and DevSecOps standards within Digital delivery teams; conduct operational readiness reviews for new platform deployments, enforce configuration management and change control processes, and partner with IT Security and NPDI to ensure all platforms meet Vertiv’s security and compliance requirements.
  • Automate Operations & Reduce Toil: Identify and eliminate manual operational toil through automation. This includes automated remediation runbooks and anomaly detection through the use of scripting, IaC tools, and approved automation platforms.
  • Capacity Planning & Performance Engineering: Analyze platform utilization trends and conduct capacity planning across Digital environments; proactively identify performance bottlenecks and recommend architectural improvements to ensure platforms scale reliably with business demand.
  • CI/CD Pipeline Reliability & Deployment Support: Partner with Digital delivery teams to ensure CI/CD pipelines are instrumented for reliability, deployment risk is managed through progressive rollout strategies, and production deployments are supported with appropriate rollback and health‑check capabilities.
  • Evaluate & Advance Observability Tooling: Stay current on advancements in observability, AIOps, and SRE tooling; evaluate and recommend new tools and practices that enhance Vertiv’s platform operations maturity, and drive adoption of modern reliability engineering standards across the Digital organization.
Requirements
  • Bachelor’s degree in Computer Science, Information Systems, Engineering, or a related field; equivalent practical experience considered.
  • 5+ years of professional experience in platform operations, site reliability engineering, DevOps, or a related software/infrastructure engineering discipline.
  • 3+ years of hands‑on experience with enterprise monitoring and observability platforms (e.g., Datadog, Grafana, Prometheus, Azure Monitor, Splunk, or equivalent) in a multi‑platform environment.
  • Demonstrated experience owning and managing incident response processes, post‑mortem facilitation, and SLA/SLO governance.
  • Experience implementing secure coding practices, DevSecOps standards, or operational governance frameworks in an enterprise software delivery environment.
Technical Skills
  • Proficiency with monitoring and observability tools (Datadog, Grafana, Prometheus, Azure Monitor, Splunk, or equivalent) for cross‑platform health and performance tracking.
  • Strong knowledge of SRE principles, including SLOs, SLIs, blameless post‑mortems, and toil reduction practices.
  • Hands‑on experience with cloud platforms (AWS preferred) and familiarity with containerized environments (Docker, Kubernetes) and infrastructure‑as‑code tooling (Terraform, Ansible, or equivalent).
  • Proficiency in multiple programming languages (Python, Ruby, PowerShell, Java, JavaScript, C#, etc.) for automation and runbook development.
  • Experience with CI/CD platforms (GitLab, Jenkins, GitHub Actions, Azure DevOps, or equivalent) and deployment reliability practices including progressive rollout, feature flags, and automated health checks.
Preferred Qualifications
  • Google SRE certification, AWS DevOps Professional, Azure certifications, or equivalent SRE/cloud operations certification.
  • Experience with AIOps tooling or AI‑assisted anomaly detection and automated remediation capabilities.
  • Familiarity with the Vertiv digital platform ecosystem: Workato, UiPath, Power Automate, Compass AI, Writer AI, or Cursor.
  • Experience applying DevSecOps practices, including SAST/DAST scanning, secrets management, and compliance‑as‑code in enterprise environments.
  • Experience working in Agile/Scrum delivery environments; familiarity with ITIL incident and change management frameworks.
Work Authorization

No calls or agencies please. Vertiv will only employ those who are legally authorized to work in the United States. This is not a position for which sponsorship will be provided. Individuals with temporary visas such as E, F‑1, H‑1, H‑2, L, B, J, or TN or who need sponsorship for work authorization now or in the future, are not eligible for hire.

Equal Opportunity Employer

Vertiv is an Equal Opportunity/Affirmative Action employer. We promote equal opportunities for all with respect to hiring, terms of employment, mobility, training, compensation, and occupational health, without discrimination as to age, race, color, religion, creed, sex, pregnancy status (including childbirth, breastfeeding, or related medical conditions), marital status, sexual orientation, gender identity / expression (including transgender status or sexual stereotypes), genetic information, citizenship status, national origin, protected veteran status, political affiliation, or disability. If you have a disability and are having difficulty accessing or using this website to apply for a position, you can request help by sending an email to help.join@vertiv.com.

#J-18808-Ljbffr

What Vertiv employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom