1

Executive Observability Engineer Jobs in Ohio (NOW HIRING)

$180 - $230/hr

Strong ability to communicate complex technical ideas clearly and influence executive, non ... observability, and service identity. * Be accountable for platform reliability and operational ...

New

Software Engineer AI/ML

Evendale, OH ยท On-site

$109K - $132K/yr

You'll also contribute to AI strategy and partner with executive stakeholders to align on ... Implement monitoring and observability for AI/ML systems to track model performance, data drift ...

VP Cloud Platform Engineering

Columbus, OH

$173K - $224K/yr

This executive leader oversees multiple platform engineering organizations and establishes ... observability, availability, recovery, and performance engineering. Drive continuous improvement ...

next page

Showing results 1-20

Executive Observability Engineer information

What is the difference between Executive Observability Engineer vs Site Reliability Engineer?

AspectExecutive Observability EngineerSite Reliability Engineer
CredentialsTypically requires expertise in observability tools, monitoring, and cloud platformsRequires skills in systems engineering, coding, and infrastructure management
Work EnvironmentFocuses on designing observability solutions, analyzing system health, and strategic monitoringManages system reliability, automates deployment, and maintains infrastructure
Industry UsageUsed in tech companies emphasizing system visibility and performance analysisCommon in cloud services, SaaS, and large-scale web services

The Executive Observability Engineer primarily focuses on implementing and optimizing observability tools to ensure system health, while the Site Reliability Engineer concentrates on maintaining system reliability and automating infrastructure. Both roles require technical expertise but differ in their strategic versus operational focus.

What are some common challenges Executive Observability Engineers face when implementing organization-wide monitoring solutions?

Executive Observability Engineers often encounter challenges such as integrating diverse monitoring tools across legacy and modern systems, ensuring data consistency, and balancing comprehensive visibility with system performance. Coordinating with multiple teams to align observability goals and fostering a culture of proactive monitoring can also require strong communication and leadership skills. Successfully managing these complexities leads to more resilient infrastructure and improved incident response times.

What are Executive Observability Engineers?

Executive Observability Engineers are specialized IT professionals who design and manage systems that monitor, analyze, and optimize the health and performance of an organization's technology infrastructure. They focus on providing high-level visibility into applications, networks, and services, enabling executives to make data-driven decisions. These engineers implement observability tools, create dashboards, and generate reports that translate complex technical metrics into actionable business insights. Their work is crucial for ensuring system reliability, quick incident response, and continuous improvement across the organization.

What are the key skills and qualifications needed to thrive as an Executive Observability Engineer, and why are they important?

To thrive as an Executive Observability Engineer, you need deep expertise in performance monitoring, distributed systems, and troubleshooting, often supported by a degree in computer science or a related field. Familiarity with observability tools such as Datadog, New Relic, Prometheus, and advanced logging systems, as well as certifications like AWS Certified DevOps Engineer, is highly beneficial. Strong analytical thinking, communication, and leadership skills set top candidates apart in this role. These skills and qualities are crucial to proactively detect issues, optimize system performance, and drive strategic decision-making across complex technical environments.
What are the most commonly searched types of Observability Engineer jobs in Ohio? The most popular types of Observability Engineer jobs in Ohio are:
What are popular job titles related to Executive Observability Engineer jobs in Ohio? For Executive Observability Engineer jobs in Ohio, the most frequently searched job titles are:
What job categories do people searching Executive Observability Engineer jobs in Ohio look for? The top searched job categories for Executive Observability Engineer jobs in Ohio are:
What cities in Ohio are hiring for Executive Observability Engineer jobs? Cities in Ohio with the most Executive Observability Engineer job openings:

Observability SME (SolarWinds) - Seattle, Alpharetta or Cincinnati

Digital Technology Solutions

Cincinnati, OH โ€ข On-site

Other

Posted 2 days ago

New


Job description

DTS is looking for Observability SME (SolarWinds) for our Client position based in Seattle, Alpharetta or cincinnati

Job Description

Overview:

Observability & Enterprise Monitoring Engineer with specialized expertise in SolarWinds platform administration and broader multi-tool observability ecosystems. Working knowledge of OpenText NNMi will be an added advantage. This role will be responsible for the end-to-end administration, optimization, integration, and operational maintenance of enterprise-scale implementation of monitoring solutions (SolarWinds). Responsible for ensuring platform health, automate alert workflows, manage hybrid/cloud monitoring integrations, and collaborate closely with cross-functional infrastructure teams to maintain high availability and performance.

Roles & Responsibilities:
  1. Platform Administration & Lifecycle Management (SolarWinds)
  • Core Module Management: Administer and optimize SolarWinds modules including NPM, NCM, NTA, SAM, and the broader Orion / SWOSH (Hybrid Cloud Observability) platform ecosystem.
  • Upgrades & Maintenance: Perform routine and major version updates across platform components; monitor platform health using Active Diagnostics and My Deployment health checks.
  • Polling Infrastructure: Manage, scale, and load-balance Additional Polling Engines (APEs) to ensure optimal performance across enterprise environments.
  • Database & Backup Operations: Perform operational tasks on the underlying MS SQL Database, manage, schedule, and verify configuration and database backup jobs.
  1. Network & Device Observability Operations
  • Discovery & Asset Management: Execute network discoveries, manage node onboarding/offboarding, assign Universal Device Pollers (UnDP), and maintain custom custom attributes and group hierarchies.
  • Configuration Management (NCM): Build and maintain NCM command templates, automate daily startup/running config backups, archive config files, and remediate compliance/transfer failures.
  • Topology & Visualization: Create and maintain dynamic, accurate network topology maps using Network Atlas and modern visual canvases based on operational requirements.
  1. Alerting, Dashboarding & ITSM Integration
  • Signal Optimization: Design, tune, and maintain custom Alert Triggers, Actions, and Thresholds to eliminate alert noise and drive actionable alerting.
  • Ticketing & Automation: Configure bi-directional ITSM/ticketing integrations to enable automatic ticket creation, routing, and lifecycle tracking.
  • Reporting & Visibility: Build custom operational and executive Dashboards, Views, and Reports tailored to stakeholder requirements.
  • Incident Support: Monitor alert channels for operational anomalies, troubleshoot lingering telemetry issues, and collaborate with domain teams to drive root cause resolution.
  1. AIOps Operations
  • Leverage AIOps, machine learning, and pattern-recognition capabilities to identify baseline anomalies, reduce event noise, and drive predictive incident management.
  • Collaborate with cross-functional teams to integrate AI-driven event correlation models and automated remediation workflows into the central monitoring platform.
  1. Integration, Vendor Coordination
  • Manage relationships and support escalations with platform vendors.
  • Work on REST API integrations across applications/tools as per requirements.
  1. Operational Troubleshooting & Diagnostics
  • Perform deep-dive troubleshooting and root-cause analysis for platform-level performance degradations, engine polling failures, and monitoring agent corruptions.
  • Utilize Active Diagnostics and system telemetry to investigate and resolve complex network configuration transfer failures, polling sync latency, and data ingestion issues.
Required Skills:
  • Multi-tool expertise (SolarWinds, OpenText NNMi, Splunk, etc.)
  • Protocol & Telemetry Knowledge: In-depth understanding of SNMP (v2c/v3), WMI, WinRM, Syslog, NetFlow/sFlow, and Observability (Metrics, Logs, Traces).
  • Automation & API Integration: Good to have skills in PowerShell/Python, and API-driven automation for monitoring workflows.
  • AIOps & Intelligent Automation: Basic understanding of AIOps concepts, machine learning algorithms for anomaly detection, automated event correlation, and predictive analytics within modern observability frameworks.
  • Cloud & Hybrid Observability: Hands-on experience extending platform monitoring into AWS, Azure, or Google Cloud Platform environments.
  • Infrastructure Knowledge

o   System Administration: Intermediate knowledge of Windows and Linux administration.

o   Database: Understanding of SQL/Database concepts and standard query execution.

o   Networking: Good understanding of networking concepts including TCP/IP, DNS, DHCP, Routing and Switching.

o   ITSM: Experience in ITSM processes and operational support.

DTS offers excellent compensation package.

Contact:

Karun Sharma

Team Lead

Digital Technology Solutions (DTS)