2

Remote Aiops Engineer Jobs in Ohio (NOW HIRING)

Remote Aiops Engineer information

What is the difference between Remote Aiops Engineer vs Cloud Operations Engineer?

AspectRemote Aiops EngineerCloud Operations Engineer
Required CredentialsCertifications in AI, Machine Learning, Cloud platforms (AWS, Azure)Certifications in Cloud platforms, DevOps, Networking
Work EnvironmentRemote, tech-focused teams managing AI-driven systemsRemote or on-site, managing cloud infrastructure and services
Employer & Industry UsageTech companies, AI startups, cloud service providersCloud service providers, enterprise IT departments
Search & Comparison IntentUnderstanding AI-focused cloud operations rolesManaging cloud infrastructure and services

The Remote Aiops Engineer focuses on maintaining AI-driven systems and automating operations using AI and machine learning, often requiring specialized certifications. In contrast, the Cloud Operations Engineer manages cloud infrastructure, ensuring system reliability and performance. Both roles are remote-friendly and prevalent in tech and cloud industries, but they emphasize different technical skills and responsibilities.

What is a remote AIOps engineer?

Remote Aiops Engineers are IT professionals who specialize in applying artificial intelligence (AI) and machine learning (ML) techniques to automate and enhance IT operations, typically while working from a location outside of a traditional office. They monitor, analyze, and optimize system performance, detect anomalies, and help prevent outages by leveraging AI-driven tools and algorithms. By working remotely, they can support organizations globally, offering flexibility and a broad skill set to manage complex IT environments efficiently.

What are the key skills and qualifications needed to thrive as a remote AIOps engineer, and why are they important?

To thrive as a Remote AIOps Engineer, you need a strong background in IT operations, machine learning, and data analytics, typically supported by a degree in computer science or a related field. Familiarity with monitoring tools (such as Splunk, Datadog, or Prometheus), scripting languages (like Python), and cloud platforms is essential, and certifications like AWS Certified Solutions Architect or Google Cloud Professional Engineer are highly valued. Strong problem-solving skills, effective communication, and the ability to work independently set outstanding candidates apart in a remote setting. These skills and qualifications are crucial for proactively identifying and resolving IT issues, automating processes, and ensuring system reliability in distributed environments.

How do remote AIOps engineers typically collaborate with cross-functional teams to resolve incidents and optimize IT operations?

Remote AIOps Engineers frequently work with IT operations, development, and security teams to monitor, analyze, and automate responses to system events and incidents. Collaboration often happens through virtual meetings, shared dashboards, and incident management tools, ensuring seamless communication despite physical distance. They play a key role in facilitating root cause analysis, sharing insights from AI-driven analytics, and implementing automation solutions to prevent future issues. Effective communication and the ability to translate technical findings for non-technical stakeholders are essential for success in this collaborative environment.
What are the most commonly searched types of Aiops Engineer jobs in Ohio? The most popular types of Aiops Engineer jobs in Ohio are:
What job categories do people searching Remote Aiops Engineer jobs in Ohio look for? The top searched job categories for Remote Aiops Engineer jobs in Ohio are:
What cities in Ohio are hiring for Remote Aiops Engineer jobs? Cities in Ohio with the most Remote Aiops Engineer job openings:

Senior Enterprise Monitoring & Event Management Engineer | 1061927

Hive + Co

Columbus, OH โ€ข Remote

$100K - $138K/yr

Full-time

Posted 3 days ago

New


Job description

The Senior Enterprise Monitoring & Event Management Engineer is responsible for the design, administration, integration, automation, and support of the client's enterprise monitoring and event management platforms. This role serves as a technical subject matter expert for IBM Netcool Operations Insight (NOI), event correlation, alert automation, monitoring integrations, ServiceNow event management, and enterprise observability solutions.
The engineer will design and maintain monitoring solutions across infrastructure, applications, databases, cloud environments, and business-critical systems while driving automation, event reduction, incident response improvements, and platform modernization initiatives.
The position requires deep technical expertise across Linux administration, monitoring platforms, scripting, event management, system integrations, and production support in a large enterprise environment.
**Location: Prefer Columbus OH - open to remote within Arkansas, Indiana, Kentucky, Louisiana, Michigan, Ohio, Oklahoma, Tennessee, Texas, Virginia and/or West Virginia
Required Technical Skills
  • The position requires deep technical expertise across Linux administration, monitoring platforms, scripting, event management, system integrations, and production support in a large enterprise environment.
  • Monitoring & Event Management Platforms
  • IBM Netcool Operations Insight (NOI)
  • IBM Netcool OMNIbus
  • ServiceNow Event Management (EM)
  • Dynatrace
  • Splunk
  • SolarWinds
  • OpenNMS
  • Enterprise monitoring and observability solutions

Operating Systems
  • Linux Administration (RHEL preferred)
  • Unix/Linux troubleshooting
  • Process and service management
  • Performance tuning and diagnostics
  • Linux interprocess communication (IPC)

Programming & Scripting
  • Python
  • Shell Scripting
  • Perl
  • JavaScript
  • SQL
  • JSON
  • XML/XSLT

Integration Technologies
  • REST APIs
  • SOAP APIs
  • Webhooks
  • SNMP
  • Syslog
  • SMTP
  • Kafka
  • Message bus technologies

Databases
  • Oracle
  • SQL Server
  • PostgreSQL
  • MySQL
  • MongoDB

Cloud & Infrastructure
  • AWS monitoring integrations
  • Azure monitoring integrations
  • Kubernetes
  • Docker
  • Virtualization technologies (VMware)

Automation Technologies
  • Ansible
  • NOI Runbooks
  • ServiceNow Flows
  • Event-driven automation
  • Custom workflow development

Analytics & AI
  • AIOps platforms
  • Event correlation
  • Root cause analysis
  • Anomaly detection
  • Log analytics
  • Predictive alerting
  • Machine learning-enabled monitoring

Preferred Skills
  • Utility industry monitoring experience
  • SOX-regulated application support
  • ITIL framework knowledge
  • ServiceNow ITOM
  • CMDB integrations
  • Event Management architecture
  • Infrastructure monitoring design
  • Enterprise observability practices
  • Network monitoring and management systems

Roles & Responsibilities
Event Management Engineering
  • Design, develop, and maintain enterprise event management and correlation solutions.
  • Build, maintain, and optimize Netcool probes, gateways, automation policies, rules, and event processing workflows.
  • Develop alarm correlation, suppression, enrichment, normalization, deduplication, and root cause analytics.
  • Design enterprise event reduction and noise suppression strategies.
  • Identify opportunities to automate operational monitoring processes.

Platform Administration & Support
  • Administer and support IBM Netcool Operations Insight (NOI) production and non-production environments.
  • Maintain platform health, availability, performance, and resiliency.
  • Perform platform upgrades, patching, capacity planning, and lifecycle management.
  • Troubleshoot platform issues involving event processing, integrations, databases, and infrastructure components.
  • Provide Level 3 support for enterprise monitoring and event management solutions.

Monitoring Architecture
  • Design monitoring solutions across servers, databases, network devices, cloud platforms, applications, middleware, and infrastructure services.
  • Develop monitoring standards, onboarding processes, and alerting strategies for business-critical systems.
  • Partner with application teams to create meaningful actionable alerts and reduce alert fatigue.
  • Improve service visibility and operational readiness across enterprise environments.

ServiceNow Integration
  • Design and support integrations between monitoring platforms and ServiceNow.
  • Implement automated incident creation, event enrichment, routing, escalation, and ticket lifecycle management.
  • Support Event Management and AIOps initiatives within ServiceNow.
  • Integrate monitoring tools with CMDB and service mapping solutions.

Integration Engineering
  • Develop and support integrations using REST, SOAP, SNMP, Syslog, webhooks, messaging technologies, and custom APIs.
  • Integrate monitoring platforms with infrastructure, application, database, and cloud environments.
  • Support enterprise onboarding activities for new monitoring sources and technologies.

Automation & Continuous Improvement
  • Develop event-driven automation and remediation workflows.
  • Create self-healing and closed-loop operational automation capabilities.
  • Implement monitoring automation using scripts, workflows, runbooks, and orchestration platforms.
  • Drive monitoring modernization initiatives and platform improvements.

Production Operations Support
  • Participate in on-call rotation and critical incident response activities.
  • Support enterprise monitoring platforms that provide critical operational visibility to IT and business stakeholders.
  • Analyze monitoring failures and implement corrective actions.
  • Serve as a technical escalation point for monitoring and integration-related incidents.

Governance, Risk & Compliance
  • Support monitoring solutions that are subject to SOX, audit, and compliance requirements.
  • Maintain documentation, runbooks, operational procedures, and technical standards.
  • Participate in audit activities and evidence collection when required.
  • Ensure monitoring solutions align with cybersecurity and operational standards.

Leadership & Soft Skills
  • Strong customer service and stakeholder management skills
  • Exceptional troubleshooting and analytical abilities
  • Ability to communicate effectively with technical and non-technical audiences
  • Ability to prioritize multiple projects and operational demands
  • Self-directed and highly accountable
  • Strong collaboration skills across infrastructure, security, cloud, and application teams
  • Ability to lead technical initiatives and mentor junior engineers
  • Strategic mindset with the ability to align monitoring capabilities with business outcomes

Reference: 1061927
Worried that you don’t meet every single requirement listed in the job ad? Studies have shown that individuals from marginalized groups are less likely to apply to jobs unless they meet every single qualification. Hive + Co. is dedicated to building a diverse, inclusive and representative workplace, so if you’re excited about this role, but worried that you don’t meet every requirement, we encourage you to apply anyways. We’d love to get to know you.