1

Observability Aiops Engineer Jobs (NOW HIRING)

Aiops engineer

Buffalo, NY ยท On-site

$94K - $129K/yr

Azure AI Gateway ยท AI Platform ยท AIOps Engineer ร  3 Roles | Start: Immediate What They'll Build ... Instrument observability -- LLM call tracing, latency/cost dashboards via Azure Monitor & App ...

AIOps Engineer

Fort Belvoir, VA ยท On-site

$190K - $218K/yr

Mission-Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified ... Engineer secure data ingestion pipelines for telemetry data from cross-domain solutions and ...

JCS Solutions LLC is seeking a Senior AIOps Engineer to support critical mission operations within ... all observability tools comply with DoW STIGs and IL5/IL6 protocols; develop and maintain ...

Mission-Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified ... Engineer secure data ingestion pipelines for telemetry data from cross-domain solutions and ...

Mission-Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified ... Engineer secure data ingestion pipelines for telemetry data from cross-domain solutions and ...

AIOps Engineer SME

Albuquerque, NM

$80K - $106K/yr

Mission-Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified ... Engineer secure data ingestion pipelines for telemetry data from cross-domain solutions and ...

AIOps Engineer SME

Fort Belvoir, VA

$93K - $124K/yr

Mission-Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified ... Engineer secure data ingestion pipelines for telemetry data from cross-domain solutions and ...

AIOps Engineer SME

Fort Belvoir, VA ยท On-site

$131K - $237K/yr

Mission-Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified ... Engineer secure data ingestion pipelines for telemetry data from cross-domain solutions and ...

AIOps Engineer SME

Fort Belvoir, VA ยท On-site

$131K - $237K/yr

Mission-Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified ... Engineer secure data ingestion pipelines for telemetry data from cross-domain solutions and ...

AIOps Engineer SME

Albuquerque, NM ยท On-site

$131K - $237K/yr

Mission-Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified ... Engineer secure data ingestion pipelines for telemetry data from cross-domain solutions and ...

AIOps Engineer SME

Albuquerque, NM ยท On-site

$131K - $237K/yr

Mission-Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified ... Engineer secure data ingestion pipelines for telemetry data from cross-domain solutions and ...

AIOps Engineer

Miami, FL ยท On-site

$54.50 - $72.50/hr

Role Overview The AIOps Engineer operates in a dual capacity, designing and building AI-powered ... Advanced use of monitoring and observability tools such as Splunk or New Relic. * Experience with ...

The Senior Observability Engineer is responsible for defining, leading, and advancing enterprise ... Apply AIOps and agentic operations capabilities to strengthen detection, event correlation ...

AIOps Engineer

Miami, FL ยท On-site

$54.50 - $72.50/hr

* SRE & Automation: 6-8 years of experience as a Site Reliability Engineer (SRE) or DevOps Engineer ... Observability: Advanced hands-on experience with Splunk or New Relic for monitoring dashboards ...

AI Ops Engineer

Charlotte, NC ยท On-site

$51.50 - $70.50/hr

Observability/AIOPs Event Management Engineer/Sr. Developer: * Expert in defining new monitoring definitions based on achieving best practice monitoring capabilities. Infrastructure (mandatory ...

next page

Showing results 1-20

Observability Aiops Engineer information

What is an Observability AIOps engineer?

An Observability Aiops Engineer is a technology professional who focuses on implementing and managing observability tools and practices, often leveraging artificial intelligence for IT operations (AIOps). Their role is to ensure system reliability, performance, and uptime by monitoring, analyzing, and automating responses to IT incidents. They integrate data from logs, metrics, and traces to gain real-time insights, helping organizations quickly detect and resolve issues. This role combines expertise in software engineering, monitoring solutions, automation, and machine learning to improve the overall health and efficiency of IT environments.

What are the key skills and qualifications needed to thrive as an Observability AIOps engineer?

To thrive as an Observability AIOps Engineer, you need expertise in systems monitoring, data analytics, automation, and a strong understanding of IT infrastructure, often supported by a degree in computer science or a related field. Familiarity with tools like Prometheus, Grafana, ELK stack, Splunk, and AIOps platforms, as well as certifications in cloud solutions (AWS, Azure, or GCP), are typically required. Strong problem-solving skills, collaboration, and a proactive mindset help you stand out in identifying and addressing system anomalies. These skills and qualities are crucial for maintaining high system reliability, reducing downtime, and enabling data-driven decision-making in complex IT environments.

What are some common challenges faced by Observability AIOps engineers in integrating monitoring solutions across diverse technology stacks?

Observability AIOps Engineers often encounter challenges when integrating monitoring and analytics tools across a mix of legacy systems, cloud-native applications, and various third-party platforms. Ensuring consistent data collection, normalization, and visualization can be complex due to differing protocols, data formats, and tool compatibility. Collaboration with development, operations, and security teams is crucial to address these challenges, streamline workflows, and maintain a unified observability platform. Staying current with evolving AIOps technologies and best practices is also vital for continued success in this dynamic role.

What is the difference between Observability Aiops Engineer vs Site Reliability Engineer?

AspectObservability Aiops EngineerSite Reliability Engineer
Primary FocusMonitoring, analyzing, and improving system observability using AI and automationEnsuring system reliability, scalability, and performance of services
Skills & CertificationsKnowledge of AI/ML, monitoring tools, scripting, cloud platformsSystems engineering, scripting, cloud infrastructure, incident management
Work EnvironmentDevOps teams, monitoring platforms, AI toolsOperations, development teams, cloud environments
Industry UsageTech companies, cloud providers, organizations focusing on AI-driven monitoringLarge-scale tech firms, SaaS providers, internet services

While both roles focus on system performance and reliability, the Observability Aiops Engineer specializes in leveraging AI and automation to enhance system observability, whereas the Site Reliability Engineer concentrates on maintaining overall system stability and scalability. Both roles often collaborate but have distinct core responsibilities.

More about Observability Aiops Engineer jobs

What cities are hiring for Observability Aiops Engineer jobs?

Cities with the most Observability Aiops Engineer job openings:

What states have the most Observability Aiops Engineer jobs?

States with the most job openings for Observability Aiops Engineer jobs include:

What are popular job titles related to Observability Aiops Engineer jobs?

For Observability Aiops Engineer jobs, the most frequently searched job titles are:

Infographic showing various Observability Aiops Engineer job openings in the United States as of September 2026, with employment types broken down into 1% Internship, 87% Full Time, 9% Part Time, and 3% Contract. Highlights an 86% Physical, 4% Hybrid, and 10% Remote job distribution.

Aiops engineer

Buffalo, NY โ€ข On-site

$94K - $129K/yr

Other

Posted 11 days ago


Job description

Azure AI Gateway ยท AI Platform ยท AIOps Engineer ร  3 Roles | Start: Immediate

What They'll Build

  • Stand up M&T's Azure AI Gateway (APIM, token quotas, cost attribution, multi-model routing)
  • Establish the AI Platform foundation โ€” Azure OpenAI, AI Foundry, model registry, fine-tuning pipelines
  • Build AIOps: automated monitoring, drift alerting, self-healing pipelines for production AI workloads
  • Instrument observability โ€” LLM call tracing, latency/cost dashboards via Azure Monitor & App Insights
  • Embed guardrails: content filtering, PII redaction, prompt injection controls aligned to M&T risk standards

Must-Have Skills

  • Azure OpenAI ยท Azure AI Foundry ยท APIM ยท AKS โ€” hands-on, not conceptual
  • AIOps tooling: MLflow, PrometheGrafana, Azure Monitor, drift detection
  • IaC (Terraform or Bicep) + CI/CD (Azure DevOps / GitHub Actions)
  • Python + LangChain or Semantic Kernel
  • 5+ yrs engineering; 2+ yrs Azure AI/ML in regulated or financial environments