1

Observability Aiops Engineer Jobs in Florida (NOW HIRING)

DevOps Engineer (AIOps)

Jacksonville, FL · On-site

$49 - $67/hr

The AIOps Engineer is a technical operational engineering role responsible for improving DevOps, ... Improve observability and operational visibility through dashboards, monitoring integrations ...

... engineering Infrastructure and managed services Integration / middleware modernization ServiceNow, observability, AIOps, automation, and security services Map client org structure, technology ...

SRE Lead

Miami, FL · On-site

$54.50 - $72.25/hr

Tata Consultancy Services is seeking an SRE Lead who will provide strategic and technical ... modern observability and cloud infrastructure. Preferred : • Familiarity with modern AIOps ...

Observability, AIOps, APM; Industry leading discovery technologies (SCCM, Tanium, Armis, Intune) and how they integrate with ServiceNow; Developing and re-engineering IT processes, capabilities, and ...

Observability, AIOps, APM; Industry leading discovery technologies (SCCM, Tanium, Armis, Intune) and how they integrate with ServiceNow; Developing and re-engineering IT processes, capabilities, and ...

Lead ML Ops Engineer

Orlando, FL

$95K - $126K/yr

Define and execute an enterprise AI/ML platform strategy, encompassing MLOps, LLMOps, and AIOps ... Deep understanding of AI system observability, including drift detection, evaluation frameworks ...

Lead ML Ops Engineer

Naples, FL

$96K - $127K/yr

Define and execute an enterprise AI/ML platform strategy, encompassing MLOps, LLMOps, and AIOps ... Deep understanding of AI system observability, including drift detection, evaluation frameworks ...

.NET Solution Architect

Doral, FL · On-site

$58.25 - $76.75/hr

... developer productivity accelerators. * Exposure to AIOps and intelligent monitoring frameworks ... Experience with observability and monitoring tools such as: * Datadog * Dynatrace * Knowledge of ...

next page

Showing results 1-20

Observability Aiops Engineer information

What are some common challenges faced by Observability AIOps Engineers in integrating monitoring solutions across diverse technology stacks?

Observability AIOps Engineers often encounter challenges when integrating monitoring and analytics tools across a mix of legacy systems, cloud-native applications, and various third-party platforms. Ensuring consistent data collection, normalization, and visualization can be complex due to differing protocols, data formats, and tool compatibility. Collaboration with development, operations, and security teams is crucial to address these challenges, streamline workflows, and maintain a unified observability platform. Staying current with evolving AIOps technologies and best practices is also vital for continued success in this dynamic role.

What is an Observability Aiops Engineer?

An Observability Aiops Engineer is a technology professional who focuses on implementing and managing observability tools and practices, often leveraging artificial intelligence for IT operations (AIOps). Their role is to ensure system reliability, performance, and uptime by monitoring, analyzing, and automating responses to IT incidents. They integrate data from logs, metrics, and traces to gain real-time insights, helping organizations quickly detect and resolve issues. This role combines expertise in software engineering, monitoring solutions, automation, and machine learning to improve the overall health and efficiency of IT environments.

What are the key skills and qualifications needed to thrive as an Observability AIOps Engineer, and why are they important?

To thrive as an Observability AIOps Engineer, you need expertise in systems monitoring, data analytics, automation, and a strong understanding of IT infrastructure, often supported by a degree in computer science or a related field. Familiarity with tools like Prometheus, Grafana, ELK stack, Splunk, and AIOps platforms, as well as certifications in cloud solutions (AWS, Azure, or GCP), are typically required. Strong problem-solving skills, collaboration, and a proactive mindset help you stand out in identifying and addressing system anomalies. These skills and qualities are crucial for maintaining high system reliability, reducing downtime, and enabling data-driven decision-making in complex IT environments.

What is the difference between Observability Aiops Engineer vs Site Reliability Engineer?

AspectObservability Aiops EngineerSite Reliability Engineer
Primary FocusMonitoring, analyzing, and improving system observability using AI and automationEnsuring system reliability, scalability, and performance of services
Skills & CertificationsKnowledge of AI/ML, monitoring tools, scripting, cloud platformsSystems engineering, scripting, cloud infrastructure, incident management
Work EnvironmentDevOps teams, monitoring platforms, AI toolsOperations, development teams, cloud environments
Industry UsageTech companies, cloud providers, organizations focusing on AI-driven monitoringLarge-scale tech firms, SaaS providers, internet services

While both roles focus on system performance and reliability, the Observability Aiops Engineer specializes in leveraging AI and automation to enhance system observability, whereas the Site Reliability Engineer concentrates on maintaining overall system stability and scalability. Both roles often collaborate but have distinct core responsibilities.

What are popular job titles related to Observability Aiops Engineer jobs in Florida? For Observability Aiops Engineer jobs in Florida, the most frequently searched job titles are:
What job categories do people searching Observability Aiops Engineer jobs in Florida look for? The top searched job categories for Observability Aiops Engineer jobs in Florida are:
What cities in Florida are hiring for Observability Aiops Engineer jobs? Cities in Florida with the most Observability Aiops Engineer job openings:
DevOps Engineer (AIOps)

DevOps Engineer (AIOps)

Allegis Group

Jacksonville, FL • On-site

$49 - $67/hr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Posted 10 days ago


Job description

Overview

Job Summary:

The AIOps Engineer is a technical operational engineering role responsible for improving DevOps, Production Support, Release Management, Test Automation, and platform operations through intelligent automation, operational tooling, observability, and AI-enabled capabilities. This role focuses on building scalable operational solutions that improve system reliability, reduce manual operational effort, strengthen production visibility, and optimize engineering support workflows across the Connected and OBE programs.

The AIOps Engineer partners closely with Engineering, DevOps, Production Support, Architecture, and Release teams to identify improvement opportunities, streamline operational processes, and implement practical solutions across monitoring, incident management, deployment operations, operational analytics, release coordination, and support workflows. This role contributes to implementing operational tooling, automation frameworks, intelligent monitoring capabilities, and AI-assisted solutions that improve operational efficiency, engineering effectiveness, and platform stability.

This is a hands-on engineering role focused on applying modern operational engineering practices to solve enterprise operational challenges through automation, operational intelligence, and scalable engineering solutions. The AIOps Engineer helps improve operational scalability, support responsiveness, operational maturity, and engineering efficiency while aligning enterprise technology and governance standards. 

Required in-office presence at least 2 days per week

Responsibilities

Essential Functions:

  • Design, develop, and maintain intelligent operational solutions, automation workflows, scripts, integrations, and orchestration capabilities across DevOps, Production Support, Release Management, Test Automation, and operational ecosystems.
  • Identify operational bottlenecks, repetitive support activities, and manual engineering processes suitable for automation and operational optimization.
  • Implement AI-enabled operational capabilities focused on incident visibility, monitoring, alert correlation, workflow automation, operational analytics, and engineering efficiency.
  • Improve observability and operational visibility through dashboards, monitoring integrations, alerting, logging, analytics, reporting, and release insight solutions across enterprise platforms.
  • Partner with Engineering, DevOps, Architecture, Production Support, and Release teams to improve operational scalability, engineering effectiveness, and platform reliability.
  • Support operational triage, incident coordination, root cause analysis, and operational response workflows through automation and intelligent operational tooling.
  • Evaluate, implement, and enhance operational tooling, integrations, automation frameworks, orchestration capabilities, and operational intelligence solutions.
  • Support release operations, deployment visibility, environment coordination, and operational readiness activities across enterprise platforms.
  • Contribute to operational standards, implementation best practices, governance alignment, and modernization initiatives.
  • Assist with operational data analysis, trend identification, operational metrics, engineering productivity insights, and reporting initiatives.
  • Collaborate with cross-functional teams to implement scalable operational processes aligned with enterprise technology and security standards.
  • Provide technical recommendations related to automation, reliability, scalability, observability, and operational optimization opportunities.
  • Support continuous improvement initiatives focused on operational maturity, engineering efficiency, support effectiveness, and platform stability.

Supervisory or Management Responsibility:

  • Provide technical collaboration and operational support across DevOps, Engineering, Production Support, Release Management, and operational teams.
  • Share operational knowledge, automation practices, and implementation guidance with engineering and operational teams.
  • Assist leadership and engineering teams with operational insights, technical recommendations, and operational improvement opportunities.

Budget Responsibility:

  • Support operational tooling evaluations, technical assessments, and operational platform recommendations.
  • Assist with operational tooling research and operational modernization initiatives aligned with long-term engineering and operational goals.
Qualifications

Minimum Education and/or Experience:

  • Bachelor's degree in computer science, Information Systems, Engineering, or equivalent combination of education and work experience.
  • 4+ years of experience in DevOps, Site Reliability Engineering (SRE), Platform Engineering, Production Support Engineering, Operational Engineering, Automation Engineering, or enterprise operational support environments.
  • Strong understanding of enterprise operational workflows, operational support models, and platform reliability practices.
  • Experience with CI/CD pipelines, release orchestration, environment management, and DevOps operational processes.
  • Experience with observability, monitoring, alerting, logging, and operational analytics platforms.
  • Experience with automation frameworks, scripting, workflow automation, tooling integrations, and orchestration capabilities.
  • Experience supporting cloud-based enterprise environments and distributed operational ecosystems.
  • Familiarity with AI-assisted operational tooling, intelligent automation platforms, operational analytics, or operational intelligence capabilities.
  • Understanding of incident management, operational triage, root cause analysis, and production support workflows.
  • Experience integrating APIs, operational services, enterprise platforms, and automation pipelines.
  • Strong communication, stakeholder collaboration, analytical thinking, and problem-solving skills.
  • Experience operating within Agile enterprise delivery environments.
Skills and Abilities:
  • Experience supporting Salesforce enterprise ecosystems and Salesforce DevOps environments.
  • Experience with GitHub Actions, release orchestration platforms, infrastructure automation, or operational workflow tooling.
  • Familiarity with AI-assisted operational tooling, operational intelligence platforms, or enterprise automation frameworks.
  • Experience supporting enterprise-scale production support operations and operational modernization initiatives.
  • Experience with operational analytics, engineering productivity tooling, intelligent monitoring solutions, or operational dashboards.
  • Familiarity with cloud-native operational tooling, enterprise observability platforms, and operational automation practices.

Core Competencies:

  • Build relationships
  • Develop people
  • Lead change
  • Inspire Others
  • Think critically
  • Communicate clearly
  • Create Accountability

Benefits Overview:

Benefits are subject to change and may be subject to specific elections, plan, or program terms.  This role is eligible for the following:

  • Medical, dental & vision
  • Hospital plans
  • 401(k) Retirement Plan - Pre-tax and Roth post-tax contributions available
  • Life Insurance (Company paid Basic Life and AD&D as well as voluntary Life & AD&D for the employee and dependents)
  • Company paid Short and long-term disability
  • Health & Dependent Care Spending Accounts (HSA & DCFSA)
  • Transportation benefits
  • Employee Assistance Program
  • Tuition Assistance
  • Time Off/Leave (PTO, Allegis Group Paid Family Leave, Parental Leave)

Salary Range:

  • $73,600.00 - $110,400.00
  • This position is bonus eligible
  • Individual compensation offered for this position within this range will depend on many factors, including qualifications, skills, relevant experience, job knowledge, geographic location, internal equity, and other pertinent job-related factors.

The company is an equal opportunity employer and will consider all applications without regard to race, sex, age, color, religion, national origin, veteran status, disability, sexual orientation, gender identity, genetic information or any characteristic protected by law. 

If you would like to request a reasonable accommodation, such as the modification or adjustment of the job application process or interviewing process due to a disability, please email Lauren Lara at llara@allegisgroup.com or call 410-579-3526 for other accommodation options.

Employment Type: FULL_TIME