2

Full Time Observability Jobs in Chicago, IL (NOW HIRING)

Miratech retains nearly 1000 full-time professionals, and our annual growth rate exceeds 25%. The ... Monitor UC platform health using observability and logging tools (Splunk). * Manage and analyze ...

Miratech retains nearly 1000 full-time professionals, and our annual growth rate exceeds 25%. The ... Monitor UC platform health using observability and logging tools (Splunk). * Manage and analyze ...

Senior DevOps Engineer

Chicago, IL ยท On-site +1

$160K - $200K/yr

Improve observability, monitoring, and production reliability. * Lead troubleshooting, incident ... full-time employees. Pay Transparency The expected pay range for this position is $160,000 to $200 ...

Lead DevOps Engineer

Chicago, IL ยท On-site

$54.25 - $74.50/hr

Chicago, IL (Hybrid) Employment Type: Full-Time Overview A financial services organization based in ... Observability & Performance Monitoring * Implement monitoring and observability solutions using ...

We are seeking a full-time, remote Principal AI Architect. The Principal AI Architect role provides ... Ensure AI solutions are designed for performance, scalability, observability, privacy, and ...

We are seeking a full-time, remote Principal AI Architect. The Principal AI Architect role provides ... Ensure AI solutions are designed for performance, scalability, observability, privacy, and ...

Hybrid in Chicago (IL) Employment Type : Full-time FLSA Classification : Exempt Start Date : ASAP ... Experience with performance testing, observability (e.g., logs/metrics), and test data management.

Software Engineer II

Chicago, IL ยท On-site

$100K - $137K/yr

Experience working with telemetry data to support observability, monitoring, and debugging in ... At least 18 paid time off days annually for full-time employees and 6 company holidays per year ...

Showing results 21-40

Full Time Observability information

See Chicago, IL salary details

$16

$62

$88

How much do full time observability jobs pay per hour?

As of Aug 23, 2026, the average hourly pay for full time observability in Chicago, IL is $62.36, according to ZipRecruiter salary data. Most workers in this role earn between $52.26 and $71.59 per hour, depending on experience, location, and employer.

What is a full time observability engineer?

A Full Time Observability Engineer is a professional who specializes in ensuring the health, performance, and reliability of software systems by implementing monitoring, logging, and tracing solutions. Their primary focus is to provide visibility into complex applications and infrastructure, enabling teams to detect and resolve issues quickly. They work with tools like Prometheus, Grafana, ELK Stack, and others to collect and analyze data from across the tech stack. This role is essential in modern DevOps and Site Reliability Engineering (SRE) teams, helping organizations maintain high service availability and improve incident response.

What are the key skills and qualifications needed to thrive as a full time observability engineer?

To thrive as a Full-Time Observability Engineer, you need a solid background in systems monitoring, troubleshooting, and infrastructure automation, often supported by a degree in computer science or a related field. Proficiency with observability tools such as Prometheus, Grafana, ELK stack, Datadog, or New Relic, and experience with cloud platforms and scripting languages, is typically required. Strong analytical thinking, attention to detail, and effective communication skills help engineers identify issues and collaborate across teams. These skills and qualities are crucial for maintaining system reliability, quickly resolving incidents, and enabling continuous improvement in complex technical environments.

What are some typical challenges faced by professionals in a full time observability role, and how can they be addressed?

Professionals in a Full Time Observability role often encounter challenges such as managing large volumes of monitoring data, ensuring the integration of multiple observability tools, and maintaining system performance while collecting detailed metrics. Addressing these challenges typically involves automating data collection, standardizing on a core set of observability platforms, and collaborating closely with development and operations teams to fine-tune monitoring coverage. Continuous learning and staying updated on the latest observability technologies are also key to overcoming these hurdles and ensuring system reliability.

What is the difference between Full Time Observability vs Full Time Monitoring?

AspectFull Time ObservabilityFull Time Monitoring
FocusComprehensive system insights, including metrics, logs, and tracesTracking predefined metrics and alerts
ToolsOpen-source and advanced observability platforms (e.g., Prometheus, Grafana)Monitoring tools like Nagios, Zabbix
Work EnvironmentDevOps, SRE teams, cloud environmentsIT operations, infrastructure teams
CertificationsOften requires knowledge of cloud, Linux, scriptingIT certifications, network, and system administration

Full Time Observability involves a broader approach to understanding system health through metrics, logs, and traces, enabling proactive troubleshooting. Full Time Monitoring typically focuses on tracking specific metrics and alerts to detect issues. While monitoring is a subset of observability, the two roles differ in scope and tools used.

Is full time observability a good career?

Full time observability is a growing field focused on monitoring and analyzing system performance using tools like Prometheus, Grafana, and data analysis skills. It offers opportunities in IT, DevOps, and software engineering with demand for technical expertise and certifications. The role typically involves working in fast-paced environments and requires continuous learning of new tools and technologies.

What are the most commonly searched types of Observability jobs in Chicago, IL?

The most popular types of Observability jobs in Chicago, IL are:

What job categories do people searching Full Time Observability jobs in Chicago, IL look for?

The top searched job categories for Full Time Observability jobs in Chicago, IL are:

Infographic showing various Full Time Observability job openings in Chicago, IL as of July 2026, with employment types broken down into 92% Full Time, 3% Part Time, and 5% Contract. Highlights an 78% Physical, 6% Hybrid, and 16% Remote job distribution, with an average salary of $129,704 per year, or $62.4 per hour.

Platform Engineering & AI Operations Lead

Intelligent Generation

Oak Brook, IL โ€ข On-site

$103K - $136K/yr

Full-time

Re-posted 14 days ago


Job description

Platform Engineering & AI Operations Lead

Full Time | Hybrid | Chicago Metro Area

Build the platform foundation for POWR:Suite and IG’s AI-assisted operating model
Intelligent Generation’s mission is to empower businesses to engage the clean energy grid. 
Intelligent Generation builds and operates POWR:Suite, a software platform that helps battery energy storage assets make highly profitable economic decisions.
POWR:Suite connects distributed energy assets to wholesale power markets while also optimizing behind-the-meter value: reducing utility bills, managing demand charges, improving asset performance, supporting resilience, and helping customers capture the full economic value of their energy assets.
Our work sits at the intersection of energy markets, grid operations, customer savings, software automation, telemetry, and AI-assisted decision-making.
We are looking for a platform engineering leader who can own the infrastructure foundation behind POWR:Suite and help IG scale an AI-assisted engineering and operating model.
This is a leadership role. You will be hands-on early, but the expectation is that you will grow into leading people, establishing platform standards, and orchestrating AI agents that improve engineering delivery, reliability, security, and operations.
Why this role matters
Battery assets are only valuable when they are operated intelligently. Every decision matters: when to charge, when to discharge, when to participate in the market, when to preserve state of charge, when to reduce customer demand charges, when to support resilience, and how to prove the economic value created.
The platform behind POWR:Suite must be reliable, observable, secure, automated, and ready to support both human operators and AI-assisted workflows.
You will help define that foundation.

What you will lead and own
Cloud platform and infrastructure
Lead the architecture and operation of IG’s cloud platform on GCP, including Cloud Run, GKE, Cloud SQL, Pub/Sub, IAM, Terraform, networking, observability, and deployment standards.
CI/CD and engineering enablement
Build and lead the practices that make engineering delivery faster, safer, and more repeatable. Own CI/CD, release automation, environment strategy, deployment quality, and platform guardrails.
Edge-to-cloud reliability
Strengthen the communication layer between field assets and cloud systems. Partner with software, controls, and operations teams to improve telemetry, command paths, failure detection, and recovery patterns.
Security and operational controls
Own practical security and IT operations standards for a company operating live energy infrastructure, including IAM lifecycle, endpoint standards, secrets management, access controls, and production system hygiene.
Agent operations foundation
Build, maintain, evaluate, and govern agents that support platform engineering work, including deployment assistants, infrastructure review agents, security review agents, incident summarizers, documentation agents, and operational troubleshooting agents.
People and agent orchestration
Over time, build and lead a platform engineering function. Establish how work is divided between engineers and agents, how agent outputs are reviewed, and how platform knowledge compounds over time.

What success looks like
First 90 days
  • Understand current platform architecture, cloud services, deployments, edge/cloud interfaces, and operational pain points
  • Establish a platform ownership map and priority risk list
  • Improve documentation around environments, deployment flow, access, and operational dependencies
  • Identify high-value opportunities for agent-assisted platform work
First 6 months
  • Improve CI/CD maturity and release reliability
  • Strengthen observability and alerting for critical platform services
  • Build or deploy early internal agents that assist with platform review, troubleshooting, documentation, or deployment support
  • Establish platform standards that engineering teams can follow

First 12 months
  • Lead the platform function for POWR:Suite
  • Improve engineering independence, reliability, and operational control
  • Mature agent-assisted engineering workflows
  • Build the foundation for a team that can scale with IG’s growth
What we are looking for
Required

  • 8+ years in platform engineering, infrastructure engineering, DevOps, SRE, cloud architecture, or senior software engineering
  • Experience leading technical work across teams or mentoring engineers
  • Strong production GCP experience
  • Terraform or infrastructure-as-code experience
  • CI/CD ownership in production environments
  • Strong Python or equivalent engineering/scripting ability
  • Strong understanding of distributed systems, reliability, observability, and operational failure modes
  • Security-first mindset around IAM, access control, secrets, production systems, and operational risk
  • Hands-on experience using AI tools as part of engineering work
  • Ability to build, maintain, evaluate, govern, or orchestrate agents that support engineering workflows
  • Strong systems thinking and ownership mindset
Strongly preferred
  • Energy, industrial, IoT, SCADA, or OT-adjacent experience
  • Familiarity with DNP3, Modbus, ICCP, or similar protocols
  • Experience with agentic AI systems, tool use, evals, or AI-assisted operations
  • GCP Professional certification
  • Experience building or leading a platform engineering team
Why this is exciting
You will help define the platform foundation for a company operating at the intersection of energy storage, virtual power plants, cloud infrastructure, market automation, customer savings, and AI-assisted engineering.
This is a chance to build systems, standards, agents, and eventually a team that will shape how IG scales POWR:Suite and optimizes the full economic value of distributed battery assets.
This position is hybrid and includes both remote and in-person work.
Intelligent Generation participates in the E-Verify process for all new hires.