1

Observability Jobs in Colorado (NOW HIRING)

Acquire-Site Reliability Engineer

Denver, CO ยท On-site

$90K - $105K/yr

Close HIPAA-aware observability gaps: PHI-safe logging, auditability, access controls, incident evidence. * DevOps and release engineering: operate and improve our GitHub Actions deploy pipelines ...

Data Engineer II

Denver, CO ยท On-site

$108K - $136K/yr

Apply and improve data quality, testing, observability, and lineage standards * Collaborate with cross-functional partners to define data contracts and interfaces * Contribute to Capital Rx's modular ...

Software Engineer Principal

Arvada, CO

$135K - $181K/yr

Linux, Windows Server, PowerShell, Python, automation, scripting, platform engineering, configuration management, drift management, drift detection, observability, monitoring, Elastic, Dynatrace ...

DevOps Engineer

Denver, CO ยท Remote

$54.25 - $74.25/hr

Improve platform observability through logging, monitoring, metrics, alerting, and dashboards. * Optimize systems for performance, scalability, availability, security, and cost efficiency. * Partner ...

DevOps Engineer

Denver, CO ยท On-site

$54.25 - $74.25/hr

Improve platform observability through logging, monitoring, metrics, alerting, and dashboards. * Optimize systems for performance, scalability, availability, security, and cost efficiency. * Partner ...

Sr. Database Administrator

Denver, CO ยท On-site

$51.25 - $70.50/hr

Implement monitoring and observability using CloudWatch, Dynatrace, Grafana, and custom automation. * Establish PostgreSQL standards, governance, and operational best practices. * Provide ...

Senior DevOps Engineer

Greenwood Village, CO ยท On-site

$131K - $169K/yr

Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. * Ensure infrastructure and automation ...

Provide thought leadership on emerging trends in cloud-native technologies, platform engineering, automation, and observability; bring point-of-view to clients on the partner and tooling ecosystem.

Establish a core observability foundation--including real-time metrics, logging, crash telemetry, and field analytics--to improve application reliability across all deployed customer devices. * Drive ...

Sr DevOps Engineer

Denver, CO ยท On-site

$133K - $171K/yr

The role requires deep expertise in deployment automation, HPA tuning, release reliability, networking, observability, and rollback strategies for large-scale microservices platforms handling high ...

DevOps Engineer

Denver, CO

$54.25 - $74.25/hr

Improve platform observability through logging, monitoring, metrics, alerting, and dashboards. * Optimize systems for performance, scalability, availability, security, and cost efficiency. * Partner ...

Showing results 41-60

Observability information

See Colorado salary details

$17

$63

$90

How much do observability jobs pay per hour?

As of Aug 11, 2026, the average hourly pay for observability in Colorado is $63.65, according to ZipRecruiter salary data. Most workers in this role earn between $53.32 and $73.03 per hour, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive in an observability role?

To thrive in an Observability role, you need a strong background in monitoring, alerting, logging, and analyzing system performance, often supported by a degree in computer science or related field. Familiarity with tools such as Prometheus, Grafana, Datadog, Splunk, and experience with cloud platforms and scripting languages is crucial. Excellent problem-solving, communication, and collaboration skills help you work effectively with cross-functional engineering and operations teams. These capabilities are essential to ensure system reliability, quickly detect issues, and maintain seamless digital experiences.

What is an observability?

An Observability job focuses on ensuring the performance, reliability, and health of software systems by collecting, analyzing, and visualizing telemetry data such as logs, metrics, and traces. Professionals in this field work with monitoring tools, distributed tracing, and alerting systems to detect and troubleshoot issues proactively. They collaborate with engineering and operations teams to improve system visibility, reduce downtime, and enhance overall system performance.

Is observability a good career?

Observability is a growing field within IT and software engineering that involves monitoring, logging, and analyzing system performance to ensure reliability and efficiency. It often requires skills in tools like Prometheus, Grafana, and data analysis, making it a valuable and in-demand career path with opportunities for advancement. The role typically involves collaboration across teams and continuous learning to keep up with evolving technologies.

What does an observability do?

In an Observability role, your daily tasks often include designing and maintaining monitoring dashboards, configuring alerts, analyzing system logs, and working closely with development and operations teams to troubleshoot issues. You'll proactively identify areas of improvement to increase system reliability, document monitoring strategies, and support incident response efforts. Collaboration is key, as you may participate in post-incident reviews and help drive architectural improvements based on the data you collect. The role is dynamic and requires a proactive approach to ensure systems stay healthy and downtime is minimized.

What are the most commonly searched types of Observability jobs in Colorado? The most popular types of Observability jobs in Colorado are:
What are popular job titles related to Observability jobs in Colorado? For Observability jobs in Colorado, the most frequently searched job titles are:
What cities in Colorado are hiring for Observability jobs? Cities in Colorado with the most Observability job openings:
Infographic showing various Observability job openings in Colorado as of August 2026, with employment types broken down into 92% Full Time, 3% Part Time, and 5% Contract. Highlights an 74% Physical, 8% Hybrid, and 18% Remote job distribution, with an average salary of $132,395 per year, or $63.7 per hour.

Acquire-Site Reliability Engineer

BehaviorSpan

Denver, CO โ€ข On-site

$90K - $105K/yr

Full-time

Medical, PTO

Re-posted 16 days ago


Job description

Site Reliability EngineerAbout Acquire Learning

Acquire Learning is a learning management platform built for ABA (Applied Behavior Analysis) therapy. Clinicians and behavior technicians use it daily with clients on the autism spectrum, and the data it captures shapes real treatment decisions. We are a small, product-focused team in a HIPAA-regulated environment. When a clinician is mid-session with a child, the platform behaving predictably is the difference between productive therapy and a disrupted session. Your work affects the communities we serve.

About the Role

Acquire is hiring its first dedicated Site Reliability Engineer, a mid-level role with a clear path to Lead SRE as we grow. You will own production health, a trustworthy release pipeline, and the reliability surface of the codebase, while raising release-quality risk and acting as the customer-facing escalation point. You report to the Lead Engineer and work regularly with the CEO and CTO. You will not inherit a mature SRE team or a thick runbook library; you will help build them.


Just as important as the technical background: we want a product-focused thinker. AI tooling makes raw implementation cheaper, so the scarce skill is judgment about why the product matters to clinicians and what actually needs building. Treat reliability and ops as ways to keep a great product healthy, and step into feature work when the team needs it.


This is a broad role today by design. As the team grows it narrows toward Lead SRE: reliability strategy, incident response, and how ops, observability, and release engineering work at Acquire.

What You'll Do
  • Production reliability: own day-to-day health across AWS and MongoDB Atlas. Triage and respond to alerts (CloudWatch, Sentry, Google Chat ops-alerts), run root-cause and incident comms, and turn retros into runbooks and alerting improvements. Close HIPAA-aware observability gaps: PHI-safe logging, auditability, access controls, incident evidence.

  • DevOps and release engineering: operate and improve our GitHub Actions deploy pipelines (backend, webapp, native), maintain Terraform infrastructure and review infra PRs for safety, improve CI signal quality, coordinate mobile releases via TestFlight and Google Play, and harden rollback, restore, and break-glass paths.

  • Reliability-focused code (TypeScript/Node): repair scripts, migrations and index management, observability instrumentation, tenant-scoped operational tooling, background job lifecycle, deploy tooling, and e2e (Playwright) and integration tests. This is your primary lane, but with agentic tooling we expect you to help on product features when it counts rather than treating "that's not infra" as a boundary.

  • Release quality and QA: validate release candidates, walk core clinical workflows on web, iOS, and Android, run our test suites (Playwright, Jest, Postman), and drive release checklists and post-release verification.

  • Customer escalations: be the first internal contact for customer-reported issues. Reproduce, isolate, document, prioritize by clinical impact, and own the loop back to the customer.

Where We Need Help

HIPAA-aware operations, multi-tenant architecture (tenant isolation, org-safe migrations and diagnostics), disaster recovery and restore confidence, data-integrity operations, security and production-access hygiene (IAM, secrets, least privilege), incident-response maturity (severity levels, SLOs, alerting standards), and scale and cost visibility across AWS and MongoDB.

Who You Are
  • 3-5 years that meaningfully includes SRE, DevOps, production, platform, or infrastructure engineering.

  • AWS proficiency is a hard requirement. Real, hands-on experience operating production workloads on AWS, and you can talk through it in specifics.

  • A product-focused thinker who asks why the product matters and what needs building, not only how the infrastructure runs, and who can contribute to feature work with agentic tooling when needed.

  • Self-sufficient and a fast ramper. Dropped into an unfamiliar codebase, cloud account, or toolchain, you have a real strategy to get productive on your own, and you should expect to demonstrate it live during interviews.

  • Curious about ABA, our product, and AI. Genuine interest in the clinical work Acquire supports is preferred, and we want people who are eager to learn the domain and to explore how AI can move it forward.

  • Honest about your background, with recent, checkable professional references.

  • Comfortable with CI/CD, Terraform or comparable IaC, MongoDB or another production database, and observability tooling (CloudWatch, Datadog, Sentry, Grafana).

  • Fluent enough in Node/TypeScript for reliability work: repair scripts, migrations, instrumentation, job lifecycle, deploy tooling, and e2e/integration tests.

  • At home on the command line, in production logs, and in cloud consoles, with real incident experience you can talk through.

  • Careful about production access, customer data, and tenant boundaries, and clear on the difference between a helpful diagnostic and an accidental data leak.

  • Comfortable on a small team where some process exists, some needs creating, and everyone stays close to the product.

Using AI Tools

We use AI where it genuinely helps. You do not need to be an AI expert, and we are not looking for someone who treats AI as a substitute for judgment. We want someone comfortable with tools like Claude or Cursor to move faster (log triage, drafting repair scripts and tests, investigating support cases, verification checklists, working past the edge of their expertise) while still checking the result like an engineer. If a tool creates uncertainty, know when to slow down and verify.

Bonus Points

Healthcare or other HIPAA-regulated experience; multi-tenant SaaS (tenant isolation, support tooling, data repair); disaster recovery, SLOs, or incident command; familiarity with our stack (AWS ECS/CloudFront/IAM/CloudWatch/SNS, MongoDB Atlas, Terraform, GitHub Actions, Node, TypeScript, Playwright, Sentry, ClickUp); React/React Native/Express in production and monorepos (pnpm, Turborepo); native release coordination; and prior work introducing SRE/DevOps/QA practices on a small team.

Path to Lead SRE

In the first months you learn the product, infrastructure, and highest-impact reliability surfaces, take over the alert-response loop, become a reliable hand on deploys and incidents, and start contributing to reliability code. Over the first year you take on reliability strategy, observability and post-incident hygiene, release engineering, and the operational and QA playbooks that keep Acquire trustworthy as it scales. The goal is Lead SRE at Acquire. Comp is revisited as you grow into that scope.

Compensation & Benefits
  • Base salary: $90,000-$105,000, revisited as the role grows into Lead SRE ownership.

  • Semi-annual performance reviews with raise potential.

  • Location: Denver, CO. We strongly prefer candidates who are already local; this is an in-person role to start.

  • Fully in office for your first 90 days. Being present while you ramp on the product, systems, and team is how you build the context this job depends on. Hybrid flexibility (1-2 days WFH) comes afterward as a privilege earned through demonstrated ownership, not a day-one default.

  • Healthcare, PTO, and additional benefits.