1

Site Reliability Engineer Jobs in Colorado (NOW HIRING)

Software Site Reliability Engineer

Golden, CO · On-site

$58.75 - $78.25/hr

Job Title Software Site Reliability Engineer As the Site Reliability Engineer, you will support CoorsTek's Databricks application and data product strategy by ensuring solutions built, migrated, and ...

Software Site Reliability Engineer

Golden, CO · On-site

$58.75 - $78.25/hr

Job Title Software Site Reliability Engineer As the Site Reliability Engineer, you will support CoorsTek's Databricks application and data product strategy by ensuring solutions built, migrated, and ...

Software Site Reliability Engineer

Golden, CO · On-site

$103.04 - $136.01/hr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

Databricks Site Reliability EngineerAs the Site Reliability Engineer, you will support CoorsTek's Databricks application and data product strategy by ensuring solutions built, migrated, and deployed ...

Staff Site Reliability Engineer

Greenwood Village, CO · On-site

$57.75 - $76.75/hr

  • Medical

  • Dental

  • Vision

  • PTO

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work on the platform. As a Staff Site ...

Staff Site Reliability Engineer

Greenwood Village, CO · On-site

$57.75 - $76.75/hr

  • Medical

  • Dental

  • Vision

  • PTO

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work on the platform. As a Staff Site ...

Principal Site Reliability Engineer

Denver, CO · On-site

$58.75 - $78/hr

We are seeking a Principal Site Reliability Engineer to define the strategic vision and own the enterprise-wide reliability, scalability, and performance of our critical production services. As a ...

We are seeking a Principal Site Reliability Engineer to define the strategic vision and own the enterprise-wide reliability, scalability, and performance of our critical production services. As a ...

Site Reliability Engineer

Centennial, CO

$104K - $156K/yr

  • Medical

  • Dental

  • Life

  • Retirement

  • PTO

Our Site Reliability Engineering (SRE) team is at the heart of our mission to deliver seamless and robust services to our users. We're a distributed team of engineers with diverse skillsets who ...

Site Reliability Engineer - TS/SCI with Poly

Colorado Springs, CO · On-site

$56.25 - $74.75/hr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Yes SITE RELIABILITY ENGINEER (SRE) Own your opportunity. Make your impact As a Site Reliability Engineer (SRE) supporting the CIO Infrastructure Services (CIS) program, you will help maintain the ...

Site Reliability Engineer

Aurora, CO · On-site

$130K - $200K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Site Reliability Engineer Position ID: TA247 Location(s): On-site; Aurora, CO; Herndon, VA Application Deadline: August 31, 2026 Security Clearance Requirement: TS/SCI Security Clearance with ...

Principal Site Reliability Engineer

Denver, CO · On-site

$58.75 - $78/hr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

We are seeking a Principal Site Reliability Engineer to define the strategic vision and own the enterprise-wide reliability, scalability, and performance of our critical production services. As a ...

Site Reliability Engineer

Centennial, CO · On-site

$104.43 - $156.65/hr

  • Medical

  • Dental

  • Life

  • Retirement

  • PTO

Our Site Reliability Engineering (SRE) team is at the heart of our mission to deliver seamless and robust services to our users. We're a distributed team of engineers with diverse skillsets who ...

Sr. Site Reliability Engineer

Denver, CO · On-site

$58.75 - $78/hr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is accountable for the full ...

Site Reliability Engineer

Centennial, CO · On-site

$58.50 - $78/hr

  • Medical

  • Dental

  • Life

  • Retirement

  • PTO

Our Site Reliability Engineering (SRE) team is at the heart of our mission to deliver seamless and robust services to our users. We're a distributed team of engineers with diverse skillsets who ...

Sr. Site Reliability Engineer

Denver, CO · Hybrid

$58.75 - $78/hr

We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is accountable for the full ...

Sr. Site Reliability Engineer

Denver, CO · Hybrid

$58.75 - $78/hr

We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is accountable for the full ...

Site Reliability Engineer, Senior

Aurora, CO · On-site

$86K - $198K/yr

  • Medical

  • Life

  • Retirement

  • PTO

Site Reliability Engineer, Senior The Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in ...

Showing results 21-40

Site Reliability Engineer information

See Colorado salary details

$11

$67

$96

How much do site reliability engineer jobs pay per hour?

As of Aug 13, 2026, the average hourly pay for site reliability engineer in Colorado is $67.03, according to ZipRecruiter salary data. Most workers in this role earn between $57.64 and $76.59 per hour, depending on experience, location, and employer.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often involves on-call duties, troubleshooting complex issues, and working with automation tools, which can contribute to work-related stress but also offers opportunities for skill development and problem-solving.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.

What are the most commonly searched types of Site Reliability Engineer jobs in Colorado?

The most popular types of Site Reliability Engineer jobs in Colorado are:

What are popular job titles related to Site Reliability Engineer jobs in Colorado?

For Site Reliability Engineer jobs in Colorado, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer jobs in Colorado look for?

The top searched job categories for Site Reliability Engineer jobs in Colorado are:

What cities in Colorado are hiring for Site Reliability Engineer jobs?

Cities in Colorado with the most Site Reliability Engineer job openings:

What are popular job titles related to Site Reliability Engineer jobs in CO?

For Site Reliability Engineer jobs in CO, the most frequently searched job titles are:

Infographic showing various Site Reliability Engineer job openings in Colorado as of August 2026, with employment types broken down into 1% As Needed, 83% Full Time, 13% Part Time, 1% Temporary, and 2% Contract. Highlights an 94% Physical, 3% Hybrid, and 3% Remote job distribution, with an average salary of $139,413 per year, or $67 per hour.

Software Site Reliability Engineer

CoorsTek, Inc.

Golden, CO • On-site

$58.75 - $78.25/hr

Full-time

Re-posted 10 days ago


CoorsTek rating

8.2

Company rating: 8.2 out of 10

Based on 28 frontline employees who took The Breakroom Quiz


Job description

It's exciting to work for a company that makes the world measurably better.
We're committed to bringing safety, quality, and customer focus to the business of advanced ceramics manufacturing.
Job Title
Software Site Reliability Engineer
As the Site Reliability Engineer, you will support CoorsTek's Databricks application and data product strategy by ensuring solutions built, migrated, and deployed on Databricks are reliable, secure, observable, supportable, and cost-effective in production. This role is not solely focused on monitoring and operational support. In this role, you will actively develop automation, platform tooling, deployment pipelines, observability capabilities, and reliability solutions that reduce operational toil and improve the scalability of Databricks-hosted applications and data products.
This role sits within Data & Analytics and partners closely with Architecture, Cybersecurity, Infrastructure, Manufacturing IT/OT, Enterprise Applications, citizen developers, and business teams. As the Databricks Site Reliability Engineer, you will support production reliability for Databricks-hosted applications (pattern B), analytics products, workflows, and AI-enabled solutions.
In this role, you will help CoorsTek move quickly without creating unmanaged technical debt by contributing to and improving support patterns, monitoring standards, deployment practices, runbooks, incident response, and operational guardrails for Databricks solutions created by both business-enabled citizen development and IT delivery teams.
Roles and Responsibilities
  • Support production reliability, operational readiness, and lifecycle support for Databricks-hosted applications, data products, dashboards, notebooks, jobs, workflows, APIs, and AI-enabled solutions.
  • Support applications migrated to Databricks, built directly in Databricks, or promoted from citizen development and IT development into governed production patterns.
  • Execute intake, review, handoff, support, and release practices for Pattern B Databricks applications, including minimum requirements before production deployment.
  • Partner with citizen developers, IT developers, data engineers, enterprise architects, and business stakeholders to convert prototypes into reliable, monitored, documented, and supportable services.
  • Implement and maintain observability standards, including logging, alerting, health checks, SLIs/SLOs, lineage, usage monitoring, cost monitoring, and operational dashboards.
  • Respond to incidents, coordinate troubleshooting, participate in root cause analysis and support corrective actions for failed jobs, broken pipelines, access issues, performance issues, data refresh failures, and application outages.
  • Maintain and update runbooks, support procedures, escalation paths, ownership models, service catalogs, and knowledge articles for Databricks applications and data products.
  • Partner with Data & Analytics on Databricks workflows, Delta Lake, Unity Catalog, data lineage, permissions, SQL warehouses, jobs, clusters, serverless capabilities, and performance tuning.
  • Partner with Cybersecurity and Architecture to ensure Databricks solutions meet standards for identity, access, secrets management, logging, data classification, responsible AI, and least-privilege access.
  • Support CI/CD, testing, environment promotion, release controls, rollback procedures, and change management for Databricks applications and related Azure or integration components.
  • Identify recurring failure patterns and assist with automating manual support work, reducing operational toil, and creating reusable templates and standards.
  • Advise teams on production-ready design, including resiliency, scalability, maintainability, cost control, data quality checks, monitoring hooks, and clear ownership.
  • Collaborate with manufacturing, finance, supply chain, quality, and other business teams to understand impact, prioritize recovery, and maintain trust in critical Databricks-supported solutions.
  • Support governance for citizen-built solutions by ensuring business-created applications have appropriate documentation, testing evidence, security review, support model, and IT transition plan before broad use.
  • Monitor and problem solve service health, support metrics, incidents, problem records, platform risks, and improvement backlog items for Databricks applications and data products.
  • Design and develop automation, self-healing workflows, monitoring integrations, and operational tooling using Python and cloud-native technologies.

Job Requirements
Education:
  • Bachelor's degree in Computer Science, Information Technology, Data Engineering, Software Engineering, Systems Engineering, or a related field required.
  • Master's degree preferred.

Experience:
  • 5 or more years of progressive experience in site reliability engineering, data platform engineering, cloud operations, DevOps, software engineering, data engineering, or production application support.
  • 3 or more years supporting cloud, data, analytics, application, or platform services in production environments preferred.
  • Experience with Databricks, Delta Lake, Unity Catalog, SQL, Python, PySpark, notebooks, jobs/workflows, SQL warehouses, clusters, or lakehouse architecture.
  • Experience operating applications through incident management, problem management, change management, monitoring, release management, and production readiness practices.
  • Preferred experience with Azure, CI/CD pipelines, Git-based development, infrastructure patterns, logging, alerting, automation, and support runbooks.
  • Preferred experience supporting data pipelines, analytics products, dashboards, APIs, AI-enabled applications, or business-critical reporting environments.

Functional / Technical Knowledge, Skills & Abilities:
  • Strong understanding of SRE, DevOps, IT operations, and production support practices, including reliability, observability, automation, incident response, and operational excellence.
  • Working knowledge of Databricks platform capabilities, including Delta tables, notebooks, workflows/jobs, SQL, Unity Catalog, lineage, permissions, compute configuration, and governed access patterns.
  • Ability to troubleshoot Databricks jobs, pipelines, notebooks, SQL queries, permissions, data refreshes, performance issues, and environment or integration failures.
  • Ability to write and review SQL and Python; PySpark, scripting, API, and automation experience preferred.
  • Ability to define operational readiness standards for applications created by citizen developers, IT teams, consultants, and data engineering teams.
  • Strong understanding of monitoring, alerting, logging, service health, SLOs, runbooks, release controls, rollback planning, and root cause analysis.
  • Ability to balance speed, business enablement, cybersecurity, supportability, cost control, and long-term platform sustainability.
  • Ability to partner effectively with Data & Analytics, Cybersecurity, Architecture, Infrastructure, Enterprise Applications, Manufacturing IT/OT, and business stakeholders.
  • Strong documentation and communication skills, including support models, knowledge articles, architecture notes, production checklists, escalation paths, and operational dashboards.
  • Ability to manage multiple production priorities, operate calmly during incidents, drive follow-through on corrective actions, and influence teams without direct authority.

Preferred Certifications:
  • Relevant Databricks certifications, including Data Engineer, Data Analyst, Machine Learning, or Lakehouse Fundamentals preferred.
  • Relevant Microsoft Azure, DevOps, cloud engineering, cybersecurity, ITIL, SRE, observability, or data engineering certifications preferred.
  • ITIL Foundation, Azure Administrator, Azure Developer, GitHub, Terraform, Kubernetes, or related platform operations certifications are a plus.

Additional Position Details:
  • Location: Golden, CO (on-site)
  • Work Authorization: Requires U.S. Person status (U.S. citizen, a Green Card holder, or a protected refugee/asylee)

Target Hiring Range
Annual Salary: USD 115,000.00 - USD 155,000.00
Actual compensation is commensurate with experience, skills and education. CoorsTek strives to give all qualified applicants equal opportunity and to make selection decisions on job related factors. Do not provide any information on the application which will indicate your race, color, religion, national origin, sex, age, disability, sexual orientation, gender identity, pregnancy, genetic information, veteran status, or any other status protected by law or regulation.
If you like working for a company that makes a real difference in the world, you'll enjoy your career with us!

What CoorsTek employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom