1

Site Reliability Engineer Manager Jobs in Spokane, WA

Our Site Reliability Engineer should help keep our systems steady, secure, and running like a well-oiled machine (except without actual oil). You'll work closely with our DevOps engineers to build ...

... management system backed by Heroic Customer Service ® and support. Our on-site and cloud-based ... As SRE you will be responsible for maintaining system reliability, automating operations, and ...

... management system backed by Heroic Customer Service and support. Our on-site and cloud-based ... As SRE you will be responsible for maintaining system reliability, automating operations, and ...

Automation Engineer

Spokane, WA · On-site

$70K - $125K/yr

Our secret is what we put into it--innovative thinking, industry-leading reliability, and a world ... Engineering Manager. This position supports/leads flow team directed Manufacturing Excellence ...

... management disciplines. Core responsibilities include site layout, grading, stormwater analysis ... The role offers mentorship from senior engineers while also mentoring junior staff, with growing ...

next page

Showing results 1-20

Site Reliability Engineer Manager information

See Spokane, WA salary details

$10

$64

$92

How much do site reliability engineer manager jobs pay per hour?

As of Aug 23, 2026, the average hourly pay for site reliability engineer manager in Spokane, WA is $64.45, according to ZipRecruiter salary data. Most workers in this role earn between $55.43 and $73.65 per hour, depending on experience, location, and employer.

What is a site reliability engineer manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.

How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How much do site reliability engineer managers get paid?

Site Reliability Engineer Managers typically earn between $120,000 and $180,000 annually, depending on experience, location, and company size. They often oversee teams responsible for system reliability, incident response, and infrastructure automation, requiring strong leadership and technical skills.

Is a Site Reliability Engineer Manager a stressful job?

A Site Reliability Engineer Manager role can be stressful due to the responsibility of maintaining system uptime, managing incident responses, and ensuring reliability across complex infrastructure. The job often involves working under pressure, handling outages, and coordinating teams, but it also offers opportunities for problem-solving and leadership. Stress levels vary depending on company size, team structure, and workload management skills.

What are the most commonly searched types of Site Reliability Engineer jobs in Spokane, WA?

The most popular types of Site Reliability Engineer jobs in Spokane, WA are:

What cities near Spokane, WA are hiring for Site Reliability Engineer Manager jobs?

Cities near Spokane, WA with the most Site Reliability Engineer Manager job openings:

Site Reliability Engineer

Corporate Tools

Post Falls, ID • On-site, Remote

$175K/yr

Full-time

Medical, Dental, Vision, Retirement, PTO

Re-posted 13 days ago


Job description

Overview
Corporate Tools is looking for a Site Reliability Engineer. You will be a traditional company employee. This is a remote position, but if you're near one of our local offices, you're welcome to come hangout with us in-office as well. Our main offices are in Post Falls, ID, and Spokane, WA; we also have satellite offices in Austin, TX, and Salt Lake City, UT. You'll be working 40 hours a week and, of course, enjoy great company benefits. Our Site Reliability Engineer should help keep our systems steady, secure, and running like a well-oiled machine (except without actual oil). You'll work closely with our DevOps engineers to build out tools and automation that make things faster, easier, and less painful for everyone.
Your main job? Stop problems before they start. And when something does break (because let's be real-it will), help us fix it quickly and learn from it so we don't do the same dumb thing twice. We're big on taking ownership here. You won't get blamed for something going wrong-but you will be expected to help make it right.
If you like digging into weird errors, thinking ahead, and making things just work-even when no one notices-this might be your kind of thing.
Wage
Up to $175,000 / year
Benefits
  • 100% employer-paid medical, dental and vision for employees
  • Annual review with raise option
  • 22 days Paid Time Off accrued annually, and 4 holidays
    • After 3 years, PTO increases to 29 days. Employees transition to flexible time off after 5 years with the company-not accrued, not capped, take time off when you want
    • The 4 holidays are: New Year's Day, Fourth of July, Thanksgiving, and Christmas Day
  • Paid Parental Leave
  • Up to 6% company matching 401(k) with no vesting period
  • Quarterly allowance
    • Use to make your remote work set up more comfortable, for continuing education classes, a plant for your desk, coffee for your coworker, a massage for yourself... really, whatever
  • Open concept office with friendly coworkers
  • Creative environment where you can make a difference
  • No dumb benefits like free dog walking on the weekends that snobby hipster places have to make you feel cool, but mathematically won't cost the company much money because you won't use it
  • Trail Mix Bar - oh yeah

Requirements
  • Bachelor's degree in Computer Science, Software Engineering, or equivalent practical experience.
  • 5+ years of experience in software engineering.
  • 2+ years of experience in site reliability engineering, DevOps, or infrastructure engineering roles.
  • Deep experience with cloud platforms (AWS, Azure, or GCP) and infrastructure as code tools such as Terraform, CloudFormation, or Pulumi.
  • Strong proficiency with Kubernetes, Docker, and container orchestration in production environments.
  • Hands-on experience with observability and monitoring tools like Prometheus, Grafana, OpenTelemetry, Sentry, or New Relic.
  • Proven ability to design and implement highly available, fault-tolerant systems and lead proactive incident response efforts.
  • Experience with performance tuning, database optimization, and caching strategies (e.g., PostgreSQL, Redis, Memcached).
  • Demonstrated ability to drive reliability improvements, reduce operational toil, and foster a culture of resilience and continuous improvement.
  • Experience leading reliability-focused initiatives such as post-incident reviews, capacity planning, and root cause analysis.
  • Experience in site reliability engineering within Ruby on Rails environments.
  • Familiarity with the Grafana observability stack and related tools (e.g., Alloy, Loki, Tempo, Prometheus).
  • In-depth experience with AWS services, including ECS, EKS, Route 53, and other related tools.
  • Proven ability to collaborate across teams to improve service reliability, reduce incident frequency, and drive operational excellence.
  • Troubleshoot and resolve complex production issues, applying SRE best practices to minimize impact and prevent recurrence.