1

Site Reliability Engineer Manager Jobs in Tooele, UT

Site Reliability Engineer

Draper, UT · On-site

$53.25 - $70.75/hr

They are seeking a hands-on Site Reliability Engineer to bridge the gap between software ... Experience managing high-throughput message queues or data streaming platforms, specifically ...

Reliability Engineer

Salt Lake City, UT

$98K - $123K/yr

Job Title Reliability Engineer Summary Reliability, Maintenance, and Engineering (RME) is hiring ... Support the Maintenance Manager in the implementation of short and long-term projects for the site ...

Reliability Engineer

Salt Lake City, UT · On-site

$99K - $125K/yr

Spare Parts Management: Support maintenance planning by identifying critical spare parts and ... Reliability engineering certification is a plus (e.g., Certified Maintenance & Reliability ...

next page

Showing results 1-20

Site Reliability Engineer Manager information

See Tooele, UT salary details

$10

$59

$86

How much do site reliability engineer manager jobs pay per hour?

As of Aug 23, 2026, the average hourly pay for site reliability engineer manager in Tooele, UT is $59.84, according to ZipRecruiter salary data. Most workers in this role earn between $51.44 and $68.37 per hour, depending on experience, location, and employer.

What is a site reliability engineer manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.

How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How much do site reliability engineer managers get paid?

Site Reliability Engineer Managers typically earn between $120,000 and $180,000 annually, depending on experience, location, and company size. They often oversee teams responsible for system reliability, incident response, and infrastructure automation, requiring strong leadership and technical skills.

Is a Site Reliability Engineer Manager a stressful job?

A Site Reliability Engineer Manager role can be stressful due to the responsibility of maintaining system uptime, managing incident responses, and ensuring reliability across complex infrastructure. The job often involves working under pressure, handling outages, and coordinating teams, but it also offers opportunities for problem-solving and leadership. Stress levels vary depending on company size, team structure, and workload management skills.

What cities near Tooele, UT are hiring for Site Reliability Engineer Manager jobs?

Cities near Tooele, UT with the most Site Reliability Engineer Manager job openings:

Site Reliability Engineer

ProdataKey

Draper, UT • On-site

$53.25 - $70.75/hr

Full-time

Re-posted 28 days ago


Job description

Job Summary:
ProdataKey is a leading innovator of cloud-based access control products and services. They are seeking a hands-on Site Reliability Engineer to bridge the gap between software development and infrastructure operations, ensuring the reliability and performance of their platform.
Responsibilities:
• Design, build, and maintain scalable, secure multi-tenant cloud infrastructure using Infrastructure as Code (IaC) principles.
• Own the availability, latency, performance, and capacity planning of the pdk.io platform and its supporting backend microservices.
• Develop and manage robust monitoring, logging, and alerting systems to gain deep visibility into cloud infrastructure, API health, and IoT endpoint performance.
• Participate in a collaborative on-call rotation. Lead rapid incident response mitigation and drive rigorous, blameless post-mortems to ensure long-term system resilience.
• Optimize and secure automated deployment pipelines to enable developers to ship code to production safely and efficiently.
• Partner closely with backend developers and hardware engineering teams to define Service Level Indicators (SLIs), Service Level Objectives (SLOs), and manage error budgets.
• Contribute to technical documentation and knowledge sharing.
Qualifications:
Required:
• Bachelor’s or Master’s degree in Computer Science, Information Systems, or related fields or equivalent experience
• 3+ years of experience in an SRE, DevOps, or Infrastructure Engineering role supporting production cloud environments
• Cloud Architecture: Deep hands-on experience with major cloud providers (AWS or GCP) and a strong command of containerization and orchestration technologies (Docker and Kubernetes).
• Coding & Scripting: Strong programming proficiency in languages such as Python, Go, TypeScript, or Bash for automation, internal tooling, and system integrations.
• Systems & Networking: Solid fundamentals in Linux/Unix administration, networking protocols (TCP/IP, DNS, HTTP/S, load balancing), and cloud security best practices.
• Mindset: A passionate problem-solver who prioritizes automation over manual operations and thrives in high-ownership environments.
• Must pass drug and criminal background check
• Work well in an onsite team environment
Preferred:
• Experience with multi-region deployments, failover strategies, and data consistency
• Messaging Systems: Experience managing high-throughput message queues or data streaming platforms, specifically RabbitMQ or Apache Kafka.
• Experience operating production systems at scale
• Data Infrastructure: Familiarity with modern data stack environments, such as Snowflake or relational databases in a self hosted environment.
• Certifications: Relevant industry certifications such as Certified Kubernetes Administrator (CKA) or AWS Certified DevOps Engineer Professional.
• Familiarity with regulatory requirements (SOC2, GDPR, etc)
• Experience in physical security or access control systems
• Familiarity with GCP ecosystem and tooling
• Experience working in a scaling startup environment
Company:
ProdataKey provides cloud-based access control and SaaS services. Founded in 2011, the company is headquartered in Draper, USA, with a team of 51-200 employees. The company is currently Growth Stage.