1

Google Site Reliability Engineer Jobs in Utah (NOW HIRING)

Sr. Observability Engineer

Lehi, UT ยท On-site

$98K - $134K/yr

Google SRE practices: toil elimination, incident management, automation for self-healing * Cross-functional influence without authority. You've improved teams that don't report to you * Governance ...

Reliability Engineer

Salt Lake City, UT

$98K - $123K/yr

Job Title Reliability Engineer Summary Reliability, Maintenance, and Engineering (RME) is hiring ... Lead the Root Cause Analysis and Permanent Corrective Action activities at the site. * Identify ...

Platform Engineer (DevOps)

Salt Lake City, UT ยท On-site

$51 - $70/hr

... Reliability Engineer (SRE), Cloud Engineer, Pipeline Engineer, or in a similar enterprise ... Google Cloud Platform (preferred)** or AWS * Working knowledge of both **Linux and Windows ...

Principal Reliability Engineer

Provo, UT

$97K - $122K/yr

US-AZ-TUCSON-801 ~ 1151 E Hermans Rd ~ BLDG 801 (External Site) Position Role Type: Onsite U.S ... Coach other Reliability engineers and work with functional leadership in identifying and ...

Showing results 21-40

Google Site Reliability Engineer information

See Utah salary details

$9

$58

$83

How much do google site reliability engineer jobs pay per hour?

As of Aug 8, 2026, the average hourly pay for google site reliability engineer in Utah is $58.03, according to ZipRecruiter salary data. Most workers in this role earn between $49.90 and $66.30 per hour, depending on experience, location, and employer.

What is a Google Site Reliability Engineer?

Google Site Reliability Engineers (SREs) are specialized engineers responsible for ensuring the reliability, scalability, and performance of Google's infrastructure and services. They bridge the gap between software development and operations by automating tasks, monitoring systems, and responding to incidents. SREs work to minimize downtime, manage large-scale distributed systems, and implement best practices for reliability. Their work involves a combination of software engineering, system administration, and process improvement to support Google's products and users.

How does a Google Site Reliability Engineer typically collaborate with software development teams?

Google Site Reliability Engineers (SREs) work closely with software development teams to ensure that systems are reliable, scalable, and efficient. They participate in design reviews, provide feedback on system architecture with an emphasis on reliability, and help define service level objectives (SLOs) and monitoring standards. SREs often engage in incident response and post-mortem analysis alongside developers, fostering a culture of shared responsibility for uptime and performance. This collaboration helps bridge the gap between development and operations, promoting continuous improvement and rapid innovation.

What is the difference between Google Site Reliability Engineer vs DevOps Engineer?

AspectGoogle Site Reliability EngineerDevOps Engineer
CredentialsTypically requires computer science or engineering degree, experience with coding, systems, and cloud platformsOften has similar technical background, certifications like AWS, Azure, or Linux are common
Work EnvironmentFocuses on maintaining large-scale systems, reliability, and automation at GoogleWorks across development and operations teams to streamline deployment and infrastructure
Industry UsagePrimarily used within Google and similar large tech companiesWidely adopted across various industries for infrastructure and deployment automation

Google Site Reliability Engineers and DevOps Engineers share many skills, including scripting, cloud knowledge, and system management. However, SREs focus more on system reliability and scalability within large-scale environments like Google, while DevOps Engineers emphasize continuous integration, deployment, and collaboration across development and operations teams.

What are the key skills and qualifications needed to thrive as a Google Site Reliability Engineer, and why are they important?

To thrive as a Google Site Reliability Engineer, you need a solid background in computer science, software engineering, and systems administration, often backed by a relevant degree or equivalent experience. Familiarity with programming languages (such as Python, Go, or Java), cloud platforms, configuration management tools, and monitoring systems is essential. Strong problem-solving abilities, effective communication, and a proactive mindset set exceptional SREs apart. These skills are crucial for ensuring the reliability, scalability, and efficiency of large-scale infrastructure and services.
What are popular job titles related to Google Site Reliability Engineer jobs in Utah? For Google Site Reliability Engineer jobs in Utah, the most frequently searched job titles are:
What job categories do people searching Google Site Reliability Engineer jobs in Utah look for? The top searched job categories for Google Site Reliability Engineer jobs in Utah are:
What cities in Utah are hiring for Google Site Reliability Engineer jobs? Cities in Utah with the most Google Site Reliability Engineer job openings:
Infographic showing various Google Site Reliability Engineer job openings in Utah as of August 2026, with employment types broken down into 1% As Needed, 82% Full Time, 12% Part Time, 1% Temporary, and 4% Contract. Highlights an 94% Physical, 3% Hybrid, and 3% Remote job distribution, with an average salary of $120,700 per year, or $58 per hour.

Site Reliability Engineer (Contractor)

Varo Bank

Salt Lake City, UT โ€ข On-site

$55.25 - $73.25/hr

Full-time

This job post hasย expired 1 day ago.ย Applications are no longer accepted.


Job description

Varo is an entirely new kind of bank. All digital, mission-driven, FDIC insured and designed for the way our customers live their lives. A bank for all of us.

About the role

Varoโ€™s SRE team is well established, designing, building, and running large-scale, distributed, fault-tolerant systems that power most of Varo\'s operations. We live and breathe AWS and Kubernetes, having an open source first and result oriented mindset.

We are an automation and observability focused team and we strive to automate ourselves out of manual / remedial tasks. We monitor and create dashboards to promote a data-driven approach to scale out our platform.

On a typical day, members of our team are hands-on scaling-out production infrastructure, building out CI/CD pipelines, and brainstorming with developers on how to make things better. We collectively strive to build and maintain a rapid-feedback platform that enables our engineers to accomplish their own goals instead of creating friction.

Responsibilities
  • EKS & Karpenter Management: Manage, upgrade, and autoscale EKS clusters across multiple environments (SIT, UAT, Prod) and AWS accounts.

  • Infrastructure as Code & GitOps: Write Terraform modules and Helm charts to support GitOps workflows using ArgoCD and GitLab CI/CD pipelines.

  • Kafka & Data Platform Support: Maintain and troubleshoot Kafka (MSK) clusters, including broker health, connectors, and CDC pipelines.

  • Observability & Cost Control: Improve observability using Prometheus, Thanos, Grafana, and ELK while proactively identifying cloud cost-optimization opportunities.

  • Automation & AIOps: Automate operational tasks with Python and leverage AI/ML techniques for predictive alerting and intelligent runbooks.

  • Service Desk & Support: Handle Platform Service Desk requests, including Terraform merge request reviews, access management, and deployment support.

  • Incident Response & On-Call: Participate in the production on-call rotation, support incident response, and contribute to blameless post-mortems.

Skills & Qualification
  • Experience: 3+ years of experience in an SRE, DevOps, or Infrastructure Engineering role, with the ability to work independently and manage multiple workstreams.

  • AWS Expertise: Strong hands-on experience with core AWS services, including EKS, EC2, RDS Aurora, MSK, S3, IAM, VPC, and Direct Connect.

  • Kubernetes & Deployment: Deep production experience with Kubernetes (upgrades, networking, RBAC) alongside Helm and GitOps tools like ArgoCD.

  • Infrastructure as Code: Advanced proficiency with Terraform, including writing modules and managing multi-account/multi-environment states.

  • Data Infrastructure: Experience supporting and maintaining data platforms such as Airflow, Databricks, EMR, Kafka/MSK, or CDC pipelines.

  • Networking & Automation: Solid understanding of networking (VPCs, security groups, Istio, DNS) paired with strong Python scripting skills for tooling and automation.

  • Observability & AI: Experience managing observability stacks (Prometheus, Grafana, ELK) and effectively leveraging AI/LLM tools for automation and incident analysis.

  • Nice to Have: Experience with Karpenter and KEDA, GitLab CI/CD pipeline experience, Hashicorp Vault for secrets management.

For cash compensation, we set standard ranges for all US-based roles based on function, level, and geographic location, benchmarked against similar-stage growth companies. Final offer amounts are determined by multiple factors as well as candidate experience and expertise and may vary from the identified range.

Varo is an equal opportunity employer. Varo embraces diversity and we are committed to building teams that represent a variety of backgrounds, perspectives, and skills. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status.

Beware of fraudulent job postings!

Varo will never ask for payment to process documents, refer you to a third party to process applications or visas, or ask you to pay costs. Never send money to anyone suggesting they can provide work with Varo.  If you suspect you have received a phony offer, please e-mail careers@varomoney.com with the pertinent information and contact information.

Notice at Collection for Employees and Applicants:

Varo Candidate Privacy Notice