1

Site Reliability Engineer Contract Jobs in Utah (NOW HIRING)

Reliability Engineer

Salt Lake City, UT ยท On-site

$98K - $123K/yr

... site operation's needs. * Manage subcontractors and suppliers to deliver goods and services against contracts and expectations * Support the Maintenance Manager in the implementation of short and ...

Principal Reliability Engineer

Provo, UT ยท On-site

$97K - $122K/yr

US-AZ-TUCSON-801 ~ 1151 E Hermans Rd ~ BLDG 801 (External Site) Position Role Type: Onsite U.S ... Coach other Reliability engineers and work with functional leadership in identifying and ...

Showing results 21-40

Site Reliability Engineer Contract information

What does a site reliability engineer contract do?

A Site Reliability Engineer (SRE) contractor is responsible for ensuring the reliability, scalability, and performance of software systems, typically on a temporary or project basis. They collaborate with development and operations teams to automate processes, monitor systems, and quickly resolve incidents. SRE contractors often design and implement tools that improve system uptime and efficiency, and may create documentation or best practices for site reliability. Their work helps organizations maintain high service availability while adapting to changing infrastructure needs.

What are the key skills and qualifications needed to thrive as a site reliability engineer contract?

Thriving as a Site Reliability Engineer Contractor requires strong expertise in systems administration, cloud platforms, automation, and programming (typically in Python, Go, or Bash), usually supported by a degree in computer science or relevant experience. Familiarity with DevOps tools like Kubernetes, Docker, Terraform, CI/CD pipelines, and monitoring solutions such as Prometheus or Grafana is essential, with certifications like AWS Certified DevOps Engineer or Google Professional SRE adding value. Exceptional problem-solving, collaboration, and communication skills help contractors quickly adapt to new environments and efficiently resolve incidents. These skills and qualities ensure reliable system performance, rapid response to outages, and seamless integration with client teams.

What are some common challenges faced by site reliability engineer contracts?

Site Reliability Engineers (SREs) on contract often face the challenge of quickly adapting to new systems, tooling, and organizational cultures. Since contracts are typically for a limited duration, contractors need to rapidly build rapport with internal teams and familiarize themselves with existing infrastructure to make impactful contributions. Additionally, they may need to balance multiple priorities, such as incident response, automation, and documentation, while ensuring that their work aligns with both short-term project goals and long-term site reliability. Effective communication and proactive collaboration with development and operations teams are essential for overcoming these challenges and delivering value within the contract period.

What is the difference between Site Reliability Engineer Contract vs Site Reliability Engineer?

AspectSite Reliability Engineer ContractSite Reliability Engineer
CredentialsTypically requires SRE or related certifications, experience with cloud platformsSame as contract, often with more emphasis on full-time experience
Work EnvironmentProject-based, temporary, often remote or on-siteFull-time, ongoing, may be remote or on-site
Employer UsageUsed by companies for specific projects or to fill short-term needsUsed as a core role within organizations for continuous reliability management
Search & Comparison IntentCommonly compared for contract vs full-time roles in SREOften compared to contract roles for career planning

In summary, a Site Reliability Engineer Contract is a temporary, project-based role focusing on specific reliability tasks, while a full-time Site Reliability Engineer is a permanent position with ongoing responsibilities. Both roles require similar skills and certifications, but differ mainly in employment type and work setup.

What are the most commonly searched types of Site Reliability Engineer jobs in Utah?

The most popular types of Site Reliability Engineer jobs in Utah are:

What job categories do people searching Site Reliability Engineer Contract jobs in Utah look for?

The top searched job categories for Site Reliability Engineer Contract jobs in Utah are:

What cities in Utah are hiring for Site Reliability Engineer Contract jobs?

Cities in Utah with the most Site Reliability Engineer Contract job openings:

Infographic showing various Site Reliability Engineer Contract job openings in Utah as of July 2026, with employment types broken down into 1% As Needed, 79% Full Time, 16% Part Time, 1% Temporary, and 3% Contract. Highlights an 94% Physical, 2% Hybrid, and 4% Remote job distribution.

Manager, Site Reliability Engineering

1 O.C. Tanner Company

Salt Lake City, UT โ€ข On-site

$120 - $180/hr

Other

Posted 25 days ago


Key responsibilities

  • Lead, mentor, and develop a team of Site Reliability Engineers to foster a culture of reliability and operational excellence.

  • Define and execute the organization's reliability strategy to improve availability, scalability, and resilience through automation and best practices.

  • Oversee production triage, incident response, and escalation processes to ensure timely service restoration and effective root cause analysis.


Job description

O.C. Tanner is the global leader in software and services that improve workplace culture through meaningful employee experiences. Our Culture Cloud is a suite of apps designed to enhance the employee experience with strategic recognition, service awards, wellbeing, leadership, and events that help people thrive at work. Our Culture by Design approach provides expert services to organizations looking to create great workplaces. Our global team of 1,500 people hail from 58 countries and speak 62 languages. As programmers, researchers, designers, client professionals and craftspeople we create the tech, tools and awards that connect employees to purpose at thousands of companies. Join us as we help people all over the world thrive at work.

Location: Salt Lake City, UT

As the Manager of Site Reliability Engineering, you will lead the strategy, execution, and evolution of reliability for our world-class employee recognition platform. You will build, mentor, and empower a team of Site Reliability Engineers while partnering closely with Engineering, Product, and Support organizations to deliver highly available, scalable, and resilient services that serve millions of users. We are seeking a leader who is passionate about operational excellence, continuous improvement, and fostering a reliability-first culture through automation, observability, and shared ownership. In this role, you will champion the development of self-healing platforms, drive incident and operational maturity, and enable engineering teams to innovate faster while delivering exceptional customer experiences.

Key Responsibilities
  • Lead, mentor, and develop a team of Site Reliability Engineers, fostering a culture of reliability, accountability, operational excellence, and continuous improvement.
  • Define and execute the organization's reliability strategy, improving availability, scalability, performance, and resilience through automation and engineering best practices.
  • Establish team priorities, goals, and success metrics aligned with business objectives, customer needs, and platform health.
  • Partner with Engineering, Product, and Support leaders to drive shared ownership of production services and embed reliability, observability, and operational excellence throughout the software development lifecycle.
  • Build and evolve observability capabilities using OpenTelemetry, Datadog, Coralogix, or similar tools, establishing enterprise standards for metrics, logs, traces, alerting, and Service Level Objectives (SLOs).
  • Oversee production triage, incident response, and escalation processes, ensuring timely service restoration, effective root cause analysis, and blameless post-incident reviews.
  • Champion a reliability-first engineering culture focused on automation, proactive risk reduction, operational readiness, shiftโ€‘left quality practices, and continuous improvement.
  • Collaborate with global engineering teams in a followโ€‘theโ€‘sun support model, ensuring seamless 24x7 coverage, effective operational handoffs, and consistent service ownership.
  • Own onโ€‘call programs, incident management practices, and operational health metrics, driving improvements in alert quality, operational efficiency, and toil reduction.
  • Manage team capacity, hiring, performance management, career development, budgeting, and workforce planning to ensure effective support of businessโ€‘critical services.
  • Provide regular reporting to engineering and executive leadership on reliability trends, incidents, risks, performance metrics, and strategic initiatives.
Required Qualifications
  • 5+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or related disciplines, including 2+ years in a technical leadership or people management role.
  • Proven experience leading teams responsible for production operations, reliability engineering, incident management, and operational excellence.
  • Experience designing and implementing SRE practices, reliability programs, or operational maturity initiatives within growing engineering organizations.
  • Experience operating large-scale, customerโ€‘facing SaaS platforms with high availability, performance, and scalability requirements.
  • Strong understanding of modern software engineering practices and partnering with development teams to build reliable, resilient systems.
  • Handsโ€‘on experience with observability platforms such as OpenTelemetry, Datadog, Coralogix, or similar technologies.
  • Strong knowledge of AWS and Kubernetes in production environments.
  • Deep understanding of monitoring, logging, distributed tracing, SLIs, SLOs, error budgets, and reliability engineering principles.
  • Demonstrated ability to lead crossโ€‘functional initiatives and influence stakeholders across Engineering, Product, and Support organizations.
  • Experience developing engineering roadmaps, defining team objectives, aligning reliability investments with business priorities, and driving continuous operational improvement through incident learning and postโ€‘incident reviews.
Preferred Qualifications
  • Experience leading distributed or globally dispersed engineering teams.
  • Experience with multiple cloud providers or cloudโ€‘agnostic platform architectures.
  • Familiarity with security, compliance, governance, and operational risk management frameworks.
  • Proficiency with modern Infrastructureโ€‘asโ€‘Code and technologies such as Terraform, Golang, Python, Playwright, and Performance Monitoring tools.
  • Experience with relational and distributed data technologies such as PostgreSQL, OpenSearch, Redis/ElastiCache, or Aurora.
  • Experience with messaging and streaming platforms such as Kafka, ActiveMQ, SNS/SQS, or similar eventโ€‘driven technologies.
  • Strong understanding of cost optimization, platform sustainability, and engineering efficiency metrics.

We create inspiring workplaces for some of the biggest and best companies in the world. And we do it within our own teams every day. Thatโ€™s one reason we made the Fortune 100 Best Companies to Work For list in 2021. Join us and watch people thrive at workโ€”including you. With seven global offices and employees working around the world, weโ€™re committed to creating an atmosphere where every person can share their talents and reach their potential.

#J-18808-Ljbffr