1

Site Reliability Manager Jobs in Tennessee (NOW HIRING)

Site Reliability Engineer

Oak Ridge, TN ยท On-site

$54.50 - $72.50/hr

Experience managing Linux/UNIX operating systems in a heterogeneous environment. * Solid ... to SRE/systems engineering. * An understanding of code review and familiarity with tools like ...

Systems Engineer - SRE Enablement

Memphis, TN ยท On-site

$55.50 - $73.75/hr

Participate in the incident management process, including post-mortem analysis, to continuously strengthen systemic reliability. * Run SRE training programs and reliability workshops for engineering ...

Systems Engineer - SRE Enablement

Memphis, TN ยท On-site

$55.25 - $73.50/hr

Participate in the incident management process, including post-mortem analysis, to continuously strengthen systemic reliability. * Run SRE training programs and reliability workshops for engineering ...

Executes the site's asset reliability strategy and program to minimize downtime and leverage plant ... Manages contractor relationships and contractor work assignments in a way that increases value ...

Senior Site Reliability Engineer

Knoxville, TN ยท On-site

$50.75 - $67.50/hr

Experience managing Linux/UNIX operating systems in a heterogeneous environment. * Solid ... to SRE/systems engineering. * An understanding of code review and familiarity with tools like ...

Site Reliability Engineer II

Nashville, TN

$55 - $73.25/hr

Kastle Systems is the leader in managed security, with a track record of introducing innovative ... Site Reliability Engineer II The SRE II sits at the intersection of software engineering and ...

Senior Site Reliability Engineer

Nashville, TN ยท On-site

$55 - $73.25/hr

Experience with infrastructure-as-code or configuration-management tools. * Production support, incident response, or SRE operational experience. * Knowledge of monitoring, alerting, centralized ...

Senior Site Reliability Engineer

Nashville, TN ยท On-site

$55 - $73.25/hr

Three or more years of experience in site reliability engineering, systems administration ... Experience with ticketing, incident-management, or change-management systems. * Strong ...

Site Reliability Engineer 2

Nashville, TN ยท On-site

$55 - $73.25/hr

Experience with ticketing, incident-management, or change-management systems. * Strong ... Experience with incident response, root-cause analysis, and SRE operational practices. * Knowledge ...

Site Reliability Engineer 2

Nashville, TN ยท On-site

$55 - $73.25/hr

Experience with ticketing, incident-management, or change-management systems. * Strong ... Experience with incident response, root-cause analysis, and SRE operational practices. * Knowledge ...

Staff Site Reliability Engineer

Brentwood, TN ยท Hybrid

$180K - $210K/yr

You've managed ElasticSearch at scale and know the difference between logs that help and logs that just cost money. You think in terms of infrastructure as code and believe Terraform plans should be ...

Lead Principal Site Reliability Engineer

Nashville, TN ยท On-site

$55 - $73.25/hr

Define and drive the site reliability engineering strategy for large-scale, distributed, and ... Incident Management and Problem Resolution * Provide technical leadership during complex, high ...

Principal Site Reliability Engineer

Nashville, TN ยท On-site

$55 - $73.25/hr

Extensive experience in site reliability engineering, systems engineering, infrastructure ... Experience managing complex or high-risk production changes. Troubleshooting and Operational ...

Principal Site Reliability Engineer

Nashville, TN ยท On-site

$55 - $73.25/hr

The Principal Site Reliability Engineer will work across Windows, Linux, Oracle Cloud ... Experience managing complex or high-risk production changes. Troubleshooting and Operational ...

Principal Site Reliability Engineer

Nashville, TN ยท On-site

$55 - $73.25/hr

Extensive experience in site reliability engineering, systems engineering, infrastructure ... Experience managing complex or high-risk production changes. Troubleshooting and Operational ...

next page

Showing results 1-20

Site Reliability Manager information

See Tennessee salary details

$56.3K

$106.6K

$152.9K

How much do site reliability manager jobs pay per year?

As of Aug 24, 2026, the average yearly pay for site reliability manager in Tennessee is $106,634.00, according to ZipRecruiter salary data. Most workers in this role earn between $85,800.00 and $127,100.00 per year, depending on experience, location, and employer.

What is the difference between Site Reliability Manager vs DevOps Engineer?

AspectSite Reliability ManagerDevOps Engineer
CredentialsTypically requires a Bachelor's in Computer Science, certifications like SRE or Cloud certificationsOften holds a Bachelor's in Computer Science or related field, with certifications in cloud platforms or automation tools
Work EnvironmentLeads teams managing large-scale systems, focusing on reliability and uptimeWorks on automation, CI/CD pipelines, and infrastructure deployment
Industry UsageCommon in tech, cloud services, and large enterprisesWidely used in startups, tech companies, and organizations adopting DevOps practices

The main difference is that Site Reliability Managers focus on ensuring system reliability and managing SRE teams, while DevOps Engineers concentrate on automation, deployment, and continuous integration. Both roles require technical expertise but serve different strategic objectives within IT operations.

What is the role of a site reliability manager?

A site reliability manager (SRM) is responsible for ensuring the reliability, availability, and performance of a company's IT systems and services. They often oversee incident response, implement automation tools, and collaborate with development teams to improve system stability and scalability, typically requiring knowledge of monitoring tools and scripting skills.

What cities in Tennessee are hiring for Site Reliability Manager jobs?

Cities in Tennessee with the most Site Reliability Manager job openings:

Infographic showing various Site Reliability Manager job openings in Tennessee as of August 2026, with employment types broken down into 90% Full Time, 9% Part Time, and 1% Contract. Highlights an 81% Physical, 2% Hybrid, and 17% Remote job distribution, with an average salary of $106,634 per year, or $51.3 per hour.

Site Reliability Engineer

ITR

Oak Ridge, TN โ€ข On-site

$54.50 - $72.50/hr

Full-time

Re-posted 3 days ago


Job description

Senior Site Reliability Engineer, HPC Infrastructure and Platforms
Overview:
Seeking highly qualified individuals to play a key role in improving the security, performance, and reliability of the HPC computing infrastructure which supports multiple highly ranked Top500 Supercomputers, including the world’s first exaflop system, Frontier.
The Team:
As a Senior Site Reliability Engineer, you will work within the HPC Infrastructure and Platforms group to support all activities of our supercomputer center. Our primary platform is the OLCF Slate Service, built on Kubernetes and Red Hat OpenShift, which provides a container orchestration service for running critical operation applications and user-managed persistent applications that run alongside our OLCF Supercomputer systems and other OLCF managed HPC clusters.
Major Duties/Responsibilities:
• Lead ongoing improvements in reliability and scalability for our Kubernetes and Linux based applications and services.
• Contribute as senior technical resource to define and implement best practices and standards for the center.
• Provide primary operational support and engineering for production applications.
• Define and implement define KPIs, processes and drive continuous improvement.
• Influence the architecture and implementation of solutions.
• Tune operating systems and applications to increase performance and reliability of services.
• Mentor junior staff and enable them for success.
• Diagnose system operational problems quickly and effectively.
• Participate in on-call rotation providing 24-hour, 7-day support and off-hours maintenance windows.
• Coordinate with vendors to resolve hardware and software problems.
• Deliver client mission by aligning behaviors, priorities, and interactions with our core values of Impact, Integrity, Teamwork, Safety, and Service. Promote diversity, equity, inclusion, and accessibility by fostering a respectful workplace – in how we treat one another, work together, and measure success.
Basic Qualifications:
Bachelor’s Degree in computer science or closely related field and a minimum of 8 years of experience as an SRE/Systems Engineer. An equivalent combination of education and experience may be considered.
Preferred Qualifications:
• Excellent interpersonal/communication skills, and the ability to work as part of a team.
• Strong working knowledge of Unix system fundamentals and common network protocols.
• Experience managing Linux/UNIX operating systems in a heterogeneous environment.
• Solid understanding of networked computing environment concepts.
• Ability to develop and maintain programs and scripts that aid in the operation and automation using various shell (primarily bash) and high-level languages (Python or Go).
• Ability to proactively identify performance issues, problems, and areas for improvement.
• Ability to identify requirements and to define, plan, and implement requisite solutions.
• Ability to plan, organize, prioritize tasks, and complete assigned projects with minimal supervision.
• Experience with continuous integration and continuous deployment software methodologies and how they apply to SRE/systems engineering.
• An understanding of code review and familiarity with tools like GitHub and GitLab
• Experience using tools such as Nagios, Grafana and Prometheus to monitor systems, metrics, and create dashboards.
• Experience designing and implement highly available systems/services utilizing virtual machines and Kubernetes resources.
• Experience participating in an opensource community with patches accepted upstream.
• Experience deploying and maintaining automated configuration management software such as Puppet or Ansible
• Experience implementing systems-level security technologies like SELinux and following security best practices.
Special Requirement:
This position requires the ability to obtain and maintain a clearance from the Department of Energy. As such, this position is a Workplace Substance Abuse program (WSAP) testing designed position which requires passing a pre-placement drug test and participation in an ongoing random drug testing program in which employees are subject to being randomly selected for testing. The occupant of this position will also be subject to an ongoing requirement to report any drug-related arrest or conviction or receipt of a positive drug test result.