1

Senior Reliability Engineer Jobs in Tennessee (NOW HIRING)

Site Reliability Engineer

Oak Ridge, TN

$54.50 - $72.50/hr

Senior Site Reliability Engineer, HPC Infrastructure and Platforms Overview: Seeking highly qualified individuals to play a key role in improving the security, performance, and reliability of the HPC ...

Senior Site Reliability Engineer

Nashville, TN · On-site

$55 - $73.25/hr

We are looking for an experienced Senior Site Reliability Engineer to join our team and drive the reliability and functionality of our critical infrastructure and applications. In this role, you will ...

Principal Site Reliability Engineer

Nashville, TN · On-site

$55 - $73.25/hr

As a Principal Site Reliability Engineer (IC4), you will be responsible for designing, building ... Lead critical production incident response and act as a senior escalation point during major ...

Who We're Looking For We're hiring a Senior or Staff SRE to join our Platform team. This is a high-impact role on a small, capable team (1 Manager + 2 ICs) where you'll have significant ownership ...

Who We're Looking For We're hiring a Senior or Staff SRE to join our Platform team. This is a high-impact role on a small, capable team (1 Manager + 2 ICs) where you'll have significant ownership ...

Showing results 21-40

Senior Reliability Engineer information

See Tennessee salary details

$19

$58

$83

How much do senior reliability engineer jobs pay per hour?

As of Aug 16, 2026, the average hourly pay for senior reliability engineer in Tennessee is $58.46, according to ZipRecruiter salary data. Most workers in this role earn between $48.22 and $70.05 per hour, depending on experience, location, and employer.

How much do senior reliability engineers get paid?

Senior reliability engineers typically earn between $90,000 and $130,000 annually, depending on experience, industry, and location. They often have expertise in systems analysis, failure modes, and reliability tools like FMEA and RCM, which can influence compensation levels.

What are the key skills and qualifications needed to thrive as a senior reliability engineer?

To thrive as a Senior Reliability Engineer, you need expertise in reliability engineering principles, root cause analysis, and a relevant engineering degree such as mechanical, electrical, or industrial engineering. Familiarity with tools like FMEA, RCA software, CMMS, and certifications such as Certified Reliability Engineer (CRE) are often required. Strong analytical thinking, communication skills, and the ability to lead cross-functional teams set top performers apart. These skills are essential for minimizing downtime, improving system reliability, and ensuring safe, efficient operations.

What are some common challenges faced by senior reliability engineers, and how are they typically addressed within the team?

Senior Reliability Engineers often encounter challenges such as diagnosing complex system failures, balancing proactive maintenance with urgent reactive fixes, and ensuring consistent communication across multidisciplinary teams. These challenges are typically addressed through root cause analysis, prioritization frameworks, and fostering a culture of knowledge sharing. Regular collaboration with operations, maintenance, and engineering teams helps in developing effective solutions and continuous improvement strategies.

What does a senior reliability engineer do?

A Senior Reliability Engineer is responsible for ensuring that systems, products, or processes operate reliably and efficiently over time. They analyze failure data, design reliability tests, develop maintenance strategies, and work with cross-functional teams to improve system performance and reduce downtime. Their expertise helps organizations minimize risk, optimize lifecycle costs, and maintain high standards of quality and safety. Senior Reliability Engineers often mentor junior team members and play a key role in developing reliability standards and best practices.

What is the difference between Senior Reliability Engineer vs Reliability Engineer?

AspectSenior Reliability EngineerReliability Engineer
CredentialsTypically requires 5+ years experience, certifications like CRE or Six SigmaEntry to mid-level, often with 2-4 years experience, similar certifications
Work EnvironmentDesigns and oversees reliability programs, leads projectsPerforms analysis, supports reliability improvements
Industry UsageUsed across manufacturing, energy, aerospaceCommon in same industries, often as a stepping stone to senior roles

The main difference between a Senior Reliability Engineer and a Reliability Engineer lies in experience, leadership responsibilities, and scope of work. Senior Reliability Engineers typically lead projects and develop strategies, while Reliability Engineers focus on analysis and supporting reliability initiatives. Both roles are vital in ensuring equipment and system dependability across industries.

What are the most commonly searched types of Reliability Engineer jobs in Tennessee?

The most popular types of Reliability Engineer jobs in Tennessee are:

What are popular job titles related to Senior Reliability Engineer jobs in Tennessee?

For Senior Reliability Engineer jobs in Tennessee, the most frequently searched job titles are:

What job categories do people searching Senior Reliability Engineer jobs in Tennessee look for?

The top searched job categories for Senior Reliability Engineer jobs in Tennessee are:

What cities in Tennessee are hiring for Senior Reliability Engineer jobs?

Cities in Tennessee with the most Senior Reliability Engineer job openings:

Infographic showing various Senior Reliability Engineer job openings in Tennessee as of August 2026, with employment types broken down into 96% Full Time, and 4% Contract. Highlights an 82% In-person, 14% Hybrid, and 4% Remote job distribution, with an average salary of $121,603 per year, or $58.5 per hour.

Site Reliability Engineer

ITR

Oak Ridge, TN

$54.50 - $72.50/hr

Full-time

Posted 25 days ago


Job description

Senior Site Reliability Engineer, HPC Infrastructure and Platforms
Overview:
Seeking highly qualified individuals to play a key role in improving the security, performance, and reliability of the HPC computing infrastructure which supports multiple highly ranked Top500 Supercomputers, including the world’s first exaflop system, Frontier.
The Team:
As a Senior Site Reliability Engineer, you will work within the HPC Infrastructure and Platforms group to support all activities of our supercomputer center. Our primary platform is the OLCF Slate Service, built on Kubernetes and Red Hat OpenShift, which provides a container orchestration service for running critical operation applications and user-managed persistent applications that run alongside our OLCF Supercomputer systems and other OLCF managed HPC clusters.
Major Duties/Responsibilities:
• Lead ongoing improvements in reliability and scalability for our Kubernetes and Linux based applications and services.
• Contribute as senior technical resource to define and implement best practices and standards for the center.
• Provide primary operational support and engineering for production applications.
• Define and implement define KPIs, processes and drive continuous improvement.
• Influence the architecture and implementation of solutions.
• Tune operating systems and applications to increase performance and reliability of services.
• Mentor junior staff and enable them for success.
• Diagnose system operational problems quickly and effectively.
• Participate in on-call rotation providing 24-hour, 7-day support and off-hours maintenance windows.
• Coordinate with vendors to resolve hardware and software problems.
• Deliver client mission by aligning behaviors, priorities, and interactions with our core values of Impact, Integrity, Teamwork, Safety, and Service. Promote diversity, equity, inclusion, and accessibility by fostering a respectful workplace – in how we treat one another, work together, and measure success.
Basic Qualifications:
Bachelor’s Degree in computer science or closely related field and a minimum of 8 years of experience as an SRE/Systems Engineer. An equivalent combination of education and experience may be considered.
Preferred Qualifications:
• Excellent interpersonal/communication skills, and the ability to work as part of a team.
• Strong working knowledge of Unix system fundamentals and common network protocols.
• Experience managing Linux/UNIX operating systems in a heterogeneous environment.
• Solid understanding of networked computing environment concepts.
• Ability to develop and maintain programs and scripts that aid in the operation and automation using various shell (primarily bash) and high-level languages (Python or Go).
• Ability to proactively identify performance issues, problems, and areas for improvement.
• Ability to identify requirements and to define, plan, and implement requisite solutions.
• Ability to plan, organize, prioritize tasks, and complete assigned projects with minimal supervision.
• Experience with continuous integration and continuous deployment software methodologies and how they apply to SRE/systems engineering.
• An understanding of code review and familiarity with tools like GitHub and GitLab
• Experience using tools such as Nagios, Grafana and Prometheus to monitor systems, metrics, and create dashboards.
• Experience designing and implement highly available systems/services utilizing virtual machines and Kubernetes resources.
• Experience participating in an opensource community with patches accepted upstream.
• Experience deploying and maintaining automated configuration management software such as Puppet or Ansible
• Experience implementing systems-level security technologies like SELinux and following security best practices.
Special Requirement:
This position requires the ability to obtain and maintain a clearance from the Department of Energy. As such, this position is a Workplace Substance Abuse program (WSAP) testing designed position which requires passing a pre-placement drug test and participation in an ongoing random drug testing program in which employees are subject to being randomly selected for testing. The occupant of this position will also be subject to an ongoing requirement to report any drug-related arrest or conviction or receipt of a positive drug test result.