1

Reliability Engineer Jobs in Tennessee (NOW HIRING)

Site Reliability Engineer

Oak Ridge, TN ยท On-site

$54.50 - $72.50/hr

Senior Site Reliability Engineer, HPC Infrastructure and Platforms Overview: Seeking highly qualified individuals to play a key role in improving the security, performance, and reliability of the HPC ...

Operations Reliability Engineer

Nashville, TN ยท On-site

$99K - $124K/yr

Amazon is seeking an Operations Reliability Engineer to support the Amazon Logistics North America (AMZL NA) Engineering team, part of Amazon's Last Mile organization. The Last Mile team helps get ...

Site Reliability Engineer - Memphis

Southaven, MS ยท On-site

$53.50 - $71.25/hr

As a Site Reliability Engineer focused on campus reliability, you will design what the campus watches and trusts, technically command cross-discipline SEVs, and build the guardrails that make the ...

New

Systems Engineer - SRE Enablement

Memphis, TN ยท On-site

$55.50 - $73.75/hr

AutoZone's Site Reliability Engineering (SRE) team is seeking a Systems Engineer with a focus on SRE Enablement. This position is responsible for promoting reliability and operational excellence ...

Systems Engineer - SRE Enablement

Memphis, TN ยท On-site

$55.25 - $73.50/hr

AutoZone's Site Reliability Engineering (SRE) team is seeking a Systems Engineer with a focus on SRE Enablement. This position is responsible for promoting reliability and operational excellence ...

SRE Engagement Manager

Memphis, TN ยท On-site

$51 - $67.50/hr

Previous experience as an SRE, DevOps Engineer with a deep understanding of Production support, * on-call rotations, and incident management. Roles & Responsibilities * Drive the adoption of SRE ...

SRE Engagement Manager

Memphis, TN ยท On-site

$51 - $67.50/hr

Previous experience as an SRE, DevOps Engineer with a deep understanding of Production support, * on-call rotations, and incident management. Roles & Responsibilities * Drive the adoption of SRE ...

Service Reliability Engineer

Nashville, TN ยท On-site

$55 - $73.25/hr

As a Service Reliability Engineer, you won't just be supporting systems; you'll be ensuring the services that connect artists and fans around the globe are always on. Job Functions: Key ...

Site Reliability Engineer - Memphis

Memphis, TN ยท On-site

$55.25 - $73.50/hr

As a Site Reliability Engineer focused on campus reliability, you will design what the campus watches and trusts, technically command cross-discipline SEVs, and build the guardrails that make the ...

Site Reliability Engineer II

Nashville, TN

$55 - $73.25/hr

Site Reliability Engineer II The SRE II sits at the intersection of software engineering and platform operations. You will own the reliability, scalability, and operational hygiene of Kastle's core ...

Senior Site Reliability Engineer

Knoxville, TN ยท On-site

$50.75 - $67.50/hr

Senior Site Reliability Engineer Founded in 1999 in the beautiful Smoky Mountains of East Tennessee, Cadre5 provides innovative technical solutions to customers locally and nationally. Our Cadre5 Lab ...

Showing results 21-40

Reliability Engineer information

See Tennessee salary details

$55.4K

$107.1K

$128K

How much do reliability engineer jobs pay per year?

As of Sep 7, 2026, the average yearly pay for reliability engineer in Tennessee is $107,074.00, according to ZipRecruiter salary data. Most workers in this role earn between $93,000.00 and $117,100.00 per year, depending on experience, location, and employer.

What is a reliability engineer?

Reliability Engineers are professionals responsible for ensuring that systems, equipment, or processes function consistently and efficiently over time. They analyze data, identify potential points of failure, and develop maintenance strategies to improve system reliability and minimize downtime. Their work spans various industries, including manufacturing, energy, and technology, and often involves collaborating with design, operations, and maintenance teams. By implementing reliability-centered maintenance and predictive analysis, they help organizations save costs and increase safety.

What does a reliability engineer do?

As a reliability engineer, your duties are to test and evaluate the manufacturing of products and components and ensure that the procedures are efficient and do not lead to abnormally high maintenance or operational costs. Your other responsibilities are to find solutions to product reliability risks. You may manage risk in a supply chain, develop loss prevention strategies, and track the entire lifecycle of product development, from building prototypes to moving a product into full-scale production. You analyze information from department heads and recommend strategies to reduce risk and ensure that the product works reliably.

What are the key skills and qualifications needed to thrive as a reliability engineer, and why are they important?

To thrive as a Reliability Engineer, you need a solid background in engineering principles, failure analysis, and reliability modeling, typically with a degree in engineering or a related field. Familiarity with tools such as FMEA, Root Cause Analysis (RCA), reliability-centered maintenance (RCM) software, and certifications like Certified Reliability Engineer (CRE) are highly valued. Strong problem-solving abilities, attention to detail, and effective communication are crucial soft skills in this role. These skills ensure systems are dependable, downtime is minimized, and organizational performance and safety are optimized.

What are some typical challenges reliability engineers face when implementing preventive maintenance strategies?

Reliability Engineers often encounter challenges such as balancing preventive maintenance schedules with production demands, ensuring buy-in from operations teams, and accurately predicting equipment failures. They must analyze large sets of historical data to identify trends and root causes, which can be complex in facilities with diverse machinery. Collaboration with maintenance, operations, and engineering teams is essential to develop effective strategies that minimize downtime while optimizing resources.

What is the difference between Reliability Engineer vs Maintenance Engineer?

AspectReliability EngineerMaintenance Engineer
CredentialsTypically requires engineering degree, certifications in reliability or asset managementOften requires engineering or technical diploma, certifications in maintenance or equipment repair
Work EnvironmentFocuses on analysis, design, and improvement of systems for reliabilityHands-on maintenance, repair, and troubleshooting of equipment
Industry UsageCommon in manufacturing, energy, aerospace, and industrial sectorsPrevalent in manufacturing, facilities, and industrial plants

Reliability Engineers focus on designing and improving systems to prevent failures, using data analysis and modeling. Maintenance Engineers perform hands-on repairs and upkeep of equipment to ensure operational continuity. While both roles aim to optimize equipment performance, Reliability Engineers work proactively on system reliability, whereas Maintenance Engineers handle reactive and scheduled maintenance tasks.

Are reliability engineers in demand?

Reliability engineers are in high demand across industries such as manufacturing, energy, and aerospace due to their role in improving system performance and reducing downtime. Employers seek professionals with skills in data analysis, failure modes, and maintenance strategies, often requiring certifications like Certified Reliability Engineer (CRE). The job outlook is positive, with steady growth expected as companies prioritize operational efficiency and risk management.

How much do reliability engineers get paid?

Reliability engineers typically earn a median annual salary ranging from $70,000 to $110,000, depending on experience, location, and industry. Senior or specialized reliability engineers with certifications and advanced skills can earn higher salaries, often exceeding $120,000 annually.

What are the most commonly searched types of Reliability Engineer jobs in Tennessee?

The most popular types of Reliability Engineer jobs in Tennessee are:

What job categories do people searching Reliability Engineer jobs in Tennessee look for?

The top searched job categories for Reliability Engineer jobs in Tennessee are:

What cities in Tennessee are hiring for Reliability Engineer jobs?

Cities in Tennessee with the most Reliability Engineer job openings:

Infographic showing various Reliability Engineer job openings in Tennessee as of August 2026, with employment types broken down into 100% Full Time. Highlights an 80% In-person, and 20% Remote job distribution, with an average salary of $107,074 per year, or $51.5 per hour.

Site Reliability Engineer

ITR

Oak Ridge, TN โ€ข On-site

$54.50 - $72.50/hr

Full-time

Re-posted 17 days ago


Job description

Senior Site Reliability Engineer, HPC Infrastructure and Platforms
Overview:
Seeking highly qualified individuals to play a key role in improving the security, performance, and reliability of the HPC computing infrastructure which supports multiple highly ranked Top500 Supercomputers, including the world’s first exaflop system, Frontier.
The Team:
As a Senior Site Reliability Engineer, you will work within the HPC Infrastructure and Platforms group to support all activities of our supercomputer center. Our primary platform is the OLCF Slate Service, built on Kubernetes and Red Hat OpenShift, which provides a container orchestration service for running critical operation applications and user-managed persistent applications that run alongside our OLCF Supercomputer systems and other OLCF managed HPC clusters.
Major Duties/Responsibilities:
•    Lead ongoing improvements in reliability and scalability for our Kubernetes and Linux based applications and services.
•    Contribute as senior technical resource to define and implement best practices and standards for the center.
•    Provide primary operational support and engineering for production applications.
•    Define and implement define KPIs, processes and drive continuous improvement.
•    Influence the architecture and implementation of solutions.
•    Tune operating systems and applications to increase performance and reliability of services.
•    Mentor junior staff and enable them for success.
•    Diagnose system operational problems quickly and effectively.
•    Participate in on-call rotation providing 24-hour, 7-day support and off-hours maintenance windows.
•    Coordinate with vendors to resolve hardware and software problems.
•    Deliver client mission by aligning behaviors, priorities, and interactions with our core values of Impact, Integrity, Teamwork, Safety, and Service. Promote diversity, equity, inclusion, and accessibility by fostering a respectful workplace – in how we treat one another, work together, and measure success.
Basic Qualifications:
Bachelor’s Degree in computer science or closely related field and a minimum of 8 years of experience as an SRE/Systems Engineer.  An equivalent combination of education and experience may be considered.
Preferred Qualifications:
•    Excellent interpersonal/communication skills, and the ability to work as part of a team.
•    Strong working knowledge of Unix system fundamentals and common network protocols.
•    Experience managing Linux/UNIX operating systems in a heterogeneous environment.
•    Solid understanding of networked computing environment concepts.
•    Ability to develop and maintain programs and scripts that aid in the operation and automation using various shell (primarily bash) and high-level languages (Python or Go).
•    Ability to proactively identify performance issues, problems, and areas for improvement.
•    Ability to identify requirements and to define, plan, and implement requisite solutions.
•    Ability to plan, organize, prioritize tasks, and complete assigned projects with minimal supervision.
•    Experience with continuous integration and continuous deployment software methodologies and how they apply to SRE/systems engineering.
•    An understanding of code review and familiarity with tools like GitHub and GitLab
•    Experience using tools such as Nagios, Grafana and Prometheus to monitor systems, metrics, and create dashboards.
•    Experience designing and implement highly available systems/services utilizing virtual machines and Kubernetes resources.
•    Experience participating in an opensource community with patches accepted upstream.
•    Experience deploying and maintaining automated configuration management software such as Puppet or Ansible
•    Experience implementing systems-level security technologies like SELinux and following security best practices.
Special Requirement:
This position requires the ability to obtain and maintain a clearance from the Department of Energy. As such, this position is a Workplace Substance Abuse program (WSAP) testing designed position which requires passing a pre-placement drug test and participation in an ongoing random drug testing program in which employees are subject to being randomly selected for testing. The occupant of this position will also be subject to an ongoing requirement to report any drug-related arrest or conviction or receipt of a positive drug test result.