1

Reliability Jobs in Tennessee (NOW HIRING)

Principal Site Reliability Engineer

Nashville, TN · On-site

$55 - $73.25/hr

Designs and architects infrastructure and service to ensure reliability and functionality. Forecasts demands and responds to capacity needs. Collaborates with software development teams to develop ...

Principal Site Reliability Engineer

Nashville, TN · On-site

$55 - $73.25/hr

Designs and architects infrastructure and service to ensure reliability and functionality. Forecasts demands and responds to capacity needs. Collaborates with software development teams to develop ...

Senior Site Reliability Engineer

Nashville, TN · On-site

$55 - $73.25/hr

We are looking for an experienced Senior Site Reliability Engineer to join our team and drive the reliability and functionality of our critical infrastructure and applications. In this role, you will ...

Lead Principal Site Reliability Engineer

Nashville, TN · On-site

$55 - $73.25/hr

Establish reliability standards, engineering practices, and operational readiness requirements across multiple teams. * Serve as a senior technical authority for system reliability, scalability ...

Site Reliability Engineer - Memphis

Memphis, TN · On-site

$55.25 - $73.50/hr

As a Site Reliability Engineer focused on campus reliability, you will design what the campus watches and trusts, technically command cross-discipline SEVs, and build the guardrails that make the ...

Principal Site Reliability Engineer

Nashville, TN · On-site

$55 - $73.25/hr

The Principal Site Reliability Engineer will work across Windows, Linux, Oracle Cloud Infrastructure, hybrid cloud, and legacy environments while partnering with engineering, operations ...

Site Reliability Engineer II

Nashville, TN · On-site

$55 - $73.25/hr

Site Reliability Engineer II The SRE II sits at the intersection of software engineering and platform operations. You will own the reliability, scalability, and operational hygiene of Kastle's core ...

Systems Engineer - SRE Enablement

Memphis, TN · On-site

$55.50 - $73.75/hr

AutoZone's Site Reliability Engineering (SRE) team is seeking a Systems Engineer with a focus on SRE Enablement. This position is responsible for promoting reliability and operational excellence ...

$60.25 - $80.25/hr

The Site Reliability Center (SRC) is focused on establishing a culture of operational excellence by ensuring infrastructure, platforms, and applications adhere to SRC onboarding standards that ...

New

Systems Engineer - SRE Enablement

Memphis, TN · On-site

$55.25 - $73.50/hr

AutoZone's Site Reliability Engineering (SRE) team is seeking a Systems Engineer with a focus on SRE Enablement. This position is responsible for promoting reliability and operational excellence ...

Site Reliability Engineer 2

Nashville, TN · On-site

$55 - $73.25/hr

Contributes to the design and architecture of infrastructure and service to ensure reliability and functionality. Responds to infrastructure demands and capacity needs. Collaborates with software ...

Senior Site Reliability Engineer

Nashville, TN · On-site

$55 - $73.25/hr

Performs data collection to maintain and optimize operations and reliability. Leverages knowledge to perform incident response and/or maintenance tasks. Provides health and performance reporting.

Site Reliability Engineer 2

Nashville, TN · On-site

$55 - $73.25/hr

Contributes to the design and architecture of infrastructure and service to ensure reliability and functionality. Responds to infrastructure demands and capacity needs. Collaborates with software ...

Senior Site Reliability Engineer

Knoxville, TN · On-site

$50.75 - $67.50/hr

Senior Site Reliability Engineer Founded in 1999 in the beautiful Smoky Mountains of East Tennessee, Cadre5 provides innovative technical solutions to customers locally and nationally. Our Cadre5 Lab ...

Senior Site Reliability Engineer

Nashville, TN · On-site

$55 - $73.25/hr

Performs data collection to maintain and optimize operations and reliability. Leverages knowledge to perform incident response and/or maintenance tasks. Provides health and performance reporting.

Showing results 41-60

Reliability information

See Tennessee salary details

$56.3K

$106.6K

$152.9K

How much do reliability jobs pay per year?

As of Sep 10, 2026, the average yearly pay for reliability in Tennessee is $106,634.00, according to ZipRecruiter salary data. Most workers in this role earn between $85,800.00 and $127,100.00 per year, depending on experience, location, and employer.

What does a reliability engineer do?

A Reliability Engineer focuses on ensuring that systems, equipment, and processes operate dependably and efficiently over time. Their main responsibilities include analyzing failure data, identifying potential risks, and developing strategies to minimize downtime and maintenance costs. They work closely with design, maintenance, and production teams to improve product quality, safety, and reliability. By implementing reliability testing and preventive maintenance plans, they help organizations achieve higher performance and customer satisfaction.

What are the key skills and qualifications needed to thrive as a reliability engineer?

To thrive as a Reliability Engineer, you need a solid background in engineering principles, data analysis, and failure mode analysis, typically supported by a degree in engineering or a related field. Familiarity with tools like FMEA, Root Cause Analysis (RCA), reliability modeling software, and relevant certifications such as Certified Reliability Engineer (CRE) are commonly required. Strong problem-solving skills, attention to detail, and effective communication help Reliability Engineers collaborate and proactively address system weaknesses. These skills are essential for minimizing downtime, optimizing maintenance strategies, and ensuring long-term operational efficiency.

What are some common challenges faced by professionals in reliability engineering roles, and how can they be addressed?

Reliability Engineers often face the challenge of balancing immediate production demands with long-term equipment health and system improvements. They must analyze complex failure data, collaborate with cross-functional teams, and drive changes that may require operational adjustments. To address these challenges, it's important to develop strong communication skills, stay current with reliability methodologies (like FMEA and RCA), and foster a culture of proactive maintenance. Building relationships with operations and maintenance teams also helps ensure that reliability recommendations are adopted and sustained.

What is the difference between Reliability vs Maintenance Technician?

AspectReliabilityMaintenance Technician
Required credentialsCertifications in reliability engineering, asset management, or predictive maintenanceTechnical diploma or certification in maintenance, HVAC, or mechanical systems
Work environmentFocus on analyzing data, improving systems, and preventing failures in industrial or manufacturing settingsHands-on repair, troubleshooting, and maintenance of equipment and machinery
Employer and industry usageUsed in manufacturing, energy, and industrial sectors to optimize equipment performanceCommonly employed in factories, plants, and facilities for equipment upkeep

Reliability professionals focus on analyzing data and implementing strategies to prevent equipment failures, while maintenance technicians perform hands-on repairs and routine maintenance. Both roles are essential for operational efficiency but differ in their focus and approach.

How to become a reliability specialist?

To become a reliability specialist, typically a bachelor's degree in engineering, manufacturing, or a related field is required. Gaining experience in maintenance, quality assurance, or engineering, along with knowledge of reliability tools like FMEA or root cause analysis, is important. Certifications such as Certified Reliability Engineer (CRE) can enhance job prospects.

What are some reliable jobs?

Reliable jobs are typically those with consistent demand and stable employment, such as roles in healthcare, education, government, and skilled trades. These positions often offer steady hours, benefits, and opportunities for advancement, making them suitable for long-term career stability.

What are the most commonly searched types of Reliability jobs in Tennessee?

The most popular types of Reliability jobs in Tennessee are:

What are popular job titles related to Reliability jobs in Tennessee?

For Reliability jobs in Tennessee, the most frequently searched job titles are:

What job categories do people searching Reliability jobs in Tennessee look for?

The top searched job categories for Reliability jobs in Tennessee are:

What cities in Tennessee are hiring for Reliability jobs?

Cities in Tennessee with the most Reliability job openings:

Infographic showing various Reliability job openings in Tennessee as of August 2026, with employment types broken down into 1% As Needed, 82% Full Time, 13% Part Time, 3% Contract, and 1% Nights. Highlights an 90% Physical, 4% Hybrid, and 6% Remote job distribution, with an average salary of $106,634 per year, or $51.3 per hour.

Principal Site Reliability Engineer

Nashville, TN • On-site

Oracle
IT Services • 10K+ employees

$55 - $73.25/hr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Re-posted 15 days ago


Oracle rating

8.7

Company rating: 8.7 out of 10

Based on 152 frontline employees who took The Breakroom Quiz


Job description


Designs and architects infrastructure and service to ensure reliability and functionality. Forecasts demands and responds to capacity needs. Collaborates with software development teams to develop reliable and scalable infrastructures. Exercises judgment when performing data collection to maintain and optimize operations and reliability. Leverages advanced knowledge to perform incident response and/or maintenance tasks. Provides comprehensive health and performance reporting. Identifies and recommends opportunities for automation. Communicates comprehensive information about services and proactively anticipates and articulates the potential impact of changes. Provides comprehensive support for technology and documents incidents. Conducts advanced experiments with new tools and develops and maintains advanced knowledge of site reliability trends.
Responsibilities
Key Responsibilities
  • Design and architect reliable, secure, scalable, and maintainable infrastructure and services. Take proactive steps to ensure solutions meet defined reliability and functionality requirements.
  • Establish technical direction, engineering standards, and operational best practices across complex infrastructure and application environments.
  • Identify system dependencies, operational risks, capacity constraints, performance issues, and potential failure points before they affect service.
  • Translate business, client, security, and application requirements into practical infrastructure and reliability solutions.
  • Lead the installation, configuration, deployment, and validation of applications across Windows Server and Linux environments.
  • Oversee structured builds and deployments using runbooks, scripts, readiness assessments, change controls, and post-deployment validation.
  • Troubleshoot complex operating system, service, application, installation, patching, permissions, certificate, and connectivity issues.
  • Define and improve monitoring, alerting, logging, observability, capacity planning, and service-health practices.
  • Develop and promote automation that reduces manual effort, improves consistency, and lowers operational risk.
  • Lead operating system, middleware, and application patching initiatives, including change planning, rollback preparation, execution, and validation.
  • Direct major incident response, root cause analysis, corrective-action planning, and prevention of recurring failures.
  • Partner with cybersecurity teams on vulnerability remediation, system hardening, STIG compliance, and other security-driven changes.
  • Evaluate emerging technologies and recommend solutions that improve reliability, resilience, security, and operational efficiency.
  • Create and maintain technical standards, architecture documentation, runbooks, deployment procedures, and troubleshooting guides.
  • Provide technical leadership, mentorship, and design guidance to engineers across multiple teams.
  • Communicate technical risks, dependencies, decisions, and recommendations clearly to leadership and stakeholders.

Core Skills and Qualifications
Technical Leadership and Architecture
  • Extensive experience in site reliability engineering, systems engineering, infrastructure architecture, production operations, or application hosting.
  • Demonstrated ability to design and support highly available, resilient, and secure enterprise systems.
  • Experience leading complex technical initiatives across engineering, operations, security, networking, and application teams.
  • Ability to make sound architectural decisions, evaluate tradeoffs, and communicate recommendations to technical and nontechnical stakeholders.
  • Experience defining engineering standards, operational controls, and reliability practices.

Windows and Linux System Administration
  • Advanced, hands-on experience administering Windows Server and/or Linux systems.
  • Ability to access deployed hosts and perform post-deployment configuration, troubleshooting, and validation.
  • Experience installing, configuring, and validating applications in Windows Server and Linux environments.
  • Ability to resolve operating-system-level, service-level, and application-level issues.
  • Strong knowledge of system services, permissions, configuration files, logs, processes, and resource utilization.

Manual Build and Deployment Experience
  • Experience leading structured build and deployment activities using runbooks, deployment guides, scripts, and technical procedures.
  • Ability to execute and troubleshoot scripts, validate outputs, and resolve build or configuration issues.
  • Experience with build handoffs, environment-readiness assessments, deployment validation, and post-build verification.
  • Ability to identify process gaps, document exceptions, and improve deployment procedures.
  • Experience managing complex or high-risk production changes.

Troubleshooting and Operational Support
  • Ability to investigate complex service failures, installation errors, patching failures, application startup problems, permissions issues, and connectivity incidents.
  • Experience reviewing logs, event viewers, service status, configuration files, ports, certificates, and access controls.
  • Strong analytical and problem-solving skills, with the ability to isolate root causes and implement sustainable solutions.
  • Experience leading major incident response and coordinating technical teams during business-critical outages.
  • Ability to document symptoms, findings, impact, corrective actions, and recommended next steps clearly.
  • Extensive experience supporting production or other mission-critical environments.

Scripting and Automation
Advanced hands-on experience with one or more of the following:
  • PowerShell
  • Bash
  • Python
  • Ansible
  • Chef

Candidates should be able to create, modify, validate, and troubleshoot scripts and automation workflows. Experience identifying automation opportunities and establishing safe, repeatable operational processes is essential.
Patching and Software Maintenance
  • Experience planning and executing operating system, middleware, and application patching.
  • Ability to troubleshoot patch failures, compatibility issues, and post-patch application problems.
  • Strong understanding of maintenance windows, change control, rollback planning, risk assessment, and post-change validation.
  • Experience coordinating patching and remediation activities across application, infrastructure, cybersecurity, and client teams.

Cloud and Oracle Cloud Infrastructure
  • Strong understanding of cloud-hosted and hybrid infrastructure.
  • Experience with Oracle Cloud Infrastructure or another major cloud platform.
  • Knowledge of cloud compute, storage, networking, identity, access management, load balancing, and environment provisioning.
  • Experience designing or supporting reliable, secure, and scalable cloud environments.
  • Familiarity with infrastructure-as-code and configuration-management practices.

Network and Connectivity Troubleshooting
  • Working knowledge of DNS, firewalls, routing, load balancers, ports, certificates, and network communication between systems.
  • Ability to identify and isolate host, application, certificate, firewall, DNS, and routing-related issues.
  • Familiarity with standard connectivity and network diagnostic tools.

Cybersecurity and Compliance
  • Experience with vulnerability remediation, system hardening, secure configuration, and compliance-driven infrastructure changes.
  • Familiarity with Security Technical Implementation Guides and federal cybersecurity requirements.
  • Ability to implement security remediation without disrupting application functionality or service availability.
  • Experience supporting federal, government-hosted, healthcare, or other regulated environments is highly valued.

Documentation and Communication
  • Ability to create and maintain architecture documentation, operational standards, technical procedures, runbooks, and change records.
  • Strong attention to detail when documenting completed work, risks, exceptions, decisions, and validation results.
  • Experience working in ticketing, incident-management, problem-management, and change-management systems.
  • Excellent written and verbal communication skills.
  • Ability to present complex technical issues, risks, and recommendations to engineers, clients, and senior leadership.
  • Demonstrated ability to mentor engineers and influence technical direction across teams.

Preferred Qualifications
  • Experience supporting Oracle Cloud Infrastructure environments.
  • Experience supporting Oracle Health, Millennium, or Cerner applications and related infrastructure.
  • Knowledge of federal cybersecurity workflows, STIG implementation, and compliance requirements.
  • Experience supporting federal clients or government-hosted environments.
  • Experience with Citrix technologies.
  • Experience supporting legacy infrastructure and business-critical legacy applications.
  • Advanced experience with infrastructure-as-code or configuration-management tools.
  • Experience with production support, incident command, problem management, and SRE operational practices.
  • Knowledge of service-level indicators, service-level objectives, error budgets, observability, and capacity planning.
  • Experience designing high-availability, disaster-recovery, backup, and service-continuity solutions.

Qualifications
Disclaimer:
Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.
Range and benefit information provided in this posting are specific to the stated locations only
US: Hiring Range in USD from: $84,900 to $209,500 per annum. May be eligible for bonus and equity.
Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.
Oracle US offers a comprehensive benefits package which includes the following:
1. Medical, dental, and vision insurance, including expert medical opinion
2. Short term disability and long term disability
3. Life insurance and AD&D
4. Supplemental life insurance (Employee/Spouse/Child)
5. Health care and dependent care Flexible Spending Accounts
6. Pre-tax commuter and parking benefits
7. 401(k) Savings and Investment Plan with company match
8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
9. 11 paid holidays
10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
11. Paid parental leave
12. Adoption assistance
13. Employee Stock Purchase Plan
14. Financial planning and group legal
15. Voluntary benefits including auto, homeowner and pet insurance
The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
Career Level - IC4
About Us
Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.
True innovation starts when everyone is empowered to contribute. That's why we're committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.
We're committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation-request_mb@oracle.com or by calling 1-888-404-2494 in the United States.
Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans' status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

What Oracle employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Oracle logo

About Oracle

Sourced by ZipRecruiter

An Oracle career can span industries, roles, Countries and cultures, giving you the opportunity to flourish in new roles and innovate, while blending work life in. Oracle has thrived through 40+ years of change by innovating and operating with integrity while delivering for the top companies in almost every industry. In order to nurture the talent that makes this happen, we are committed to an inclusive culture that celebrates and values diverse insights and perspectives, a workforce that inspires thought leadership and innovation. Oracle offers a highly competitive suite of Employee Benefits designed on the principles of parity, consistency, and affordability. The overall package includes certain core elements such as Medical, Life Insurance, access to Retirement Planning, and much more. We also encourage our employees to engage in the culture of giving back to the communities where we live and do business. At Oracle, we believe that innovation starts with diversity and inclusion and to create the future we need talent from various backgrounds, perspectives, and abilities. We ensure that individuals with disabilities are provided reasonable accommodation to successfully participate in the job application, interview process, and in potential roles. to perform crucial job functions. That's why we're committed to creating a workforce where all individuals can do their best work. It's when everyone's voice is heard and valued that we're inspired to go beyond what's been done before.

Industry

It services

Company size

10,000+ Employees

Headquarters location

Redwood City, CA, US

Year founded

1977

Social media