1

Observability Site Reliability Engineer Jobs in Tennessee

Systems Engineer - SRE Enablement

Memphis, TN · On-site

$55.50 - $73.75/hr

Build, maintain, and standardize shared observability platforms, specifically leveraging Dynatrace ... Run SRE training programs and reliability workshops for engineering teams. * Coach and mentor teams ...

Systems Engineer - SRE Enablement

Memphis, TN

$55.25 - $73.50/hr

Hands-on experience building, administering, and optimizing observability and APM pipelines, with a ... Run SRE training programs and reliability workshops for engineering teams. * Coach and mentor teams ...

Site Reliability Engineer II

Nashville, TN

$55 - $73.25/hr

Site Reliability Engineer II The SRE II sits at the intersection of software engineering and ... Reliability & Observability * Define, instrument, and enforce SLIs and SLOs in partnership with ...

Site Reliability Engineer

Oak Ridge, TN · On-site

$54.50 - $72.50/hr

Senior Site Reliability Engineer, HPC Infrastructure and Platforms Overview: Seeking highly qualified individuals to play a key role in improving the security, performance, and reliability of the HPC ...

Senior Site Reliability Engineer

Knoxville, TN

$50.75 - $67.50/hr

Senior Site Reliability Engineer Founded in 1999 in the beautiful Smoky Mountains of East Tennessee, Cadre5 provides innovative technical solutions to customers locally and nationally. Our Cadre5 Lab ...

Service Reliability Engineer

Nashville, TN

$55 - $73.25/hr

Develop and maintain robust monitoring, alerting, and observability systems (e.g., using AWS ... Partner with engineering and IT stakeholders to embed SRE best practices (SLOs, error budgets) into ...

Set standards for code quality, testing, observability, and operational readiness. Engineering ... Experience as a Site Reliability Engineer * Open-source contributions * Code generation frameworks

The Software Reliability Engineer (SRE) will play a critical role in ensuring that our Warehouse ... Improve observability by suggesting/implement better logging practices and metric coverage. 4. ...

Senior Platform Engineer

Chattanooga, TN · On-site

$95K - $130K/yr

... observability platforms. • Champion Site Reliability Engineering (SRE) practices including incident response, root cause analysis, runbook development, service reliability, and operational ...

next page

Showing results 1-20

Observability Site Reliability Engineer information

What engineer makes $500,000 a year?

A senior or principal Site Reliability Engineer (SRE) or Observability Engineer with extensive experience, specialized skills, and working at large tech companies can earn $500,000 or more annually. Compensation often includes base salary, bonuses, and stock options, especially in high-demand markets and organizations with complex infrastructure.

Is AI replacing SRE?

AI is augmenting the work of Site Reliability Engineers (SREs) by automating tasks such as monitoring, incident detection, and response. However, SREs are still essential for designing systems, managing complex issues, and making strategic decisions that require human judgment. AI tools are considered complementary rather than replacements for SREs' expertise and problem-solving skills.

What engineers make $200,000 a year?

Senior Site Reliability Engineers and Observability Engineers with extensive experience, advanced skills in cloud platforms, automation, and monitoring tools can earn $200,000 or more annually. High compensation often correlates with working at large tech companies, possessing specialized certifications, and managing complex, scalable systems.

What is the difference between Observability Site Reliability Engineer vs Monitoring Engineer?

AspectObservability Site Reliability EngineerMonitoring Engineer
FocusEnsuring system reliability through observability, automation, and incident responseImplementing and managing monitoring tools and dashboards
SkillsCloud platforms, scripting, incident management, observability toolsMonitoring tools, alerting systems, data analysis
Work EnvironmentDevOps teams, cloud infrastructure, large-scale systemsOperations teams, infrastructure monitoring

While both roles involve system health, the Observability Site Reliability Engineer focuses on comprehensive system reliability using observability practices, whereas Monitoring Engineers primarily manage monitoring tools and alerts. The SRE role emphasizes automation, incident response, and system resilience, making it broader in scope.

What engineers make $300,000 a year?

Senior Site Reliability Engineers and Observability Engineers with extensive experience, advanced skills in cloud platforms, automation, and monitoring tools can earn $300,000 or more annually. High compensation often correlates with working at large tech companies, possessing specialized certifications, and taking on leadership or highly technical roles.
What job categories do people searching Observability Site Reliability Engineer jobs in Tennessee look for? The top searched job categories for Observability Site Reliability Engineer jobs in Tennessee are:
What cities in Tennessee are hiring for Observability Site Reliability Engineer jobs? Cities in Tennessee with the most Observability Site Reliability Engineer job openings:
Systems Engineer - SRE Enablement

Systems Engineer - SRE Enablement

AutoZone

Memphis, TN • On-site

$55.50 - $73.75/hr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Posted 8 days ago


AutoZone rating

5.2

Company rating: 5.2 out of 10

Based on 1,904 frontline employees who took The Breakroom Quiz

35th of 39 rated national retailers


Job description


AutoZone's Site Reliability Engineering (SRE) team is seeking a Systems Engineer with a focus on SRE Enablement. This position is responsible for promoting reliability and operational excellence throughout the engineering organization. The successful candidate will play a key role in establishing standards, developing shared tools, providing guidance to development teams, and cultivating a culture of reliability across our hybrid infrastructure, which includes a primary emphasis on the Google Cloud Platform (GCP), as well as on-premises servers and applications.
The SRE Enablement Engineer collaborates closely with application, infrastructure, and architecture teams to integrate SRE best practices early in the software development lifecycle and ensure platforms adhere to rigorous production readiness standards.
Responsibilities
  • Define enterprise-wide reliability standards, Service-Level Objective (SLO) frameworks, and error budget policies.
  • Establish production readiness criteria that teams must meet prior to launching and conduct production readiness reviews across teams.
  • Own, document, and maintain the internal SRE handbook and reliability playbooks.
  • Build, maintain, and standardize shared observability platforms, specifically leveraging Dynatrace to be consumed by all engineering teams.
  • Provide templates for alerting, dashboards, and runbooks across both cloud and on-premises application workloads.
  • Participate in the incident management process, including post-mortem analysis, to continuously strengthen systemic reliability.
  • Run SRE training programs and reliability workshops for engineering teams.
  • Coach and mentor teams on SLO-based thinking and error budget management.
  • Embed proactive SRE practices and a continuous improvement mindset into the broader engineering culture.
  • Track and report reliability metrics across the enterprise, rather than just a single service.
  • Identify systemic reliability gaps and trends across cross-functional teams.
  • Report organizational reliability health to leadership and hold teams accountable to agreed-upon operational standards.
  • Act as an internal consultant during architecture and system design reviews.
  • Advise development teams on reliability design patterns (e.g., circuit breakers, retries, graceful degradation) suitable for a hybrid GCP and on-premises environment.
  • Engage early in new product development to influence system reliability from the outset.

Qualifications
  • Bachelor's degree in computer science, MIS, Information Technology, or a related field, or equivalent practical experience.
    4 to 7 years of experience in Systems Engineering, DevOps, or SRE-related roles.
  • Deep understanding of Site Reliability Engineering principles, particularly regarding SLOs, SLIs, error budgets, and TOIL reduction.
  • Hands-on experience building, administering, and optimizing observability and APM pipelines, with a strong focus on Dynatrace.
  • Strong experience deploying and supporting workloads in Google Cloud Platform (GCP), as well as maintaining legacy on-premises servers and applications.
  • Experience with container orchestration platforms (e.g., Kubernetes).
  • Strong programming/scripting skills (e.g., Python, Golang, Java) and experience with IaC tools (e.g., Terraform, Ansible).
  • Exceptional communication and consulting skills, with the ability to influence architecture decisions and translate technical concepts to non-technical leadership.

About Us
Since opening our first store in 1979, AutoZone has grown into a leading retailer and distributor of automotive parts and accessories across the Americas. Our customer-first mindset and commitment to Going the Extra Mile define who we are, for both our customers and AutoZoners. Working at AutoZone means being part of a team that values dedication, teamwork, and growth. Whether you're helping customers or building your career, we provide tools and support to help you succeed and drive your future.
Benefits at AutoZone
AutoZone offers thoughtful benefits programs with one-on-one benefits guidance designed to improve AutoZoners' physical, mental and financial well-being.
All AutoZoners (Full-Time and Part-Time):
  • Competitive pay
  • Unrivaled company culture
  • Medical, dental and vision plans
  • Exclusive discounts and perks, including an AutoZone in-store discount
  • 401(k) with company match and Stock Purchase Plan
  • AutoZoners Living Well Program for free mental health support
  • Opportunities for career growth

Additional Benefits for Full-Time AutoZoners:
  • Paid time off
  • Life, and short- and long-term disability insurance options
  • Health Savings and Flexible Spending Accounts with wellness rewards
  • Tuition reimbursement

Minimum age requirements may apply. Eligibility and waiting period requirements may apply; benefits for AutoZoners in Puerto Rico, Hawaii, or the U.S. Virgin Islands may differ. Learn more about all that AutoZone has to offer at Careers.AutoZone.com.
We proudly support Veterans, Active-duty Service Members, Reservists, National Guard and Military Families. Your experience is highly valued, and we encourage you to apply to join our team.
Online Application:
An online application is required. Click the Apply button to complete your application. For step-by-step instructions on how to apply visit careers.autozone.com/candidateresources.
AutoZone, and its subsidiary, ALLDATA are equal opportunity employers. All applicants will be considered for employment without attention to age, race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status, or any other legally protected categories.

What AutoZone employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


AutoZone logo

About AutoZone

Sourced by ZipRecruiter

AutoZone Inc (AutoZone) is a retailer and distributor of automotive replacement parts and accessories. The company provides new and remanufactured automotive hard parts, maintenance items, accessories, and non-automotive products. AutoZone sells automotive diagnostic and repair software through its subsidiary ALLDATA.

Industry

Motor vehicle and motor vehicle parts wholesalers

Company size

10,000+ Employees

Headquarters location

Memphis, TN, US

Year founded

1979