1

Director Site Reliability Engineering Jobs in Georgia

Senior Site Reliability Engineer

Atlanta, GA · On-site

$54.75 - $72.75/hr

RESPONSIBILITIES Reliability Engineering * Define and manage SLIs, SLOs, and Error Budgets for ... Mentor engineers on SRE best practices * Promote a culture of engineering-driven reliability over ...

SRE

Atlanta, GA · On-site

$54.75 - $72.75/hr

Leverage deep expertise in software engineering, distributed systems, cloud infrastructure, and SRE principles to influence technology choices, establish best practices, and foster a proactive ...

Site Reliability Engineer

Atlanta, GA · On-site

$54.75 - $72.75/hr

Pyramid Consulting, Inc. is seeking a Site Reliability Engineer for a 12+ months contract ... The role involves engineering software within an AWS cloud infrastructure and designing, building ...

Site Reliability Engineer (SRE)

Atlanta, GA · On-site

$54.75 - $72.75/hr

Job Summary : eTeam is a company seeking a Site Reliability Engineer (SRE) for a contract position ... Bachelor's or Master of Engineering • Must have skills: Ansible • Must have skills: GitLab Duo ...

Site Reliability Engineer

Alpharetta, GA

$55.75 - $74/hr

Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to ... Expect real engineering deep dives, not top-down mandates. We love digging into a hard problem ...

Site Reliability Engineer

Atlanta, GA

$54.75 - $72.75/hr

Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to ... Expect real engineering deep dives, not top-down mandates. We love digging into a hard problem ...

Site Reliability Engineer III

Alpharetta, GA · On-site

$55.75 - $74/hr

You can learn more about LexisNexis Risk at About our Team Our Site Reliability Engineering (SRE) team plays a critical role in ensuring the reliability, availability, security, and performance of ...

Site Reliability Engineer III

Alpharetta, GA · On-site

$55.75 - $74/hr

You can learn more about LexisNexis Risk at About our Team Our Site Reliability Engineering (SRE) team plays a critical role in ensuring the reliability, availability, security, and performance of ...

Site Reliability Engineer III

Alpharetta, GA · On-site

$55.75 - $74/hr

You can learn more about LexisNexis Risk at About our Team Our Site Reliability Engineering (SRE) team plays a critical role in ensuring the reliability, availability, security, and performance of ...

Site Reliability Engineer

Atlanta, GA · On-site +1

$100K - $120K/yr

Overview The Site Reliability Engineer is a key force behind improving Origami's time to resolution ... Provides an actionable feedback loop to Observability and Engineering teams toward improving MELT ...

SRE Engineer

Atlanta, GA · On-site

$54.75 - $72.75/hr

Leverage deep expertise in software engineering, distributed systems, cloud infrastructure, and SRE principles to influence technology choices, establish best practices, and foster a proactive ...

Showing results 21-40

Director Site Reliability Engineering information

What is a director site reliability engineering?

A Director of Site Reliability Engineering (SRE) leads teams responsible for ensuring the availability, performance, and scalability of software systems. They define reliability best practices, drive automation, and collaborate with engineering and product teams to improve system resilience. This role requires strong leadership, technical expertise, and a focus on balancing innovation with operational stability.

What are the key skills and qualifications needed to thrive as a director site reliability engineering?

To thrive as a Director Site Reliability Engineering, you need extensive experience in software engineering, infrastructure management, incident response, and people leadership, often supported by a degree in computer science or a related field. Familiarity with cloud platforms (such as AWS, GCP, or Azure), automation tools (Terraform, Ansible), monitoring systems (Prometheus, Datadog), and relevant certifications like CKA or AWS Solutions Architect is valued. Outstanding communication, stakeholder management, and strategic vision are key soft skills that set leaders apart in this role. These abilities ensure the reliability, scalability, and efficiency of critical systems while effectively guiding and motivating technical teams.

What are the main challenges faced by a director site reliability engineering, and how can I prepare for them?

A Director of Site Reliability Engineering often encounters challenges such as balancing rapid feature delivery with system stability, managing complex incident responses, and fostering a culture of continuous improvement. Additionally, aligning reliability goals with business objectives and securing cross-functional buy-in can be demanding. To prepare, it is helpful to gain experience in high-scale system management, develop strong leadership and communication abilities, and cultivate a proactive approach to risk management and automation. Staying up to date with the latest SRE practices and building relationships with both engineering and business teams will also support your success in this pivotal role.

How much do Director Site Reliability Engineers get paid?

Director Site Reliability Engineers typically earn between $130,000 and $200,000 annually, depending on experience, location, and company size. They often oversee large teams, require advanced skills in cloud platforms and automation tools, and may receive bonuses or stock options as part of compensation.

What job categories do people searching Director Site Reliability Engineering jobs in Georgia look for?

The top searched job categories for Director Site Reliability Engineering jobs in Georgia are:

What cities in Georgia are hiring for Director Site Reliability Engineering jobs?

Cities in Georgia with the most Director Site Reliability Engineering job openings:

Infographic showing various Director Site Reliability Engineering job openings in Georgia as of August 2026, with employment types broken down into 1% As Needed, 81% Full Time, 13% Part Time, 1% Temporary, 3% Contract, and 1% Nights. Highlights an 94% Physical, 2% Hybrid, and 4% Remote job distribution.

Senior Site Reliability Engineer

Inspire Brands

Atlanta, GA • On-site

$54.75 - $72.75/hr

Full-time

Posted 20 days ago


Inspire Brands rating

5.8

Company rating: 5.8 out of 10

Based on 56 frontline employees who took The Breakroom Quiz

29th of 107 rated fast food restaurants


Job description

Inspire Brands is hiring two Senior Site Reliability Engineers to help build and scale reliable, resilient, and observable systems supporting high-traffic, customer-facing digital platforms. These role blends software engineering, systems thinking, and operational excellence to reduce toil, prevent incidents, and improve system reliability at scale.

The ideal candidate has hands-on experience applying and implementing SRE principles - not just supporting production systems, but engineering reliability into them.

RESPONSIBILITIES

Reliability Engineering

  • Define and manage SLIs, SLOs, and Error Budgets for critical services
  • Drive production readiness reviews and reliability requirements into architecture and design
  • Perform capacity planning, failure mode analysis, and dependency risk assessments
  • Identify systemic reliability risks and drive remediation before they cause customer impact

Observability

  • Design monitoring, alerting, logging, and tracing solutions using modern observability tooling
  • Improve signal-to-noise ratio and reduce alert fatigue
  • Build dashboards and telemetry that reflect true service health, not just infrastructure metrics

Incident Management

  • Lead technical response for high-severity incidents
  • Drive blameless postmortems and root cause analysis focused on systemic fixes
  • Continuously improve detection, response, and recovery processes
  • Participate in an on-call rotation

Automation & Toil Reduction

  • Identify and eliminate manual, repetitive operational work through automation
  • Build self-healing systems, tooling, and scripts to reduce human intervention
  • Improve CI/CD pipelines and deployment safety (canary, rollback, blue-green)
  • Support Infrastructure as Code (Terraform, Bicep, or similar)

Performance & Scalability

  • Conduct load testing, performance benchmarking, and bottleneck analysis
  • Partner with engineering to design systems for horizontal scalability and fault tolerance

Collaboration & Culture

  • Partner with engineering teams to implement resiliency patterns (circuit breakers, retries, graceful degradation, rate limiting)
  • Mentor engineers on SRE best practices
  • Promote a culture of engineering-driven reliability over reactive operations

EDUCATION AND EXPERIENCE QUALIFICATIONS

Required Qualifications

  • 5+ years experience in Site Reliability Engineering, Software Engineering, or Platform Engineering
  • 2+ years experience with Kubernetes and containerized workloads
  • 4-year degree in Computer Scienceor related field

Preferred Qualifications

  • Experience with chaos engineering or resiliency testing
  • Experience with high-volume, high-availability transactional systems
  • Experience with AI-assisted observability or operational automation
  • Experience making meaningful contributions to internal SRE tooling, frameworks, or platforms

REQUIRED KNOWLEDGE, SKILLS, OR ABILITIES

  • Strong programming/scripting skills (Python, Go, Java, or Node.js)
  • Demonstrated experience defining and operating against SLOs/Error Budgets
  • Strong skills in leading incident response and root cause analysis for production systems
  • Solid understanding of distributed systems and microservices architecture
  • Deep knowledge and expertise in at least one major cloud platform (Azure, AWS, or GCP)
  • Expertise with observability platforms and monitoring strategy

This position is based in our Atlanta Support Center, with an expected on-site presence of 80%.


Inspire is a multi-brand restaurant company whose portfolio includes more than 33,300 Arby's, Baskin-Robbins, Buffalo Wild Wings, Dunkin', Jimmy John's, and SONIC restaurants worldwide. We're made up of some of the world's most iconic restaurant brands, but we're much more than just a restaurant company. We're a team of hundreds of thousands who individually and collectively are changing the way people eat, drink, and gather around the table. We know that food is much more than a staple-it's an experience. At Inspire, that's our purpose: to ignite and nourish flavorful experiences.

What Inspire Brands employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Inspire Brands logo

About Inspire Brands

Sourced by ZipRecruiter

Inspire Brands Inc., located in Atlanta, GA, United States, operates in the foodservice industry as a multi-brand restaurant company, making it among the biggest restaurant companies globally. Their portfolio includes well-known restaurant brands such as Arby's, Buffalo Wild Wings, Sonic, and Jimmy John's, reflecting their commitment to innovation and quality. Founded in 2018 as a result of a consolidation of various restaurant brands under one corporate umbrella, Inspire Brands was formed with a vision to invigorate excellent brands and supercharge their long-term growth.

Industry

Food services and drinking places

Company size

10,000+ Employees

Headquarters location

Atlanta, GA, US

Year founded

2018