1

Director Site Reliability Engineering Jobs in Chelsea, AL

next page

Showing results 1-20

Director Site Reliability Engineering information

See Chelsea, AL salary details

$9

$57

$83

How much do director site reliability engineering jobs pay per hour?

As of Aug 30, 2026, the average hourly pay for director site reliability engineering in Chelsea, AL is $57.80, according to ZipRecruiter salary data. Most workers in this role earn between $49.71 and $66.06 per hour, depending on experience, location, and employer.

What is a director site reliability engineering?

A Director of Site Reliability Engineering (SRE) leads teams responsible for ensuring the availability, performance, and scalability of software systems. They define reliability best practices, drive automation, and collaborate with engineering and product teams to improve system resilience. This role requires strong leadership, technical expertise, and a focus on balancing innovation with operational stability.

What are the key skills and qualifications needed to thrive as a director site reliability engineering?

To thrive as a Director Site Reliability Engineering, you need extensive experience in software engineering, infrastructure management, incident response, and people leadership, often supported by a degree in computer science or a related field. Familiarity with cloud platforms (such as AWS, GCP, or Azure), automation tools (Terraform, Ansible), monitoring systems (Prometheus, Datadog), and relevant certifications like CKA or AWS Solutions Architect is valued. Outstanding communication, stakeholder management, and strategic vision are key soft skills that set leaders apart in this role. These abilities ensure the reliability, scalability, and efficiency of critical systems while effectively guiding and motivating technical teams.

What are the main challenges faced by a director site reliability engineering, and how can I prepare for them?

A Director of Site Reliability Engineering often encounters challenges such as balancing rapid feature delivery with system stability, managing complex incident responses, and fostering a culture of continuous improvement. Additionally, aligning reliability goals with business objectives and securing cross-functional buy-in can be demanding. To prepare, it is helpful to gain experience in high-scale system management, develop strong leadership and communication abilities, and cultivate a proactive approach to risk management and automation. Staying up to date with the latest SRE practices and building relationships with both engineering and business teams will also support your success in this pivotal role.

How much do Director Site Reliability Engineers get paid?

Director Site Reliability Engineers typically earn between $130,000 and $200,000 annually, depending on experience, location, and company size. They often oversee large teams, require advanced skills in cloud platforms and automation tools, and may receive bonuses or stock options as part of compensation.

What cities near Chelsea, AL are hiring for Director Site Reliability Engineering jobs?

Cities near Chelsea, AL with the most Director Site Reliability Engineering job openings:

Manager of Site Reliability Engineering (SRE)

Birmingham, AL • Hybrid


GPC - Genuine Parts Company
Retail • 10K+ employees

7.3

Company rating: 7.3 out of 10

Based on 63 frontline employees who took The Breakroom Quiz

215th of 426 rated retail wholesalers

People enjoy working here

Good employer

Paid breaks


$53.50 - $71/hr

Full-time

Re-posted 24 days ago


Job description

SUMMARY:

The Manager of Site Reliability Engineering leads and develops a team of SRE practitioners focused on delivering highly reliable, scalable, and performant cloud-based infrastructure and services. This role ensures the implementation of SRE principles, drives automation, observability, and incident management practices to enhance system reliability, and collaborates across development and operations teams to support continuous delivery and robust cloud platform operations.

You must be eligible to work in the US without Visa Sponsorship

JOB DUTIES

Lead, mentor, and grow a high-performing team of Site Reliability Engineers, fostering a culture of ownership, continuous improvement, and operational excellence.

Implement and champion Site Reliability Engineering principles and DevOps best practices within the team to ensure service reliability, availability, and performance.

Define and track key SRE metrics such as service uptime, incident response and resolution times.

Drive automation efforts including CI/CD pipeline enhancements, infrastructure-as-code practices, and self-service infrastructure provisioning to increase deployment velocity while reducing manual toil.

Own and continuously improve observability practices including system monitoring, logging, alerting, and diagnostics to ensure rapid issue detection and resolution.

Participate in incident response processes including incident management, root cause analysis, post-mortems, and continuous improvement to enhance system resilience.

Partner closely with software engineering, product management, architecture, and security teams to embed reliability and security early in the software development lifecycle (SDLC).

Oversee the management and scalability of cloud infrastructure environments, primarily on Google Cloud Platform (GCP), with a focus on Kubernetes, container orchestration, and hybrid cloud integrations.

Advocate for and apply best practices in performance tuning, capacity planning, and system design for high availability.

Develop and execute a long-term roadmap for our hybrid cloud platform, aligning with evolving business objectives and technology trends.

Establish and monitor key performance indicators (KPIs) service level indicators (SLIs) and service level objectives (SLOs) to drive system health and stability.

EDUCATION & EXPERIENCE

Typically requires a bachelor's degree and 7 years of experience in a technology and/or software engineering role or an equivalent combination

KNOWLEDGE, SKILLS, ABILITIES

Experience & Leadership

Proven experience working in large, complex enterprise environments (Fortune 500 or equivalent).

Site Reliability Engineering & DevOps Practices

Strong understanding and demonstrated implementation of Site Reliability Engineering (SRE) principles at scale.

Hands-on experience with infrastructure-as-code (IaC) tools such as Terraform, and ArgoCD.

In-depth knowledge and practical experience with CI/CD pipelines and automation of software delivery.

Championing DevOps practices and embedding reliability early in the SDLC.

Significant hands-on experience in Site Reliability Engineering or related roles focused on cloud infrastructure reliability.

Strong software engineering background with proficiency in infrastructure-as-code tools (e.g., Terraform, ArgoCD) and CI/CD automation.

Deep knowledge of cloud platforms, specifically Google Cloud Platform (GCP), Kubernetes, container orchestration, and cloud-native architecture.

Familiarity with monitoring and observability tools such as Dynatrace, Datadog, or equivalents.

Experience managing high-availability systems in 24/7 operational environments.

Ability to collaborate cross-functionally and drive alignment across engineering, product, and security teams.

Tools & Monitoring

Experience with monitoring, logging, and observability platforms.

Familiarity with incident management and performance monitoring tools, including Dynatrace and Datadog.

Proficient in cloud deployment tooling and automation frameworks.

Experience with Azure DevOps (ADO) or equivalent CI/CD tools.

Core Technical Skills

Strong software engineering and infrastructure background.

Solid understanding of Kubernetes, container orchestration, cluster management, and elastic scalability.

Experience with API-driven, event driven and microservices architectures.

Skilled in performance diagnostics, capacity planning, tuning, and system architecture for high-availability systems.

Not the right fit? Let us know you're interested in a future opportunity by joining our Talent Community on jobs.genpt.com or create an account to set up email alerts as new job postings become available that meet your interest!

GPC conducts its business without regard to sex, race, creed, color, religion, marital status, national origin, citizenship status, age, pregnancy, sexual orientation, gender identity or expression, genetic information, disability, military status, status as a veteran, or any other protected characteristic. GPC's policy is to recruit, hire, train, promote, assign, transfer and terminate employees based on their own ability, achievement, experience and conduct and other legitimate business reasons.



What Genuine Parts Company employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom