1

Sre Manager Jobs in Alabama (NOW HIRING)

.**Site Reliability Engineer II****SUMMARY:**Under general supervision, the Site Reliability Systems ... Manages the load configuration of a central data communication processor under limited guidance and ...

Reliability Engineer

Mcintosh, AL · On-site

$109K - $160K/yr

... Site Maintenance personnel, and others to accomplish this goal. Reliability Engineer Essential ... Basic understanding of project management, instrumentation, and equipment. * Proficiency in ERP ...

... Site Maintenance personnel, and others to accomplish this goal. Reliability Engineer Essential ... Basic understanding of project management, instrumentation, and equipment. * Proficiency in ERP ...

next page

Showing results 1-20

Sre Manager information

See Alabama salary details

$56.2K

$106.5K

$152.7K

How much do sre manager jobs pay per year?

As of Aug 19, 2026, the average yearly pay for sre manager in Alabama is $106,489.00, according to ZipRecruiter salary data. Most workers in this role earn between $85,700.00 and $126,900.00 per year, depending on experience, location, and employer.

What is an SRE manager?

SRE Managers are leaders responsible for overseeing Site Reliability Engineering (SRE) teams. They ensure the reliability, scalability, and performance of software systems by guiding engineers in implementing best practices, automation, and monitoring processes. SRE Managers collaborate closely with development and operations teams to balance feature development with system stability. Their role also includes mentoring SREs, managing incident response, and driving improvements in system reliability and operational efficiency.

What are some common challenges SRE managers face when leading Site Reliability Engineering teams?

SRE Managers often encounter challenges balancing reliability with rapid development, ensuring their teams have the right mix of software engineering and operations skills. They must also foster a culture of continuous improvement while managing on-call rotations and incident response without causing burnout. Additionally, collaborating effectively with development and product teams to set realistic service level objectives (SLOs) and drive adoption of SRE best practices can require strong communication and negotiation skills.

What are the key skills and qualifications needed to thrive as an SRE manager, and why are they important?

To thrive as an SRE Manager, you need a deep understanding of site reliability engineering principles, strong experience with systems architecture, and a background in computer science or a related field. Familiarity with tools such as Kubernetes, Prometheus, cloud platforms, and CI/CD pipelines, as well as certifications like AWS Certified Solutions Architect, are commonly expected. Leadership, effective communication, and problem-solving skills are crucial for driving team performance and collaborating across departments. These skills ensure high system reliability, efficient incident management, and a culture of continuous improvement within technical organizations.

What is the difference between Sre Manager vs DevOps Engineer?

AspectSre ManagerDevOps Engineer
CredentialsTypically requires a Bachelor's/Master's in CS or related field, with certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentLeads teams, manages incident response, and oversees reliability strategiesFocuses on automation, CI/CD pipelines, and infrastructure deployment
Industry UsageCommon in large tech companies, financial services, and cloud providersWidely used across startups, tech firms, and enterprises adopting DevOps practices

The Sre Manager and DevOps Engineer roles share overlapping skills in cloud computing, automation, and infrastructure management. While the Sre Manager oversees reliability and team coordination, the DevOps Engineer focuses on implementing automation tools and deployment pipelines. Both roles are crucial for modern IT operations, but the Sre Manager typically has a broader leadership responsibility, whereas the DevOps Engineer is more hands-on with technical implementation.

What does an SRE manager do?

An SRE manager oversees the reliability and performance of software systems, leading teams that implement automation, monitoring, and incident response processes. They coordinate efforts to ensure system availability, scalability, and efficiency, often using tools like monitoring dashboards and incident management platforms. Strong leadership, technical expertise, and understanding of service level objectives are essential in this role.

What job categories do people searching Sre Manager jobs in Alabama look for?

The top searched job categories for Sre Manager jobs in Alabama are:

Infographic showing various Sre Manager job openings in Alabama as of August 2026, with employment types broken down into 42% Full Time, and 58% Contract. Highlights an 74% In-person, and 26% Remote job distribution, with an average salary of $106,489 per year, or $51.2 per hour.

Manager of Site Reliability Engineering (SRE)

Motion

Birmingham, AL • On-site

$130 - $160/hr

Other

Posted 14 days ago


Job description

.Manager of Site Reliability Engineering (SRE) page is loaded## Manager of Site Reliability Engineering (SRE)remote type: Hybridlocations: Birmingham, AL, USAtime type: Full timeposted on: Posted Todayjob requisition id: R26\_0000009885SUMMARY:The Manager of Site Reliability Engineering leads and develops a team of SRE practitioners focused on delivering highly reliable, scalable, and performant cloud-based infrastructure and services. This role ensures the implementation of SRE principles, drives automation, observability, and incident management practices to enhance system reliability, and collaborates across development and operations teams to support continuous delivery and robust cloud platform operations.You must be eligible to work in the US without Visa SponsorshipJOB DUTIES• Lead, mentor, and grow a high-performing team of Site Reliability Engineers, fostering a culture of ownership, continuous improvement, and operational excellence.• Implement and champion Site Reliability Engineering principles and DevOps best practices within the team to ensure service reliability, availability, and performance.• Define and track key SRE metrics such as service uptime, incident response and resolution times.• Drive automation efforts including CI/CD pipeline enhancements, infrastructure-as-code practices, and self-service infrastructure provisioning to increase deployment velocity while reducing manual toil.• Own and continuously improve observability practices including system monitoring, logging, alerting, and diagnostics to ensure rapid issue detection and resolution.• Participate in incident response processes including incident management, root cause analysis, post-mortems, and continuous improvement to enhance system resilience.• Partner closely with software engineering, product management, architecture, and security teams to embed reliability and security early in the software development lifecycle (SDLC).• Oversee the management and scalability of cloud infrastructure environments, primarily on Google Cloud Platform (GCP), with a focus on Kubernetes, container orchestration, and hybrid cloud integrations.• Advocate for and apply best practices in performance tuning, capacity planning, and system design for high availability.• Develop and execute a long-term roadmap for our hybrid cloud platform, aligning with evolving business objectives and technology trends.• Establish and monitor key performance indicators (KPIs) service level indicators (SLIs) and service level objectives (SLOs) to drive system health and stability.EDUCATION & EXPERIENCETypically requires a bachelor's degree and 7 years of experience in a technology and/or software engineering role or an equivalent combinationKNOWLEDGE, SKILLS, ABILITIESExperience & Leadership• Proven experience working in large, complex enterprise environments (Fortune 500 or equivalent).Site Reliability Engineering & DevOps Practices• Strong understanding and demonstrated implementation of Site Reliability Engineering (SRE) principles at scale.• Hands-on experience with infrastructure-as-code (IaC) tools such as Terraform, and ArgoCD.• In-depth knowledge and practical experience with CI/CD pipelines and automation of software delivery.• Championing DevOps practices and embedding reliability early in the SDLC.• Significant hands-on experience in Site Reliability Engineering or related roles focused on cloud infrastructure reliability.• Strong software engineering background with proficiency in infrastructure-as-code tools (e.g., Terraform, ArgoCD) and CI/CD automation.• Deep knowledge of cloud platforms, specifically Google Cloud Platform (GCP), Kubernetes, container orchestration, and cloud-native architecture.• Familiarity with monitoring and observability tools such as Dynatrace, Datadog, or equivalents.• Experience managing high-availability systems in 24/7 operational environments.• Ability to collaborate cross-functionally and drive alignment across engineering, product, and security teams.Tools & Monitoring• Experience with monitoring, logging, and observability platforms.• Familiarity with incident management and performance monitoring tools, including Dynatrace and Datadog.• Proficient in cloud deployment tooling and automation frameworks.• Experience with Azure DevOps (ADO) or equivalent CI/CD tools.Core Technical Skills• Strong software engineering and infrastructure background.• Solid understanding of Kubernetes, container orchestration, cluster management, and elastic scalability.• Experience with API-driven, event driven and microservices architectures.• • Skilled in performance diagnostics, capacity planning, tuning, and system architecture for high-availability systems.or create an account to set up email alerts as new job postings become available that meet your interest!GPC conducts its business without regard to sex, race, creed, color, religion, marital status, national origin, citizenship status, age, pregnancy, sexual orientation, gender identity or expression, genetic information, disability, military status, status as a veteran, or any other protected characteristic. GPC's policy is to recruit, hire, train, promote, assign, transfer and terminate employees based on their own ability, achievement, experience and conduct and other legitimate business reasons.remote type: Hybridlocations: Birmingham, AL, USAtime type: Full timeposted on: Posted 30+ Days Ago #J-18808-Ljbffr