1

Site Reliability Engineer Manager Jobs in Georgia

Site Reliability Engineer (SRE)

Atlanta, GA · On-site

$54.75 - $72.75/hr

Job Summary : eTeam is a company seeking a Site Reliability Engineer (SRE) for a contract position ... management, batch opsIntegrate automation with ServiceNow workflows • Maintain reusable ...

Site Reliability Engineer - SRE

Atlanta, GA · On-site +1

$54.25 - $72/hr

Site Reliability Engineer * Location: Atlanta, GA OR Dallas OR Austin, TX * Duration: Long Term or 6+ Months contract to Hire Note: Remote Possible, however candidates will move to work onsite/Hybrid ...

Site Reliability Engineer - SRE

Atlanta, GA · On-site

$54.25 - $72/hr

Site Reliability Engineer * Location: Atlanta, GA OR Dallas OR Austin, TX * Duration: Long Term or 6+ Months contract to Hire Note: Remote Possible, however candidates will move to work onsite/Hybrid ...

Site Reliability engineer (SRE)

Atlanta, GA · On-site

$54.75 - $72.75/hr

Site Reliability engineer(SRE) Location: Atlanta, GA ( Hybrid - 3days Office - 2 days WFH) Duration ... We specialize in Big Data & Analytics, Digital Transformation, IT Service Management, Cognitive ...

SRE Architect

Atlanta, GA · On-site

$54.75 - $72.75/hr

SRE Architect Thought Leadership & Enterprise Resilience Job Summary Mphasis is seeking a visionary ... Lead capacity management and elastic infrastructure design to accommodate irregular traffic bursts ...

Site Reliability Engineer

Atlanta, GA · Remote

$54.75 - $72.75/hr

Site Reliability Engineer Company: AutoRABIT Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, GCP, Azure, Kubernetes, Docker, Ansible, Jenkins, Terraform ...

Site Reliability Engineer

Atlanta, GA · On-site +1

$100K - $120K/yr

Strong knowledge of SRE best practices and incident management protocols * Deep experience using and/or configuring New Relic, Data Dog, SumoLogic or similar observability tools * Proficiency in ...

Infrastructure & Automation Design, deploy, and manage cloud infrastructure across AWS and Azure ... SRE best practices across the organization Lead Kubernetes adoption efforts and educate teams on ...

Site Reliability Engineer

Atlanta, GA · On-site +1

$100K - $120K/yr

Strong knowledge of SRE best practices and incident management protocols * Deep experience using and/or configuring New Relic, Data Dog, SumoLogic or similar observability tools * Proficiency in ...

Site Reliability Engineer

Alpharetta, GA · On-site

$55.75 - $74/hr

I have an opportunity for a " Site Reliability Engineer " - Alpharetta, GA (Onsite). and I am looking for a candidate who can join Immediately if you are interested, reply to me with your updated ...

SRE Lead/ Architect

Atlanta, GA · On-site

$54.75 - $72.75/hr

Deep understanding and practical application of SRE principles (SLIs/SLOs, error budgets, toil reduction, automation, incident management, postmortems) * Expertise in cloud computing platforms (e.g ...

next page

Showing results 1-20

Site Reliability Engineer Manager information

See Georgia salary details

$9

$53

$77

How much do site reliability engineer manager jobs pay per hour?

As of Jul 27, 2026, the average hourly pay for site reliability engineer manager in Georgia is $53.82, according to ZipRecruiter salary data. Most workers in this role earn between $46.30 and $61.49 per hour, depending on experience, location, and employer.

Will AI replace SRE jobs?

AI is expected to augment Site Reliability Engineer (SRE) roles by automating routine tasks such as monitoring, incident response, and data analysis, allowing SREs to focus on complex problem-solving and system design. While AI can improve efficiency, it is unlikely to fully replace SREs, as human expertise is essential for managing system architecture, making strategic decisions, and handling unforeseen issues.

What is a Site Reliability Engineer Manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What engineer makes $500,000 a year?

A senior or principal Site Reliability Engineer (SRE) with extensive experience, specialized skills, and often working at large tech companies or in high-cost-of-living areas can earn $500,000 or more annually. Compensation may include base salary, bonuses, and stock options, especially for those in leadership or highly technical roles. Advanced certifications and expertise in cloud platforms, automation, and system architecture are common among top earners in this field.

Is SRE a stressful job?

Site Reliability Engineer (SRE) roles can be stressful due to the high responsibility for system uptime, incident response, and maintaining service reliability. The job often involves working under pressure, handling outages, and balancing automation with manual intervention, but it also offers opportunities for skill development and process improvement. Effective SREs use monitoring tools and incident management practices to manage stress and ensure system stability.

What is the role of site reliability engineer manager?

A Site Reliability Engineer Manager oversees a team responsible for maintaining the reliability, availability, and performance of software systems. They coordinate incident response, implement automation, and ensure system scalability, often using tools like monitoring and alerting platforms. The role requires strong leadership, technical expertise, and knowledge of cloud infrastructure and DevOps practices.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How does a Site Reliability Engineer Manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What are the key skills and qualifications needed to thrive as a Site Reliability Engineer Manager, and why are they important?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.
What are the most commonly searched types of Site Reliability Engineer jobs in Georgia? The most popular types of Site Reliability Engineer jobs in Georgia are:
What cities in Georgia are hiring for Site Reliability Engineer Manager jobs? Cities in Georgia with the most Site Reliability Engineer Manager job openings:
Infographic showing various Site Reliability Engineer Manager job openings in Georgia as of July 2026, with employment types broken down into 93% Full Time, 4% Part Time, and 3% Contract. Highlights an 87% Physical, 5% Hybrid, and 8% Remote job distribution, with an average salary of $111,951 per year, or $53.8 per hour.
Site Reliability Engineer (SRE)

Site Reliability Engineer (SRE)

Merican Inc

Atlanta, GA • On-site

$54.75 - $72.75/hr

Other

Posted 19 days ago


Job description

Job Title: Site Reliability Engineer (SRE)

Location: Atlanta, GA (Hybrid)
Local Candidates Only

Work Schedule

Hybrid work model with rotating teams:

  • Team A: Onsite Thursday & Friday; Remote Monday Wednesday
  • Team B: Remote Thursday & Friday; Onsite Monday Wednesday
  • Teams rotate schedules weekly.
Job Summary

We are seeking a skilled Site Reliability Engineer (SRE) to support highly available, business-critical applications across on-premises and AWS cloud environments. The ideal candidate will have strong expertise in DevOps, cloud infrastructure, automation, CI/CD, monitoring, and troubleshooting complex production systems. This role offers the opportunity to work with modern cloud technologies and contribute to the reliability, scalability, and performance of enterprise applications.

Key Responsibilities
  • Manage and optimize data streaming and API components in OpenShift (On-Premises) and AWS.
  • Review application APIs and processes to identify performance optimization opportunities.
  • Automate testing, including data quality validation, production deployments, and release processes.
  • Develop integrations between on-premises, AWS, and third-party tools such as ServiceNow, VersionOne, and Sumo Logic.
  • Collaborate with teams to define and implement SLIs and SLOs.
  • Monitor production environments, troubleshoot performance issues, conduct root cause analysis, and document findings.
  • Design, build, and maintain CI/CD pipelines for application artifacts, APIs, and data processing jobs.
  • Configure monitoring, alerting, and observability solutions to enable proactive issue detection.
  • Implement AWS security best practices, including IAM, HSM, encryption, and access controls.
  • Monitor cloud costs, generate usage reports, and recommend cost optimization strategies.
  • Design and implement solutions to address security vulnerabilities and compliance requirements.
  • Analyze infrastructure capacity and performance to support scalable and resilient systems.
  • Develop backup and disaster recovery strategies for critical applications and data.
  • Collaborate with architecture, infrastructure, and application teams to continuously improve system performance, reliability, and security.
Required Skills
  • Strong experience with AWS cloud services and cloud operations.
  • Hands-on experience with OpenShift, CloudFormation, Terraform, Ansible, Shell scripting, and Python.
  • Experience with Linux administration and enterprise infrastructure.
  • Knowledge of virtualization, networking, load balancers, firewalls, storage, backup, and monitoring tools.
  • Experience with CI/CD tools such as GitLab, GitHub, Jenkins, Maven, Gradle, and Nexus.
  • Experience with Software Release Management.
  • Strong troubleshooting and incident management skills for mission-critical systems.
  • Experience with automation, infrastructure orchestration, and configuration management.
Preferred Qualifications
  • Bachelor's degree in Computer Science or a related technical field (or equivalent experience).
  • 3+ years of DevOps/SysOps engineering experience with a focus on AWS.
  • 2+ years of application development experience involving data streaming and high-availability applications.
  • 1+ year of experience in a Site Reliability Engineering (SRE) environment preferred.
  • Overall 4 6 years of IT experience.

Merican logo

About Merican

Sourced by ZipRecruiter

Merican is a IT Service consulting firm, specialized in Digital adoption and Business automation. With our diverse collection of skilled and committed consultants, technology companies, businesses and digital experts, we provide our subject expertise and our unique client service approach, a best-in-class global model of delivery suited to the business demands of our clients. We ensure that we implement future-oriented solutions for our clients via investments in people, solutions, technologies, competencies and infrastructure.

Industry

It services

Company size

51 - 200 Employees

Headquarters location

Columbia , MD, US

Year founded

2020

Social media