1

Reliability Engineer Manager Jobs in Canton, GA (NOW HIRING)

Site Reliability Engineer

Alpharetta, GA ยท On-site

$55.75 - $74/hr

Atlanta-based Incident IQ is the leading workflow management platform built exclusively for K-12 ... Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to ...

Site Reliability Engineer

Atlanta, GA

$54.75 - $72.75/hr

Atlanta-based Incident IQ is the leading workflow management platform built exclusively for K-12 ... Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to ...

Site Reliability Engineer

Alpharetta, GA

$55.75 - $74/hr

Atlanta-based Incident IQ is the leading workflow management platform built exclusively for K-12 ... Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to ...

Site Reliability engineer (SRE)

Atlanta, GA ยท On-site

$54.75 - $72.75/hr

Site Reliability engineer(SRE) Location: Atlanta, GA ( Hybrid - 3days Office - 2 days WFH) Duration ... We specialize in Big Data & Analytics, Digital Transformation, IT Service Management, Cognitive ...

Site Reliability Engineer

Alpharetta, GA ยท On-site

$65 - $75/hr

Infrastructure & Automation Design, deploy, and manage cloud infrastructure across AWS and Azure ... E best practices across the organization Lead Kubernetes adoption efforts and educate teams on ...

Experience leading SRE Team/s, including rotating staff across products / platforms, mentoring and development, objective setting and general line management * Experience building and operating ...

Experience leading SRE Team/s, including rotating staff across products / platforms, mentoring and development, objective setting and general line management * Experience building and operating ...

Site Reliability Engineering Lead

Alpharetta, GA ยท On-site

$55.75 - $74/hr

Experience leading SRE Team/s, including rotating staff across products / platforms, mentoring and development, objective setting and general line management * Experience building and operating ...

Site Reliability Engineer

Atlanta, GA ยท On-site

$155K - $222K/yr

We are one of several SRE teams working together to support a platform that serves more than 500,000 customers and manages over 18 million devices worldwide. The team operates with a high degree of ...

SRE Lead/ Architect

Atlanta, GA ยท On-site

$54.75 - $72.75/hr

Deep understanding and practical application of SRE principles (SLIs/SLOs, error budgets, toil reduction, automation, incident management, postmortems) * Expertise in cloud computing platforms (e.g ...

Site Reliability Engineer

Atlanta, GA ยท On-site +1

$100K - $120K/yr

Strong knowledge of SRE best practices and incident management protocols * Deep experience using and/or configuring New Relic, Data Dog, SumoLogic or similar observability tools * Proficiency in ...

SRE/DevOps Engineer

Johns Creek, GA ยท On-site

$52.75 - $70.25/hr

The manager specifically wants engineers already incorporating AI into their daily development process--not someone simply familiar with AI concepts. Required Skills: DevOps / SRE: * Strong Site ...

Site Reliability Engineer

Atlanta, GA ยท On-site +1

$100K - $120K/yr

Strong knowledge of SRE best practices and incident management protocols * Deep experience using and/or configuring New Relic, Data Dog, SumoLogic or similar observability tools * Proficiency in ...

next page

Showing results 1-20

Reliability Engineer Manager information

See Canton, GA salary details

$57.6K

$111.4K

$133.1K

How much do reliability engineer manager jobs pay per year?

As of Aug 5, 2026, the average yearly pay for reliability engineer manager in Canton, GA is $111,388.00, according to ZipRecruiter salary data. Most workers in this role earn between $96,800.00 and $121,800.00 per year, depending on experience, location, and employer.

What does a reliability engineer manager do?

A Reliability Engineer Manager oversees teams responsible for improving the reliability and performance of systems, machinery, or processes within an organization. They develop maintenance strategies, lead root cause analyses of failures, and implement best practices to minimize downtime and costs. Additionally, they collaborate with other departments to ensure that reliability goals align with business objectives and compliance standards. Their role is crucial in industries such as manufacturing, energy, and technology, where system uptime and safety are critical.

What are some common challenges reliability engineer managers face when balancing long-term reliability improvements with immediate operational demands?

Reliability Engineer Managers often need to prioritize urgent maintenance issues while also driving long-term reliability initiatives. Balancing these competing demands can be challenging, as immediate equipment failures may require quick fixes that temporarily interrupt ongoing improvement projects. Effective managers work closely with operations, maintenance, and engineering teams to communicate priorities, allocate resources, and implement sustainable solutions that address root causes rather than just symptoms. This role typically involves using data-driven decision-making and fostering a culture of proactive maintenance and continuous improvement.

What are the key skills and qualifications needed to thrive as a reliability engineer manager?

To thrive as a Reliability Engineer Manager, you need a strong background in engineering principles, reliability analysis, and maintenance strategies, typically supported by a degree in engineering and experience in reliability roles. Familiarity with reliability-centered maintenance (RCM), failure mode and effects analysis (FMEA), and asset management software such as SAP or Maximo is common, along with certifications like Certified Reliability Engineer (CRE). Leadership, problem-solving, and effective communication are vital soft skills for managing teams and driving cross-functional initiatives. These competencies are crucial for minimizing downtime, optimizing equipment performance, and ensuring long-term operational efficiency.

What is the difference between Reliability Engineer Manager vs Reliability Engineer?

AspectReliability EngineerReliability Engineer Manager
Required CredentialsBachelor's in Engineering or related field; certifications like CRC, CRESame as Reliability Engineer, plus leadership experience
Work EnvironmentDesign, analyze, and improve system reliability; often in teamsOversees Reliability Engineers; manages projects and teams
Employer & Industry UsageManufacturing, aerospace, energy, automotiveSame industries, with added managerial responsibilities
Common Search & ComparisonFocuses on technical skills and hands-on reliability tasksFocuses on leadership, team management, and strategic planning

The main difference between a Reliability Engineer and a Reliability Engineer Manager lies in their responsibilities. The Reliability Engineer focuses on technical analysis and system improvements, while the Reliability Engineer Manager oversees teams, manages projects, and develops strategies to enhance reliability across the organization.

Infographic showing various Reliability Engineer Manager job openings in Canton, GA as of July 2026, with employment types broken down into 36% Full Time, 9% Temporary, and 55% Contract. Highlights an 91% In-person, and 9% Remote job distribution, with an average salary of $111,388 per year, or $53.6 per hour.

Site Reliability Engineer (SRE)

Merican Inc

Atlanta, GA โ€ข On-site

$54.75 - $72.75/hr

Other

Posted 28 days ago


Job description

Job Title: Site Reliability Engineer (SRE)

Location: Atlanta, GA (Hybrid)
Local Candidates Only

Work Schedule

Hybrid work model with rotating teams:

  • Team A: Onsite Thursday & Friday; Remote Monday Wednesday
  • Team B: Remote Thursday & Friday; Onsite Monday Wednesday
  • Teams rotate schedules weekly.
Job Summary

We are seeking a skilled Site Reliability Engineer (SRE) to support highly available, business-critical applications across on-premises and AWS cloud environments. The ideal candidate will have strong expertise in DevOps, cloud infrastructure, automation, CI/CD, monitoring, and troubleshooting complex production systems. This role offers the opportunity to work with modern cloud technologies and contribute to the reliability, scalability, and performance of enterprise applications.

Key Responsibilities
  • Manage and optimize data streaming and API components in OpenShift (On-Premises) and AWS.
  • Review application APIs and processes to identify performance optimization opportunities.
  • Automate testing, including data quality validation, production deployments, and release processes.
  • Develop integrations between on-premises, AWS, and third-party tools such as ServiceNow, VersionOne, and Sumo Logic.
  • Collaborate with teams to define and implement SLIs and SLOs.
  • Monitor production environments, troubleshoot performance issues, conduct root cause analysis, and document findings.
  • Design, build, and maintain CI/CD pipelines for application artifacts, APIs, and data processing jobs.
  • Configure monitoring, alerting, and observability solutions to enable proactive issue detection.
  • Implement AWS security best practices, including IAM, HSM, encryption, and access controls.
  • Monitor cloud costs, generate usage reports, and recommend cost optimization strategies.
  • Design and implement solutions to address security vulnerabilities and compliance requirements.
  • Analyze infrastructure capacity and performance to support scalable and resilient systems.
  • Develop backup and disaster recovery strategies for critical applications and data.
  • Collaborate with architecture, infrastructure, and application teams to continuously improve system performance, reliability, and security.
Required Skills
  • Strong experience with AWS cloud services and cloud operations.
  • Hands-on experience with OpenShift, CloudFormation, Terraform, Ansible, Shell scripting, and Python.
  • Experience with Linux administration and enterprise infrastructure.
  • Knowledge of virtualization, networking, load balancers, firewalls, storage, backup, and monitoring tools.
  • Experience with CI/CD tools such as GitLab, GitHub, Jenkins, Maven, Gradle, and Nexus.
  • Experience with Software Release Management.
  • Strong troubleshooting and incident management skills for mission-critical systems.
  • Experience with automation, infrastructure orchestration, and configuration management.
Preferred Qualifications
  • Bachelor's degree in Computer Science or a related technical field (or equivalent experience).
  • 3+ years of DevOps/SysOps engineering experience with a focus on AWS.
  • 2+ years of application development experience involving data streaming and high-availability applications.
  • 1+ year of experience in a Site Reliability Engineering (SRE) environment preferred.
  • Overall 4 6 years of IT experience.

Merican logo

About Merican

Sourced by ZipRecruiter

Merican is a IT Service consulting firm, specialized in Digital adoption and Business automation. With our diverse collection of skilled and committed consultants, technology companies, businesses and digital experts, we provide our subject expertise and our unique client service approach, a best-in-class global model of delivery suited to the business demands of our clients. We ensure that we implement future-oriented solutions for our clients via investments in people, solutions, technologies, competencies and infrastructure.

Industry

It services

Company size

51 - 200 Employees

Headquarters location

Columbia , MD, US

Year founded

2020

Social media