1

Observability Site Reliability Engineer Jobs in Georgia

Site Reliability Engineer

Atlanta, GA

$54.75 - $72.75/hr

Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to ... You'll work with leading-edge observability and reliability tooling, and the calls you make will ...

New

Site Reliability Engineer

Alpharetta, GA

$55.75 - $74/hr

Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to ... You'll work with leading-edge observability and reliability tooling, and the calls you make will ...

New

Site Reliability Engineer

Alpharetta, GA ยท On-site

$55.75 - $74/hr

Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to ... You'll work with leading-edge observability and reliability tooling, and the calls you make will ...

New

Site Reliability Engineer (SRE)

Atlanta, GA ยท On-site

$54.75 - $72.75/hr

... observability data to detect, predict, and prevent incidents โ€ข Define and maintain SLOs, SLIs ... SRE principles in a production environment โ€ข Hands-on experience with cloud platforms (AWS ...

Site Reliability Engineer (SRE)

Atlanta, GA ยท On-site

$54.75 - $72.75/hr

Configure monitoring, alerting, and observability solutions to enable proactive issue detection ... (SRE) environment preferred. * Overall 4 6 years of IT experience.

Site Reliability Engineer

Atlanta, GA ยท On-site +1

$100K - $120K/yr

Overview The Site Reliability Engineer is a key force behind improving Origami's time to resolution ... Maintains and configures core observability tools to ensure optimum performance and key metrics ...

SRE Lead/ Architect

Atlanta, GA ยท On-site

$54.75 - $72.75/hr

Mandatory skills are Observability, Resiliency, Chaos engineering, strong python, and Dynatrace As an SRE Architect, you will be a pivotal technical leader responsible for designing, building, and ...

Site Reliability Engineer

Alpharetta, GA ยท On-site

$65 - $75/hr

Observability & Monitoring Architect and implement comprehensive observability solutions using ... SRE best practices across the organization Lead Kubernetes adoption efforts and educate teams on ...

Site Reliability Engineer

Atlanta, GA ยท On-site +1

$100K - $120K/yr

Overview The Site Reliability Engineer is a key force behind improving Origami's time to resolution ... Maintains and configures core observability tools to ensure optimum performance and key metrics ...

Site Reliability Engineer

Atlanta, GA ยท On-site +1

$100K - $120K/yr

Overview The Site Reliability Engineer is a key force behind improving Origami's time to resolution ... Maintains and configures core observability tools to ensure optimum performance and key metrics ...

next page

Showing results 1-20

Observability Site Reliability Engineer information

What engineer makes $500,000 a year?

A senior or principal Site Reliability Engineer (SRE) or Observability Engineer with extensive experience, specialized skills, and working at large tech companies can earn $500,000 or more annually. Compensation often includes base salary, bonuses, and stock options, especially in high-demand markets and organizations with complex infrastructure.

Is AI replacing SRE?

AI is augmenting the work of Site Reliability Engineers (SREs) by automating tasks such as monitoring, incident detection, and response. However, SREs are still essential for designing systems, managing complex issues, and making strategic decisions that require human judgment. AI tools are considered complementary rather than replacements for SREs' expertise and problem-solving skills.

What engineers make $200,000 a year?

Senior Site Reliability Engineers and Observability Engineers with extensive experience, advanced skills in cloud platforms, automation, and monitoring tools can earn $200,000 or more annually. High compensation often correlates with working at large tech companies, possessing specialized certifications, and managing complex, scalable systems.

What is the difference between Observability Site Reliability Engineer vs Monitoring Engineer?

AspectObservability Site Reliability EngineerMonitoring Engineer
FocusEnsuring system reliability through observability, automation, and incident responseImplementing and managing monitoring tools and dashboards
SkillsCloud platforms, scripting, incident management, observability toolsMonitoring tools, alerting systems, data analysis
Work EnvironmentDevOps teams, cloud infrastructure, large-scale systemsOperations teams, infrastructure monitoring

While both roles involve system health, the Observability Site Reliability Engineer focuses on comprehensive system reliability using observability practices, whereas Monitoring Engineers primarily manage monitoring tools and alerts. The SRE role emphasizes automation, incident response, and system resilience, making it broader in scope.

What engineers make $300,000 a year?

Senior Site Reliability Engineers and Observability Engineers with extensive experience, advanced skills in cloud platforms, automation, and monitoring tools can earn $300,000 or more annually. High compensation often correlates with working at large tech companies, possessing specialized certifications, and taking on leadership or highly technical roles.
What are popular job titles related to Observability Site Reliability Engineer jobs in Georgia? For Observability Site Reliability Engineer jobs in Georgia, the most frequently searched job titles are:
What job categories do people searching Observability Site Reliability Engineer jobs in Georgia look for? The top searched job categories for Observability Site Reliability Engineer jobs in Georgia are:
What cities in Georgia are hiring for Observability Site Reliability Engineer jobs? Cities in Georgia with the most Observability Site Reliability Engineer job openings:

Site Reliability Engineer (SRE)

Florence Healthcare - US

Atlanta, GA โ€ข On-site

$54.75 - $72.75/hr

Other

Re-posted yesterday


Job description

What You'll Bring to the Team:

We are seeking a Site Reliability Engineer (SRE) to join one of our Scrum teams and help ensure the reliability, scalability, and performance of the Florence platform. AI-driven tooling and automation are a cornerstone of how we build, operate, and scale our systems.

In this role, you will work closely with product engineers while actively leveraging AI to improve observability, incident response, automation, and overall platform reliability. Coding assignments in this role will require working with AI-assisted development workflows as a core part of how solutions are designed and delivered.

You Will:
  • Be an embedded member of a Scrum team, participating in planning, refinement, reviews, and retrospectives
  • Use AI-powered tools to enhance system reliability, operational efficiency, and developer productivity
  • Design, build, and operate reliable, scalable cloud infrastructure supporting platform and product services
  • Apply AI-assisted analysis to monitoring, alerting, and observability data to detect, predict, and prevent incidents
  • Define and maintain SLOs, SLIs, and error budgets to guide reliability decisions
    Collaborate with software engineers to embed reliability and AI-driven automation into the software development lifecycle
  • Lead and participate in incident response, root cause analysis, and postmortems, leveraging AI insights where appropriate
  • Automate operational tasks and reduce toil through AI-enabled and traditional automation approaches
  • Contribute to disaster recovery planning, testing, and operational readiness
  • Produce and maintain documentation such as runbooks, operational guides, and system diagrams
  • Contribute code as a secondary responsibility, with coding assignments focused on building reliability tooling, automation, and integrations using AI-assisted development practices
An Ideal Candidate Is / Has:
  • Passionate about building reliable, scalable systems using modern, AI-enabled approaches
  • Strong understanding of cloud-native and distributed system architectures
  • Experience applying SRE principles in a production environment
  • Hands-on experience with cloud platforms (AWS preferred)
  • Experience using AI-assisted tools for coding, debugging, automation, or operational analysis
  • Strong background in Linux, networking, and system operations
  • Experience with infrastructure-as-code and automation tools (e.g., Terraform, CI/CD pipelines)
  • Familiarity with modern observability practices (metrics, logs, tracing), including AI-enhanced analysis
  • Comfortable working as part of an agile, cross-functional Scrum team
  • Strong problem-solving, communication, and collaboration skills
  • 4+ years of experience in SRE, DevOps, or similar roles
  • Experience supporting production systems at scale