2

Site Reliability Engineer Remote Jobs in Decatur, GA

Atlanta / Remote What You'll Do * Design, build, and maintain infrastructure-as-code (Terraform ... SRE, or cloud infrastructure engineering roles * Strong hands-on experience with a major cloud ...

Senior AI Engineer (Remote)

Atlanta, GA ยท On-site +1

$99K - $136K/yr

The Senior AI Engineer is responsible for designing, building, scaling, and optimizing production ... AIOps & Deployment Reliability: Experience building automated CI/CD pipelines for AI/agentic ...

Senior Transmission Line Engineer - REMOTE

Atlanta, GA ยท Remote

$100K - $138K/yr

Title: Senior Transmission Line Engineer Location: Remote US Ready to make a difference? We are ... and reliability > ICF Power Delivery Services #GEA25 #POWERDELIVERY #INDEED #LI-CC1 Bot and Third ...

Staff Software Engineer, Freestyle

Atlanta, GA ยท On-site +1

$164K - $274K/yr

Work closely with SRE, Partner Engineering, and Product teams to resolve dependencies and deliver robust solutions. What will you bring to Omnissa? * 8+ years of software engineering experience, with ...

Reporting to the Engineering Manager, South East Region. This is a remote role but would like the ... Responding to information requests from site and completing project documentation including as ...

Senior Cloud Security Engineer

Atlanta, GA ยท On-site +1

$110K - $151K/yr

... remote employees worldwide-we are committed to building a diverse and inclusive workplace. We ... Familiarity with policy management tools such as OPA or Kyverno Bonus points: * SRE Experience

Staff Software Engineer, Freestyle

Atlanta, GA ยท On-site +1

$164K - $274K/yr

Work closely with SRE, Partner Engineering, and Product teams to resolve dependencies and deliver robust solutions. What will you bring to Omnissa? * 8+ years of software engineering experience, with ...

Senior Data Engineer/Devops

Atlanta, GA ยท Remote

$100K - $136K/yr

Senior Healthcare Data Engineer Remote TRC Talent Solutions is seeking a Senior Healthcare Data ... Ensure data quality, security, reliability, and compliance within healthcare data environments ...

Showing results 41-60

Site Reliability Engineer Remote information

See Decatur, GA salary details

$10

$62

$89

How much do site reliability engineer remote jobs pay per hour?

As of Aug 22, 2026, the average hourly pay for site reliability engineer remote in Decatur, GA is $62.23, according to ZipRecruiter salary data. Most workers in this role earn between $53.51 and $71.11 per hour, depending on experience, location, and employer.

What is a site reliability engineer remote?

A Site Reliability Engineer (SRE) in a remote role is responsible for ensuring the reliability, performance, and scalability of software systems while working from a remote location. They bridge the gap between development and operations by implementing automation, monitoring, and incident response strategies. Remote SREs collaborate with distributed teams to improve infrastructure, troubleshoot issues, and optimize system performance. Strong communication skills, proficiency in cloud technologies, and expertise in software development are essential for success in this role.

What are the key skills and qualifications needed to thrive as a site reliability engineer remote?

To thrive as a Site Reliability Engineer Remote, you need expertise in systems administration, cloud infrastructure, automation, coding (often in Python or Go), and a solid grasp of networking fundamentals, usually demonstrated with a degree in computer science or equivalent experience. Familiarity with tools such as Docker, Kubernetes, AWS/GCP/Azure, monitoring platforms like Prometheus, and certifications like AWS Certified SysOps Administrator are highly valued. Excellent problem-solving, communication, and collaboration skills are essential, especially when troubleshooting incidents and passing information across distributed teams. These abilities ensure reliable, scalable services and smooth coordination in a remote work environment.

What are some common challenges faced by site reliability engineers working remotely, and how are they addressed?

Site Reliability Engineers working remotely may encounter challenges like coordinating across multiple time zones, maintaining clear communication during urgent incidents, and managing complex systems without direct on-site access. These are often addressed by leveraging collaborative tools (like Slack, Zoom, and incident management platforms), implementing well-documented processes, and participating in regular team syncs or on-call rotations. Remote SREs also benefit from automation and observability practices that provide in-depth systems insights without needing physical presence. Many organizations support their success through robust onboarding, continuous training, and establishing clear lines of communication for rapid response scenarios. This blend of technical and teamwork strategies helps remote SREs maintain service reliability and stay connected with their colleagues.

What are popular job titles related to Site Reliability Engineer Remote jobs in Decatur, GA?

For Site Reliability Engineer Remote jobs in Decatur, GA, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer Remote jobs in Decatur, GA look for?

The top searched job categories for Site Reliability Engineer Remote jobs in Decatur, GA are:

What cities near Decatur, GA are hiring for Site Reliability Engineer Remote jobs?

Cities near Decatur, GA with the most Site Reliability Engineer Remote job openings:

Infographic showing various Site Reliability Engineer Remote job openings in Decatur, GA as of August 2026, with employment types broken down into 1% As Needed, 81% Full Time, 16% Part Time, and 2% Contract. Highlights an 93% Physical, 3% Hybrid, and 4% Remote job distribution, with an average salary of $129,445 per year, or $62.2 per hour.

Senior Software Engineer [REMOTE]

Upbound - Job Posting

Atlanta, GA โ€ข On-site, Remote

$117K - $155K/yr

Full-time

Re-posted 23 days ago


Job description

Upbound is redefining how modern infrastructure is built for the Agentic AI Era. We're the creators and primary maintainers of Crossplane, and we're building the Intelligent Control Plane-a new platform layer that makes infrastructure programmable, autonomous, and composable.
Our mission is to power the AI-native enterprise with a foundational platform layer that helps teams provision, operate, and adapt infrastructure at scale-so platforms are ready for both humans and AI agents. We partner with leading cloud providers, ISVs, and open-source communities to help organizations move faster with greater confidence.
Today, Upbound supports Fortune 500 companies and platform engineers across 100+ countries. Crossplane has surpassed 100M+ downloads and is used by 1,000+ teams worldwide. We're a Series B company backed by GV (formerly Google Ventures), Altimeter Capital, and Intel Capital, and we've raised $69M to date. Learn more at upbound.io.
Upbound is hiring a Senior Software Engineer to help us build and operate Upbound Spaces, the multiple control plane management software at the heart of the Upbound Platform. As part of the Spaces team, you will help us scale Upbound to reliably support thousands of control planes, while also extending enterprise control plane management and operations both in the cloud and on premises. Our team is expanding, and this is the perfect opportunity for you to make a significant engineering impact in both development and production operations.
What You'll Do
  • Actively build and operate Upbound Spaces in production, troubleshooting and resolving issues across multi-tenant SaaS environments, as well as contributing to Upbound's open-source projects, including Crossplane.
  • Take ownership of building features in high demand by Upbound's customers and deliver new functionality that will delight and amaze our users.
  • Investigate and debug complex issues in customer environments, including multi-control plane scenarios, resource reconciliation problems, and performance bottlenecks.
  • Communicate through thoughtful and thorough design documents for new initiatives and detailed post-incident reviews that drive system improvements.
  • Support the full project lifecycle for highly scalable and reliable services running in a cloud environment - discovery, analysis, architecture, design, review, documentation, building, migration, automation, deployment, production-readiness, and ongoing operational support.
  • Write and maintain Go code that interfaces with the Kubernetes API, such as operators, controllers, add-ons, etc., with a focus on observability, debuggability, and operational excellence.
  • Deploy, manage, and troubleshoot our Kubernetes services in production, using metrics, logs, and traces to identify and resolve issues quickly.
  • Build and maintain operational tooling for debugging customer environments, analyzing control plane health, and automating incident response.
  • Author documentation, user guides, runbooks, and blog posts to support and promote new features that you release.
  • Support the software release cycle for Spaces self-hosted distributions, including diagnosing issues in customer-managed deployments.
  • Participate in on-call rotation to support Upbound Cloud, responding to incidents and driving them to resolution.
What You'll Bring
  • Have experience operating production cloud services at scale: monitoring, alerting, incident response, post-mortems, and continuous improvement of service reliability.
  • Have strong debugging skills across distributed systems, including experience with observability tools (Prometheus, Grafana, OpenTelemetry, distributed tracing) and techniques for diagnosing issues in production environments.
  • Have experience building and operating controllers that interact with the Kubernetes API server, including troubleshooting reconciliation loops, managing API rate limits, and optimizing controller performance.
  • Are comfortable working directly with customers to understand, reproduce, and resolve complex technical issues in their environments.
  • Take responsibility and ownership for solving problems even if they are outside your lane, especially during incidents affecting customer workloads.
  • Demonstrate excellence in your work, constantly trying to improve your skills and the operational posture of the systems you build.
  • Have empathy for customers and keep them in mind as you build solutions, understanding that reliability and debuggability are features.
  • Realize the importance of clear communication and effective collaboration to work as a team, deliver great results, and support customers through technical challenges.
  • Help create a safe environment where everyone can contribute, learn from failures, share on-call knowledge, and help each other grow as operators and engineers.

#LI-REMOTE
Why Upbound?
At Upbound, you'll help shape the systems and strategies that drive predictable, scalable growth in a product-led company embracing usage-based models. If you're excited to build from the ground up, work with cutting-edge cloud technologies, and directly impact how revenue is generated and scaled-this is your seat at the table.
About Upbound
Upbound is pioneering infrastructure platforms for the Agentic AI Era, serving Fortune 500 companies and platform engineers across more than 100 countries. The company empowers infrastructure and platform teams with Intelligent Control Planes - based on Kubernetes and Crossplane - that provision, operate, and adapt so platforms are ready for both humans and AI agents. Upbound is the creator and primary maintainer of Crossplane, the popular open-source framework for building cloud-native control planes, with over 100 million downloads and adoption by more than 1,000 teams worldwide. A Series B startup backed by GV (formerly Google Ventures), Altimeter Capital, and Intel Capital, Upbound has raised $69M to date. For more information, visit www.upbound.io.