1

Principal Site Reliability Engineer Jobs (NOW HIRING)

Principal Site Reliability Engineer

Atlanta, GA · Hybrid

$54.25 - $72/hr

PowerPlan is seeking a Principal Site Reliability Engineer to sit at the heart of our cloud platform's reliability, scalability, and operational maturity. You'll work hands-on across AWS and Azure ...

$200 - $250/hr

The Principal Site Reliability Engineer partners with development teams by designing availability and resiliency patterns in applications and infrastructure.**Essential Functions:*** Design and ...

Symmetrio is recruiting a Principal Site Reliability Engineer (SRE) for our customer, a rapidly growing healthcare technology organization focused on advanced healthcare technology solutions. This ...

Principal Site Reliability Engineer

Atlanta, GA · Hybrid

$54.25 - $72/hr

Overview PowerPlan is seeking a Principal Site Reliability Engineer to sit at the heart of our cloud platform's reliability, scalability, and operational maturity. You'll work hands-on across AWS and ...

Digital - Principal SRE

Columbus, OH · On-site

$125 - $150/hr

The Digital - Principal SRE (AI Engineer) role is a position that blends expertise in artificial intelligence, machine learning, and reliability engineering. This professional is responsible for ...

Principal Site Reliability Engineer

Atlanta, GA · On-site

$54.25 - $72/hr

Overview PowerPlan is seeking a Principal Site Reliability Engineer to sit at the heart of our cloud platform's reliability, scalability, and operational maturity. You'll work hands-on across AWS and ...

Digital - Principal SRE

Columbus, OH · On-site +1

$53.50 - $71.25/hr

Description The Digital - Principal SRE (AI Engineer) role is a position that blends expertise in artificial intelligence, machine learning, and reliability engineering. This professional is ...

Digital - Principal SRE

Columbus, OH · On-site +1

$55 - $73.25/hr

Description The Digital - Principal SRE (AI Engineer) role is a position that blends expertise in artificial intelligence, machine learning, and reliability engineering. This professional is ...

Digital - Principal SRE

Columbus, OH · On-site +1

$53.50 - $71.25/hr

Description The Digital - Principal SRE (AI Engineer) role is a position that blends expertise in artificial intelligence, machine learning, and reliability engineering. This professional is ...

Digital - Principal SRE

Columbus, OH · On-site +1

$53.50 - $71.25/hr

Description The Digital - Principal SRE (AI Engineer) role is a position that blends expertise in artificial intelligence, machine learning, and reliability engineering. This professional is ...

Principal Site Reliability Engineer

Nashville, TN · On-site

$55 - $73.25/hr

The Principal Site Reliability Engineer will work across Windows, Linux, Oracle Cloud Infrastructure, hybrid cloud, and legacy environments while partnering with engineering, operations ...

Principal SRE / Hybrid / Tempe

Tempe, AZ · Hybrid

$55.50 - $73.75/hr

This organization is hiring a Principal Site Reliability Engineer for a full-time hybrid role based in Tempe, Arizona (3 days onsite). The company builds and operates large-scale consumer technology ...

Showing results 21-40

Principal Site Reliability Engineer information

See salary details

$10

$63

$91

How much do principal site reliability engineer jobs pay per hour?

As of Sep 9, 2026, the average hourly pay for principal site reliability engineer in the United States is $63.74, according to ZipRecruiter salary data. Most workers in this role earn between $54.81 and $72.84 per hour, depending on experience, location, and employer.

What is a principal site reliability engineer?

Principal Site Reliability Engineers (SREs) are senior technical experts who lead the design, implementation, and maintenance of reliable, scalable, and highly available systems. They oversee complex infrastructure and work closely with engineering teams to optimize system performance, automate processes, and ensure operational excellence. Principal SREs also mentor other engineers, set technical standards, and drive improvements in incident response, monitoring, and system resilience. Their work is critical in minimizing downtime and ensuring a seamless experience for users.

How does a principal site reliability engineer contribute to setting technical direction and mentoring within an SRE team?

As a Principal Site Reliability Engineer, you play a critical role in shaping the technical vision of the SRE team by establishing best practices for infrastructure reliability, scalability, and incident response. You are often expected to mentor junior and mid-level engineers, guiding them through complex troubleshooting, architectural decisions, and automation strategies. Additionally, you collaborate closely with software engineering, product, and operations teams to ensure that reliability and performance goals align with business needs. This role offers significant influence over technical roadmaps and provides opportunities to lead cross-functional initiatives, making it ideal for those seeking both leadership and hands-on impact.

What are the key skills and qualifications needed to thrive as a principal site reliability engineer, and why are they important?

To thrive as a Principal Site Reliability Engineer, you need deep expertise in systems engineering, cloud infrastructure, automation, and strong programming skills, typically supported by a degree in computer science or a related field. Familiarity with tools like Kubernetes, Terraform, Prometheus, and CI/CD platforms, as well as certifications such as AWS Certified Solutions Architect or Google Professional Cloud DevOps Engineer, are often required. Exceptional problem-solving, leadership, and communication skills help you guide teams and drive reliability initiatives across organizations. These skills ensure reliable, scalable systems and foster a culture of continuous improvement and operational excellence.

What is the difference between Principal Site Reliability Engineer vs Site Reliability Engineer?

AspectPrincipal Site Reliability EngineerSite Reliability Engineer
CredentialsAdvanced certifications (e.g., AWS, Google Cloud), extensive experienceEntry to mid-level certifications, relevant experience
Work EnvironmentStrategic planning, architecture design, mentoringOperational tasks, automation, monitoring
Employer UsageLarge tech companies, cloud providers, enterprisesTech firms, startups, cloud services

The Principal Site Reliability Engineer typically holds more advanced certifications and has a strategic, leadership role in designing systems and mentoring teams. In contrast, the Site Reliability Engineer focuses on operational tasks, automation, and maintaining system reliability. Both roles are vital in ensuring system stability but differ in scope and seniority.

More about Principal Site Reliability Engineer jobs

What cities are hiring for Principal Site Reliability Engineer jobs?

Cities with the most Principal Site Reliability Engineer job openings:

What are popular job titles related to Principal Site Reliability Engineer jobs?

For Principal Site Reliability Engineer jobs, the most frequently searched job titles are:

Infographic showing various Principal Site Reliability Engineer job openings in the United States as of September 2026, with employment types broken down into 1% As Needed, 84% Full Time, 12% Part Time, 2% Contract, and 1% Nights. Highlights an 94% Physical, 2% Hybrid, and 4% Remote job distribution, with an average salary of $132,583 per year, or $63.7 per hour.

Principal Site Reliability Engineer

Atlanta, GA • Hybrid

PowerPlan, Inc
Software Development • 201 - 500 employees

$54.25 - $72/hr

Full-time

This job post has expired today. Applications are no longer accepted.


Job description

PowerPlan is seeking a Principal Site Reliability Engineer to sit at the heart of our cloud platform's reliability, scalability, and operational maturity. You'll work hands-on across AWS and Azure environments, solving complex production problems while systematically eliminating the manual toil that creates them.

This role offers significant autonomy, deep technical impact, and the opportunity to shape how reliability engineering is practiced across the organization.

About PowerPlan

PowerPlan helps the companies that power the world unlock greater value from their infrastructure investments. We combine deep industry expertise, trusted technology, and AI-driven innovation to help capital-intensive organizations manage the financial complexity of their assets with confidence.

We're in the middle of one of the most significant transformations in our history. Decades of proven functionality are being reimagined on a modern SaaS foundation, while AI becomes central to how we build, deliver, and evolve our products. The next generation of the PowerPlan platform is being designed right now. Join us, and you'll have the opportunity to influence the technology, engineering practices, and capabilities that will shape it for years to come.


Your Impact
  • Resolve escalated infrastructure cases across AWS and Azure and deliver targeted automations that reduce manual resolution time.
  • Eliminate or significantly reduce manual intervention for the highest-frequency operational issues through automation and tooling.
  • Establish a consistent, high-quality incident response and post-incident review process for critical production incidents.
  • Deliver a mature, SLO-aligned observability platform with dashboards, tuned alerts, and clear reporting.
  • Coach teams on effective incident communication and decision-making.
What Success Looks Like

Within 90 days, you'll have shipped automations that measurably cut manual resolution time on recurring issues. By month six, you'll have eliminated toil on the highest-frequency operational problems. By month twelve, on-call and engineering teams will be running on an observability platform you built — one with dashboards, tuned alerts, and SLIs/SLOs that make reliability a data-driven practice instead of a guessing game.


What You'll Bring
  • Deep hands-on experience operating production systems in AWS and Azure environments.
  • Strong automation skills using Python and PowerShell in operational contexts.
  • Proven ability to identify repetitive operational work and eliminate it through automation.
  • Experience leading incident response and blameless post-incident reviews.
  • Strong observability expertise, particularly with Grafana and SLI/SLO-driven monitoring.
  • Ability to influence engineering practices without formal authority.
  • Clear written and verbal communication skills across technical and non-technical audiences.
Education & Experience
  • Extensive experience in cloud operations, site reliability engineering, or infrastructure engineering roles, or equivalent professional experience.

PowerPlan is an EOE

Applicant and Candidate Privacy Notice

Please note that this is a hybrid role that involves a combination of onsite work from our corporate office as well as work from home. While we strive to accommodate flexible working arrangements when sensible, there will be times when onsite work is required. This could include scheduled office days, team meetings, client meetings, or special events.