1

Cloud Site Reliability Engineer Jobs (NOW HIRING)

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

This role will help establish and mature SRE practices across AWS cloud environments, with a focus on reliability, automation, scalability, observability, incident response, and operational ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

This role will help establish and mature SRE practices across AWS cloud environments, with a focus on reliability, automation, scalability, observability, incident response, and operational ...

Site Reliability Engineer

Waltham, MA ยท On-site

$61.50 - $81.75/hr

This is not a product development role, and it isn't a traditional cloud-SRE role either. You are the frontline for keeping fielded imaging systems alive: the person field personnel and customer ...

$53.25 - $70.75/hr

Overview We are seeking an experienced Site Reliability Engineer (SRE) - Microsoft Hyper-V ... Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V. * Optimize, and ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

This role will help establish and mature SRE practices across AWS cloud environments, with a focus on reliability, automation, scalability, observability, incident response, and operational ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

This role will help establish and mature SRE practices across AWS cloud environments, with a focus on reliability, automation, scalability, observability, incident response, and operational ...

Showing results 41-60

Cloud Site Reliability Engineer information

See salary details

$10

$63

$91

How much do cloud site reliability engineer jobs pay per hour?

As of Sep 9, 2026, the average hourly pay for cloud site reliability engineer in the United States is $63.74, according to ZipRecruiter salary data. Most workers in this role earn between $54.81 and $72.84 per hour, depending on experience, location, and employer.

What is a cloud site reliability engineer?

A Cloud Site Reliability Engineer (Cloud SRE) is responsible for ensuring the reliability, availability, and performance of cloud-based systems and services. They apply software engineering principles to infrastructure and operations tasks, automating processes like monitoring, incident response, and scaling. Cloud SREs work closely with development and operations teams to build resilient cloud environments, implement observability tools, and manage system reliability. Their goal is to minimize downtime, optimize performance, and enhance the overall efficiency of cloud services.

What does a cloud site reliability engineer do?

Cloud Site Reliability Engineers (SREs) are responsible for maintaining, monitoring, and improving the reliability and performance of cloud-based services. This involves automating repetitive tasks, designing robust monitoring and alerting systems, managing incidents, and working proactively to prevent outages. SREs often collaborate closely with software development, infrastructure, and security teams to deploy new features and ensure scalability. Additionally, they participate in on-call rotations and root cause analyses to address and learn from incidents. These varied tasks help ensure the stability and efficiency of cloud environments while allowing SREs to develop a broad skill set and contribute to continuous service improvement.

What are the key skills and qualifications needed to thrive as a cloud site reliability engineer?

To thrive as a Cloud Site Reliability Engineer, you need a solid background in cloud infrastructure, automation, monitoring, and reliability engineering, supported by a degree in computer science or related field. Proficiency with cloud platforms like AWS, Azure, or Google Cloud, along with expertise in scripting languages, container orchestration tools (such as Kubernetes), and relevant certifications (e.g., AWS Certified DevOps Engineer) is highly valued. Strong analytical thinking, clear communication, and effective problem-solving skills help SREs excel in dynamic, cross-functional teams. These capabilities ensure the design, operation, and continuous improvement of resilient and scalable cloud systems, minimizing downtime and enhancing user experience.

More about Cloud Site Reliability Engineer jobs

What are popular job titles related to Cloud Site Reliability Engineer jobs?

For Cloud Site Reliability Engineer jobs, the most frequently searched job titles are:

Infographic showing various Cloud Site Reliability Engineer job openings in the United States as of September 2026, with employment types broken down into 1% As Needed, 84% Full Time, 12% Part Time, 2% Contract, and 1% Nights. Highlights an 94% Physical, 2% Hybrid, and 4% Remote job distribution, with an average salary of $132,583 per year, or $63.7 per hour.

Site Reliability Engineer (SRE)

Mclean, VA โ€ข Remote

System One
Business Consulting Servicesย โ€ขย 5 - 10K employees

$58.25 - $77.50/hr

Contractor

Medical, Dental, Vision, Life, Retirement

Re-posted 18 days ago


Job description

Site Reliability Engineer (SRE)

Remote No sponsorship available. Must be able to obtain a Public Trust clearance.

What You Will Do

We are seeking a Site Reliability Engineer (SRE) to support the SBA Disaster Lending Platform modernization effort in a remote capacity. This role will help establish and mature SRE practices across AWS cloud environments, with a focus on reliability, automation, scalability, observability, incident response, and operational excellence.

In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to improve system resilience, reduce downtime, strengthen deployment practices, and support reliable cloud-based application delivery in an Agile environment.

Responsibilities include:

• Help establish and mature SRE practices within an Agile Scrum delivery environment. • Support system design reviews to identify reliability risks, failure points, scalability concerns, and opportunities for automation. • Improve operational readiness by contributing to code reviews, deployment reviews, monitoring practices, and reliability-focused engineering standards. • Support incident management activities, including troubleshooting, root-cause analysis, mitigation planning, and post-incident improvements. • Build and maintain automation to improve reliability, reduce manual effort, and support self-healing cloud infrastructure. • Support AWS cloud platform operations across monitoring, logging, security, scalability, and availability. • Work with CI/CD and Infrastructure as Code tools to support repeatable, secure, and reliable deployments. • Create and maintain clear technical documentation for systems, processes, runbooks, and operational procedures. • Collaborate with cross-functional teams and stakeholders to promote DevOps, automation, and reliability best practices.

What You Will Need

• Minimum of four years of experience supporting the reliability, scalability, security, and operational excellence of AWS cloud platforms. • Bachelor’s degree required, or four additional years of relevant experience in lieu of a degree. • Hands-on experience with CI/CD and Infrastructure as Code tools such as Terraform, Ansible Automation Platform, GitLab, Artifactory, and Packer. • Strong scripting and automation experience using Python, PowerShell, and Bash; Python experience is preferred. • Experience supporting Windows and Linux environments. • Strong understanding of networking concepts, cloud troubleshooting, monitoring, logging, and incident response. • Experience designing, deploying, or supporting cloud-based systems with a focus on reliability, scalability, security, and performance. • Knowledge of source control best practices. • Experience working in Agile delivery environments, including Scrum, Kanban, SAFe, or similar methodologies. • Strong analytical, troubleshooting, and problem-solving skills, including the ability to resolve complex technical issues in high-pressure situations. • Strong communication skills and the ability to collaborate effectively with technical teams, stakeholders, and cross-functional partners. • Must be authorized to work in the United States without sponsorship and able to obtain a Public Trust clearance.

Nice to Have

• Current or prior government contracting experience. • Red Hat, CompTIA, AWS, or related technical certifications. • Experience mentoring technical teams or helping promote DevOps/SRE practices across engineering groups.

System One, and its subsidiaries including Joulé and Mountain Ltd., are leaders in delivering outsourced services and workforce solutions across North America. We help clients get work done more efficiently and economically, without compromising quality. System One not only serves as a valued partner for our clients, but we offer eligible employees health and welfare benefits coverage options including medical, dental, vision, spending accounts, life insurance, voluntary plans, as well as participation in a 401(k) plan.

System One is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity, age, national origin, disability, family care or medical leave status, genetic information, veteran status, marital status, or any other characteristic protected by applicable federal, state, or local law.

#M1 #LI-CS1 Ref: #851-Rockville-S1


System One logo

About System One

Sourced by ZipRecruiter

System One helps employers get work done more efficiently and economically without compromising quality. Over our 35+ year history, we've helped connect thousands of talented people with innovative companies. The excitement of a perfect fit motivates us every single day.

Industry

Business consulting services and recruiting and staffing services

Company size

5,001 - 10,000 Employees

Headquarters location

Pittsburgh, PA, US