1

Observability Site Reliability Engineer Jobs in Washington, DC

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

This role will help establish and mature SRE practices across AWS cloud environments, with a focus on reliability, automation, scalability, observability, incident response, and operational ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

This role will help establish and mature SRE practices across AWS cloud environments, with a focus on reliability, automation, scalability, observability, incident response, and operational ...

Site Reliability Engineer (SRE)

Mclean, VA ยท Remote

$58.25 - $77.50/hr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

This role will help establish and mature SRE practices across AWS cloud environments, with a focus on reliability, automation, scalability, observability, incident response, and operational ...

SRE Engineer

Washington, DC ยท On-site

$64.50 - $85.75/hr

Observability & Monitoring: Standardize and automate Dynatrace installations, integrate telemetry ... Champion SRE metrics including Service Level Indicators (SLIs), Service Level Objectives (SLOs ...

Site Reliability Engineer IV

Sterling, VA ยท On-site

$120 - $160/hr

As a member of the SRE team, you will focus on building and operating highly reliable application platforms by applying SRE principles such as automation, observability, resilience and continuous ...

Site Reliability Engineer - Hybrid

Reston, VA ยท On-site

$59.25 - $78.75/hr

Experience with Observability using tools such as AWS CloudWatch, Splunk/SignalFX, Dynatrace, and ... The SRE at Fannie Mae doesn't work 24*7. They get scheduled on a rotation basis. 20% of their job ...

Site Reliability Engineer, Lead

Chantilly, VA ยท On-site

$99K - $225K/yr

  • Medical

  • Life

  • Retirement

  • PTO

This role leads the design and implementation of observability, automation, incident response, and ... The Lead SRE also drives root cause analysis, capacity planning, reliability standards, and ...

Site Reliability Engineer

Washington, DC ยท On-site

$112K - $179K/yr

The SRE will drive automation initiatives, observability improvements, and incident response operations. Site Reliability Engineer responsibilities: * Design and implement automation solutions for ...

Site Reliability Engineer

Washington, DC ยท On-site

$112K - $179K/yr

The SRE will drive automation initiatives, observability improvements, and incident response operations. Site Reliability Engineer responsibilities: * Design and implement automation solutions for ...

Site Reliability Engineer

Washington, DC ยท On-site

$112K - $179K/yr

The SRE will drive automation initiatives, observability improvements, and incident response operations. Site Reliability Engineer responsibilities: * Design and implement automation solutions for ...

The SRE will work closely with software developers and operations teams to improve system ... Experience working with observability and incident management tools (Datadog, OpsGenie, PagerDuty)

Site Reliability Engineering (SRE) Director

Washington, DC ยท On-site

$158K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

This role drives automation, observability, incident response maturity, and continuous improvement while building a SRE function that supports both enterprise continuity and innovation.

Showing results 21-40

Observability Site Reliability Engineer information

See Washington, DC salary details

$12

$72

$104

How much do observability site reliability engineer jobs pay per hour?

As of Aug 17, 2026, the average hourly pay for observability site reliability engineer in Washington, DC is $72.19, according to ZipRecruiter salary data. Most workers in this role earn between $62.07 and $82.50 per hour, depending on experience, location, and employer.

What is the difference between Observability Site Reliability Engineer vs Monitoring Engineer?

AspectObservability Site Reliability EngineerMonitoring Engineer
FocusEnsuring system reliability through observability, automation, and incident responseImplementing and managing monitoring tools and dashboards
SkillsCloud platforms, scripting, incident management, observability toolsMonitoring tools, alerting systems, data analysis
Work EnvironmentDevOps teams, cloud infrastructure, large-scale systemsOperations teams, infrastructure monitoring

While both roles involve system health, the Observability Site Reliability Engineer focuses on comprehensive system reliability using observability practices, whereas Monitoring Engineers primarily manage monitoring tools and alerts. The SRE role emphasizes automation, incident response, and system resilience, making it broader in scope.

What job categories do people searching Observability Site Reliability Engineer jobs in Washington, DC look for?

The top searched job categories for Observability Site Reliability Engineer jobs in Washington, DC are:

Site Reliability Engineer (SRE)

System One

Mclean, VA โ€ข Remote

$58.25 - $77.50/hr

Contractor

Medical, Dental, Vision, Life, Retirement

Re-posted 25 days ago


Job description

Site Reliability Engineer (SRE)

Remote No sponsorship available. Must be able to obtain a Public Trust clearance.

What You Will Do

We are seeking a Site Reliability Engineer (SRE) to support the SBA Disaster Lending Platform modernization effort in a remote capacity. This role will help establish and mature SRE practices across AWS cloud environments, with a focus on reliability, automation, scalability, observability, incident response, and operational excellence.

In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to improve system resilience, reduce downtime, strengthen deployment practices, and support reliable cloud-based application delivery in an Agile environment.

Responsibilities include:

• Help establish and mature SRE practices within an Agile Scrum delivery environment. • Support system design reviews to identify reliability risks, failure points, scalability concerns, and opportunities for automation. • Improve operational readiness by contributing to code reviews, deployment reviews, monitoring practices, and reliability-focused engineering standards. • Support incident management activities, including troubleshooting, root-cause analysis, mitigation planning, and post-incident improvements. • Build and maintain automation to improve reliability, reduce manual effort, and support self-healing cloud infrastructure. • Support AWS cloud platform operations across monitoring, logging, security, scalability, and availability. • Work with CI/CD and Infrastructure as Code tools to support repeatable, secure, and reliable deployments. • Create and maintain clear technical documentation for systems, processes, runbooks, and operational procedures. • Collaborate with cross-functional teams and stakeholders to promote DevOps, automation, and reliability best practices.

What You Will Need

• Minimum of four years of experience supporting the reliability, scalability, security, and operational excellence of AWS cloud platforms. • Bachelor’s degree required, or four additional years of relevant experience in lieu of a degree. • Hands-on experience with CI/CD and Infrastructure as Code tools such as Terraform, Ansible Automation Platform, GitLab, Artifactory, and Packer. • Strong scripting and automation experience using Python, PowerShell, and Bash; Python experience is preferred. • Experience supporting Windows and Linux environments. • Strong understanding of networking concepts, cloud troubleshooting, monitoring, logging, and incident response. • Experience designing, deploying, or supporting cloud-based systems with a focus on reliability, scalability, security, and performance. • Knowledge of source control best practices. • Experience working in Agile delivery environments, including Scrum, Kanban, SAFe, or similar methodologies. • Strong analytical, troubleshooting, and problem-solving skills, including the ability to resolve complex technical issues in high-pressure situations. • Strong communication skills and the ability to collaborate effectively with technical teams, stakeholders, and cross-functional partners. • Must be authorized to work in the United States without sponsorship and able to obtain a Public Trust clearance.

Nice to Have

• Current or prior government contracting experience. • Red Hat, CompTIA, AWS, or related technical certifications. • Experience mentoring technical teams or helping promote DevOps/SRE practices across engineering groups.

System One, and its subsidiaries including Joulé and Mountain Ltd., are leaders in delivering outsourced services and workforce solutions across North America. We help clients get work done more efficiently and economically, without compromising quality. System One not only serves as a valued partner for our clients, but we offer eligible employees health and welfare benefits coverage options including medical, dental, vision, spending accounts, life insurance, voluntary plans, as well as participation in a 401(k) plan.

System One is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity, age, national origin, disability, family care or medical leave status, genetic information, veteran status, marital status, or any other characteristic protected by applicable federal, state, or local law.

#M1 #LI-CS1 Ref: #851-Rockville-S1