2

Overnight Site Reliability Engineer Remote Jobs in Washington

Site Reliability Engineer

Columbia, MD · On-site +1

$55.50 - $73.75/hr

Hybrid Columbia MD 3 times per week OR Remote (as applicable to role) Work Authorization ... Job Overview Cogent People Inc. is seeking a Site Reliability to support system reliability ...

Site Reliability Engineer

Herndon, VA · On-site +1

$87K - $157K/yr

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. You'll work across infrastructure, Linux systems, networking ...

Site Reliability Engineer

Sterling, VA · On-site +1

$87K - $157K/yr

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. You'll work across infrastructure, Linux systems, networking ...

Site Reliability Engineer

Reston, VA · On-site +1

$87K - $157K/yr

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. You'll work across infrastructure, Linux systems, networking ...

Site Reliability Engineer

Burke, VA · On-site +1

$87K - $157K/yr

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. You'll work across infrastructure, Linux systems, networking ...

Site Reliability Engineer

Springfield, VA · On-site +1

$87K - $157K/yr

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. You'll work across infrastructure, Linux systems, networking ...

SRE - Linux

Reston, VA · On-site +1

$164K - $222K/yr

The SRE is part of a highly skilled engineering and infrastructure team responsible for the design, administration, security, and operation of Verisign's application delivery and application security ...

Site Reliability Engineer

Merrifield, VA · On-site +1

$87K - $157K/yr

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. You'll work across infrastructure, Linux systems, networking ...

Site Reliability Engineer

Manassas, VA · On-site +1

$87K - $157K/yr

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. You'll work across infrastructure, Linux systems, networking ...

Site Reliability Engineer

Falls Church, VA · On-site +1

$87K - $157K/yr

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. You'll work across infrastructure, Linux systems, networking ...

Site Reliability Engineer

Chantilly, VA · On-site +1

$87K - $157K/yr

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. You'll work across infrastructure, Linux systems, networking ...

SRE - Linux

Reston, VA · On-site +1

$164K - $222K/yr

The SRE is part of a highly skilled engineering and infrastructure team responsible for the design, administration, security, and operation of Verisign's application delivery and application security ...

Showing results 21-40

Overnight Site Reliability Engineer Remote information

What is an overnight site reliability engineer?

An Overnight Site Reliability Engineer (SRE) is a professional responsible for ensuring the reliability, performance, and uptime of software systems during overnight or off-peak hours, typically working remotely. Their main tasks include monitoring system health, responding to incidents, troubleshooting outages, and implementing fixes to maintain service availability. SREs also work to automate processes, improve system resilience, and collaborate with other engineering teams to prevent future issues. Working overnight ensures that critical systems remain operational and issues are addressed promptly, even outside of standard business hours.

What skills and qualifications are needed to be an overnight site reliability engineer?

To thrive as an Overnight Site Reliability Engineer (Remote), you need strong expertise in systems administration, incident response, automation, and a solid background in computer science or related fields. Proficiency with monitoring tools (like Prometheus or Datadog), cloud platforms (such as AWS or GCP), scripting languages (Python, Bash), and certifications like AWS Certified SysOps Administrator are highly beneficial. Exceptional problem-solving skills, attention to detail, and effective remote communication help you excel in high-pressure overnight scenarios. These skills ensure system reliability, minimize downtime, and maintain seamless operations during critical off-hours.

What are the unique challenges and expectations for an overnight site reliability engineer working remotely?

As an Overnight Site Reliability Engineer working remotely, you'll often handle critical incidents that arise outside of standard business hours, so strong problem-solving skills and the ability to work independently are crucial. Communication is key, as you'll need to coordinate with team members in different time zones and document incidents clearly for seamless handoffs. You may also be tasked with proactive monitoring and maintenance activities during quieter periods, making self-motivation and attention to detail especially important. The role offers valuable exposure to high-impact issues and can accelerate your expertise in incident management and system reliability.

What is the difference between Overnight Site Reliability Engineer Remote vs Overnight DevOps Engineer Remote?

AspectOvernight Site Reliability Engineer RemoteOvernight DevOps Engineer Remote
Primary FocusEnsuring system reliability, uptime, and incident responseAutomating deployment, integration, and infrastructure management
Required SkillsMonitoring, incident management, scripting, system troubleshootingCI/CD pipelines, automation, cloud platforms, scripting
Work EnvironmentRemote, on-call shifts, collaboration with SRE teamsRemote, development and deployment focus, collaboration with development teams
CertificationsLinux, AWS, Google Cloud, or Azure certifications often preferredCloud certifications, Docker, Kubernetes, CI/CD tools

While both roles involve working remotely and require cloud and scripting skills, the Overnight Site Reliability Engineer Remote primarily focuses on maintaining system reliability and incident response, whereas the Overnight DevOps Engineer Remote emphasizes automation, deployment, and infrastructure management. Understanding these differences helps candidates align their skills with the right role.

What are the most commonly searched types of Site Reliability Engineer Remote jobs in Washington?

The most popular types of Site Reliability Engineer Remote jobs in Washington are:

What are popular job titles related to Overnight Site Reliability Engineer Remote jobs in Washington?

For Overnight Site Reliability Engineer Remote jobs in Washington, the most frequently searched job titles are:

What job categories do people searching Overnight Site Reliability Engineer Remote jobs in Washington look for?

The top searched job categories for Overnight Site Reliability Engineer Remote jobs in Washington are:

What cities in Washington are hiring for Overnight Site Reliability Engineer Remote jobs?

Cities in Washington with the most Overnight Site Reliability Engineer Remote job openings:

Site Reliability Engineer, Lead

Booz Allen Hamilton, Inc.

Chantilly, VA • On-site, Remote

$99K - $225K/yr

Full-time, Part-time

Medical, Life, Retirement, PTO

Re-posted 14 days ago


Key responsibilities

  • Design and develop observability, automation, incident response, and operational best practices for production systems and platforms.

  • Work with development and operation teams to evaluate system health, stability, and reliability, and support the deployment of debugging tools.

  • Lead efforts to improve system resilience, implement reliability standards, and drive continuous improvement initiatives.


Booz Allen Hamilton rating

8.9

Company rating: 8.9 out of 10

Based on 49 frontline employees who took The Breakroom Quiz

11th of 72 rated business consultants


Job description


Remote Work:
No
Job Number:
R0245130
Location:
Chantilly,VA,US
Share job via:
Share
Site Reliability Engineer, Lead
The Opportunity:
As a Lead Site Reliability Engineer (SRE) on our team, you'll be responsible for ensuring the reliability, performance, scalability, and security of critical production systems and platforms. This role leads the design and implementation of observability, automation, incident response, and operational best practices across cloud and air-gapped environments, while partnering closely with DevOps, infrastructure, and security teams to improve system resilience and reduce operational risk. The Lead SRE also drives root cause analysis, capacity planning, reliability standards, and continuous improvement initiatives to support highly available, efficient, and scalable services. This is your chance to further your skills in cloud infrastructure and technologies while continuing to grow your SRE experience. You'll build and support a reliable site for the environment in order to meet the development and maintenance requirements of systems and platforms. Work with the development and operation teams to evaluate the health, stability and reliability of systems and platforms. Design and develop technical tools to debug problems that occur in the deployment of applications, within specific platforms and systems.
Join our efforts to strengthen our security posture and safeguard national interests.
Join us. The world can't wait.
You Have:
  • 8+ years of experience with monitoring, logging, and observability platforms, such as Prometheus, Grafana, and ELK stack
  • 8+ years of experience with Linux systems administration and networking fundamentals within AWS
  • Experience with Python scripting and automation
  • Experience with Infrastructure as Code using Terraform and Terragrunt
  • Knowledge of Kubernetes administration, troubleshooting, and operations.
  • TS/SCI clearance with a polygraph
  • Bachelor's degree and 8+ years of experience in Site Reliability Engineering, DevOps Engineering, or Platform Engineering, or 12+ years of experience in Site Reliability Engineering, DevOps Engineering, or Platform Engineering in lieu of a degree
  • Ability to obtain a Security+ CE, SSCP, CCNA-Security, or GSEC Certification within 6 months of start date

Nice If You Have:
  • Experience with deploying and managing OpenTelemetry.
  • Experience with AWS CloudWatch, AWS EKS, and related AWS services
  • Experience managing Kubernetes environments through Rancher
  • Experience implementing SRE practices such as SLOs, SLIs, error budgets, and incident management
  • Experience with Jenkins, Git, Docker, Kubernetes, Nessus, JIRA, and Confluence
  • Knowledge of distributed systems, microservices architectures, and containerized workloads
  • Knowledge of NIST 800-53 and NIST-190
  • Master's degree in a relevant field
  • Security+ CE, SSCP, CCNA-Security, or GSEC Certification

Clearance:
Applicants selected will be subject to a security investigation and may need to meet eligibility requirements for access to classified information; TS/SCI clearance with polygraph is required.
Compensation
At Booz Allen, we celebrate your contributions, provide you with opportunities and choices, and support your total well-being. Our offerings include health, life, disability, financial, and retirement benefits, as well as paid leave, professional development, tuition assistance, work-life programs, and dependent care. Our recognition awards program acknowledges employees for exceptional performance and superior demonstration of our values. Full-time and part-time employees working at least 20 hours a week on a regular basis are eligible to participate in Booz Allen's benefit programs. Individuals that do not meet the threshold are only eligible for select offerings, not inclusive of health benefits. We encourage you to learn more about our total benefits by visiting the Resource page on our Careers site and reviewing Our Employee Benefits page.
Salary at Booz Allen is determined by various factors, including but not limited to location, the individual's particular combination of education, knowledge, skills, competencies, and experience, as well as contract-specific affordability and organizational requirements. The projected compensation range for this position is $99,000.00 to $225,000.00 (annualized USD). The estimate displayed represents the typical salary range for this position and is just one component of Booz Allen's total compensation package for employees. This posting will close within 90 days from the Posting Date.
Identity Statement
As part of the hiring process, we will ask you to complete an identity verification process that leverages advanced biometrics and artificial intelligence to ensure authenticity and protect against identity fraud. You are expected to be on camera during interviews and assessments. We reserve the right to take your picture to verify your identity and prevent fraud.
Candidate AI Usage Policy
AI is a part of our daily work at Booz Allen, and we are committed to the responsible and ethical use of AI tools. However, we want to ensure a fair candidate process based on your own skills and knowledge. As part of this commitment, the use of artificial intelligence (AI) or other tools to assist with responses during interviews (whether in-person or virtual) is prohibited unless permission is explicitly provided.
Work Model
Our people-first culture prioritizes the benefits of collaboration, whether it occurs in person or virtually. To support engagement and effective communication, employees working virtually are generally expected to have their cameras on during meetings.
  • Remote: If this position is listed as remote, there may still be occasions when you are required to work in person at a Booz Allen or customer facility.
  • Hybrid: If this position is listed as hybrid, you will be expected to work from a Booz Allen facility frequently, in alignment with leadership expectations and the needs of the role. You may also be required to work from or visit a customer facility.
  • Onsite: If this position is listed as onsite, work will primarily be performed at a Booz Allen office or customer facility, where employees will collaborate directly with colleagues and customers as required by the role.

Commitment to Non-Discrimination
All qualified applicants will receive consideration for employment without regard to disability, status as a protected veteran or any other status protected by applicable federal, state, local, or international law.
Not ready to apply? Join our Talent Community and sign up for job alerts.

What Booz Allen Hamilton employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Booz Allen Hamilton logo

About Booz Allen Hamilton

Sourced by ZipRecruiter

Booz Allen Hamilton is a leading provider of management and technology consulting services to the US government in defense, intelligence, and civil markets. Headquartered in McLean, Virginia, the firm also serves major corporations, institutions, and not-for-profit organizations. Founded in 1914 by Edwin G. Booz, the company has a long-standing tradition of helping clients achieve success by delivering a wide range of consulting services that include strategic planning, human capital and learning, communication, systems development, and others. The company's mission is to empower people to change the world, and it has a reputation for maintaining the highest standards of integrity and-excellence.

Industry

It services

Company size

10,000+ Employees

Headquarters location

McLean, VA, US

Year founded

1914