1

Site Reliability Engineer Jobs in Springfield, VA

Everforth ECS is seeking a Senior Site Reliability Engineer to work remotely . Everforth ECS is seeking talented professionals to join our successful and growing team in building the next-generation ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

The SRE Engineer will improve the reliability, availability, performance, and operational resilience of mission-critical systems for a federal enterprise program. Responsibilities : • Define ...

Site Reliability Engineer (SRE)

Vienna, VA · On-site

$57.25 - $76/hr

Up to 2 years in duration MUST HAVES: • Minimum of 8 years of experience as a Site Reliability Engineerwith a strong understanding of SRE principles for highly scalable and reliable systems • ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

Spatial Front, Inc. is a recognized workplace seeking a SRE Engineer to enhance their Infrastructure, Production, and Compliance Support team. The role focuses on improving the reliability and ...

The SRE will work closely with software developers and operations teams to improve system reliability, automate processes, and minimize downtime. Responsibilities * Design, implement, and maintain ...

SRE Engineer

Washington, DC · On-site

$64.50 - $85.75/hr

Job Title: SRE Engineer Location: Washington, DC Duration: 12+ Months Rate: DOE Reliability Engineer (SRE) to champion system availability, performance, and automation across their enterprise cloud ...

Position: Site Reliability Engineer IV Experience: 9-12 Years Location: Bangalore (Ecospace) Candescent Site Reliability Engineering (SRE) mission is to proactively ensure the reliability ...

Site Reliability Engineer - Hybrid

Reston, VA · On-site

$59.25 - $78.75/hr

Title: Site Reliability Engineer V Location: Reston, VA (Hybrid onsite - 3 days a week from day 1) Assignment duration: 24 months with possibility of extension Interview process: 2 rounds. First ...

Showing results 21-40

Site Reliability Engineer information

See Springfield, VA salary details

$11

$66

$95

How much do site reliability engineer jobs pay per hour?

As of Aug 18, 2026, the average hourly pay for site reliability engineer in Springfield, VA is $66.58, according to ZipRecruiter salary data. Most workers in this role earn between $57.26 and $76.06 per hour, depending on experience, location, and employer.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often requires strong problem-solving skills, familiarity with monitoring tools, and the ability to work in high-pressure situations, but it also offers opportunities for skill development and process improvements.

What are the most commonly searched types of Site Reliability Engineer jobs in Springfield, VA?

The most popular types of Site Reliability Engineer jobs in Springfield, VA are:

What are popular job titles related to Site Reliability Engineer jobs in Springfield, VA?

For Site Reliability Engineer jobs in Springfield, VA, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer jobs in Springfield, VA look for?

The top searched job categories for Site Reliability Engineer jobs in Springfield, VA are:

What cities near Springfield, VA are hiring for Site Reliability Engineer jobs?

Cities near Springfield, VA with the most Site Reliability Engineer job openings:

Infographic showing various Site Reliability Engineer job openings in Springfield, VA as of August 2026, with employment types broken down into 1% As Needed, 80% Full Time, 14% Part Time, 1% Temporary, and 4% Contract. Highlights an 92% Physical, 2% Hybrid, and 6% Remote job distribution, with an average salary of $138,486 per year, or $66.6 per hour.

Site Reliability Engineer (SRE) (TS)

Koniag Government Services

Washington, DC • On-site

$64.25 - $85.50/hr

Full-time

Re-posted 29 days ago


Koniag Government Services rating

8.6

Company rating: 8.6 out of 10

Based on 6 frontline employees who took The Breakroom Quiz

57th of 492 rated business services


Job description

Job Summary:
Koniag Government Services (KGS) is hiring a Site Reliability Engineer (SRE) to ensure the reliability, availability, and performance of mission-critical applications. The role focuses on automation, observability, and incident response while maintaining strict Service Level Objectives (SLOs) and requires an active Top Secret/SCI clearance.
Responsibilities:
• Ensure application reliability, performance, and availability through automation, monitoring, and systems engineering.
• Develop infrastructure-as-code (IaC) solutions using Terraform, Ansible, and Desired State Configuration (DSC).
• Build and manage containerized workloads using Kubernetes, Rancher, Docker, Helm, and related ecosystem tools.
• Support service mesh and networking constructs such as Cilium, load balancing, ingress management, and distributed storage.
• Engineer and maintain storage and object systems including Rook, Ceph, MinIO, and S3-compatible platforms.
• Implement and maintain comprehensive observability platforms (metrics, logging, tracing) to support SLO monitoring and incident response.
• Lead and participate in incident response activities, postmortem analysis, and reliability engineering improvements.
• Develop automations, scripts, and tools using Python, PowerShell, and shell scripting.
• Support CI/CD pipelines and cloud-native deployment methodologies.
• Collaborate with development and operations teams to embed SRE practices into the application lifecycle.
Qualifications:
Required:
• Active Top Secret/SCI clearance with ability to obtain additional security requirements.
• At least two required technical certifications from the following: Security +, Cloud Associate (such as AWS Solutions Architect Associate, Azure AZ-104, or Google Cloud Associate Cloud Engineer), Terraform Associate, Cloud Professional/Architect (such as AWS Solutions Architect Professional or Azure Architect Expert), CKA (Certified Kubernetes Administrator)
• Strong understanding of the following technologies: Kubernetes, Rancher, Helm, Docker, Cilium, Rook, Ceph, MinIO, S3, PortWorx, Load balancing, ingress, and service networking, Ansible, Terraform, Desired State Configuration, Python, PowerShell, and scripting/automation, Distributed systems, cloud computing, and microservices architecture, Monitoring/observability practices and tools, Incident response frameworks and SLO-based operations.
• TS/SCI security clearance required, candidate will not be considered without.
Preferred:
• CKA (if not used to meet required cert)
• RHCSA
• AWS DevOps Engineer or AZ-400
• CCSP
• Advanced observability certifications (Datadog, New Relic, Dynatrace, etc.)
• Formal incident management or SRE-focused training
• Building scalable, fault-tolerant cloud-native systems across hybrid or multi-cloud environments.
• Developing or supporting enterprise CI/CD pipelines.
• Managing complex Kubernetes clusters across on-prem and cloud platforms.
• Implementing enterprise observability stacks (e.g., Prometheus, Loki, Grafana, ELK, Open Telemetry).
• Supporting large-scale infrastructure within DoD or Intelligence Community environments.
Company:
Koniag Government Services is a Professional Services and Operational Management to Federal Government. Founded in 1971, the company is headquartered in Chantilly, USA, with a team of 1001-5000 employees. The company is currently Late Stage.

What Koniag Government Services employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom