1

Observability Site Reliability Engineer Jobs in Virginia

SRE Engineer

Arlington, VA ยท On-site

$65.75 - $87.25/hr

... observability solutions for production systems. โ€ข Support production and pre-production ... Site reliability engineering, monitoring, automation, incident response, performance optimization ...

Site Reliability Engineer

Richmond, VA ยท On-site

$56.50 - $75/hr

Role: Site Reliability Engineer One and Done Virtual Interview Needs to be onsite from day 1 in ... Log Data The more Observability tools the better( Datadog, New Relic, Splunk, ELK, etc) Programming ...

SRE Engineer

Arlington, VA ยท On-site

$65.75 - $87.25/hr

... observability solutions for production systems. โ€ข Support production and pre-production ... Site reliability engineering, monitoring, automation, incident response, performance optimization ...

Site Reliability Engineer - Hybrid

Reston, VA ยท On-site

$59.25 - $78.75/hr

Experience with Observability using tools such as AWS CloudWatch, Splunk/SignalFX, Dynatrace, and ... The SRE at Fannie Mae doesn't work 24*7. They get scheduled on a rotation basis. 20% of their job ...

Site Reliability Engineer

Mclean, VA ยท Remote

$57.50 - $76.50/hr

Site Reliability Engineer Job number: 880 This is a remote position. Ad Hoc is a technology company ... Building and maintaining observability tooling, including metrics, logging, alerting, and ...

Site Reliability Engineer

Mclean, VA ยท On-site

$57.50 - $76.50/hr

Site Reliability Engineer Job number: 880 This is a remote position. Ad Hoc is a technology company ... Building and maintaining observability tooling, including metrics, logging, alerting, and ...

Staff Site Reliability Engineer

Reston, VA ยท On-site

$59.25 - $78.75/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards ... Expert-level command of monitoring, observability, and alerting platforms (e.g., Datadog ...

next page

Showing results 1-20

Observability Site Reliability Engineer information

What engineer makes $500,000 a year?

A senior or principal Site Reliability Engineer (SRE) or Observability Engineer with extensive experience, specialized skills, and working at large tech companies can earn $500,000 or more annually. Compensation often includes base salary, bonuses, and stock options, especially in high-demand markets and organizations with complex infrastructure.

Is AI replacing SRE?

AI is augmenting the work of Site Reliability Engineers (SREs) by automating tasks such as monitoring, incident detection, and response. However, SREs are still essential for designing systems, managing complex issues, and making strategic decisions that require human judgment. AI tools are considered complementary rather than replacements for SREs' expertise and problem-solving skills.

What engineers make $200,000 a year?

Senior Site Reliability Engineers and Observability Engineers with extensive experience, advanced skills in cloud platforms, automation, and monitoring tools can earn $200,000 or more annually. High compensation often correlates with working at large tech companies, possessing specialized certifications, and managing complex, scalable systems.

What is the difference between Observability Site Reliability Engineer vs Monitoring Engineer?

AspectObservability Site Reliability EngineerMonitoring Engineer
FocusEnsuring system reliability through observability, automation, and incident responseImplementing and managing monitoring tools and dashboards
SkillsCloud platforms, scripting, incident management, observability toolsMonitoring tools, alerting systems, data analysis
Work EnvironmentDevOps teams, cloud infrastructure, large-scale systemsOperations teams, infrastructure monitoring

While both roles involve system health, the Observability Site Reliability Engineer focuses on comprehensive system reliability using observability practices, whereas Monitoring Engineers primarily manage monitoring tools and alerts. The SRE role emphasizes automation, incident response, and system resilience, making it broader in scope.

What engineers make $300,000 a year?

Senior Site Reliability Engineers and Observability Engineers with extensive experience, advanced skills in cloud platforms, automation, and monitoring tools can earn $300,000 or more annually. High compensation often correlates with working at large tech companies, possessing specialized certifications, and taking on leadership or highly technical roles.
What are popular job titles related to Observability Site Reliability Engineer jobs in Virginia? For Observability Site Reliability Engineer jobs in Virginia, the most frequently searched job titles are:
What job categories do people searching Observability Site Reliability Engineer jobs in Virginia look for? The top searched job categories for Observability Site Reliability Engineer jobs in Virginia are:
What cities in Virginia are hiring for Observability Site Reliability Engineer jobs? Cities in Virginia with the most Observability Site Reliability Engineer job openings:
Elastic SRE - Security & Observability with Security Clearance

Elastic SRE - Security & Observability with Security Clearance

Zachary Piper Solutions, LLC

Hampton, VA โ€ข On-site

$180K - $200K/yr

Other

Re-posted 23 days ago


Job description

Zachary Piper Solutions is seeking an experienced Elastic Site Reliability Engineer (SRE) to support a high-visibility federal engagement focused on observability, platform reliability, and security operations across classified environments. This position will support mission-critical Elastic infrastructure deployments at Langley AFB, VA. The ideal candidate will have deep expertise supporting enterprise Elastic Stack environments, Kubernetes-based deployments, and production SRE operations within secure or regulated infrastructure environments. SECRET CLEARANCE REQUIRED Key Responsibilities * Operate, maintain, and optimize large-scale Elastic Stack environments supporting logging, search, observability, and telemetry operations. * Ensure platform reliability, uptime, scalability, and performance across production mission systems. * Manage Kubernetes-based Elastic deployments, including ECK operator environments. * Develop and maintain automation for deployment workflows, monitoring, alerting, and incident response processes. * Integrate Elastic infrastructure with SIEM and security tooling including Splunk, EDR platforms, and telemetry systems. * Troubleshoot complex issues across distributed systems, infrastructure, and application environments. * Implement and support observability frameworks including logging, metrics, tracing, and monitoring solutions. * Support CI/CD pipelines and infrastructure-as-code initiatives within DevOps environments. * Maintain operational runbooks, escalation procedures, and technical documentation. * Participate in on-call support rotations and incident response activities. Qualifications * 5+ years of experience supporting Site Reliability Engineering, DevOps, or infrastructure operations environments. * Strong hands-on experience with Elastic Stack in enterprise production environments. * Advanced Kubernetes experience, including ECK operator deployments. * Strong Linux/Unix administration and networking fundamentals. * Experience supporting observability, telemetry, logging, and monitoring platforms. * Experience working within secure, classified, federal, or highly regulated environments. * Ability to work onsite at Langley AFB (VA).
Nice-to-Haves * Elastic certifications including Elastic Engineer, Security, or Observability. * Experience with Terraform, Ansible, and CI/CD pipeline automation. * Exposure to SIEM and EDR technologies including Splunk, CrowdStrike, or Trellix. * Experience supporting GovCloud, DoD, or federal infrastructure environments. * Prior experience supporting distributed logging or telemetry platforms. Soft Skills * Strong incident response and operational troubleshooting mindset. * Ability to remain calm and effective during production outages or high-pressure situations. * Strong collaboration skills across security, infrastructure, DevOps, and operations teams. * Excellent communication skills for escalation and operational coordination environments. * Self-sufficient and capable of operating independently within classified environments. Compensation & Benefits * Compensation: $180,000 - $200,000 annually. * Long-term federal engagement supporting mission-critical infrastructure initiatives. * Opportunity to support advanced observability and security operations within classified environments. Keywords elastic sre, site reliability engineer, elastic stack, elasticsearch, kibana, logstash, beats, observability, telemetry, logging infrastructure, distributed systems, kubernetes, eck, elastic cloud on kubernetes, sre, devops, platform engineering, infrastructure engineering, production support, linux administration, unix systems, networking, monitoring, tracing, metrics, incident response, automation, ci/cd, infrastructure as code, terraform, ansible, cloud infrastructure, distributed logging, telemetry systems, siem, splunk, edr, crowdstrike, trellix, platform reliability, reliability engineering, scalability, uptime, performance tuning, root cause analysis, operational excellence, federal infrastructure, dod, govcloud, classified systems, mission systems, secret clearance, hanscom afb, langley afb, secure environments, production engineering, elastic observability, elastic security, sre engineer, kubernetes engineer, platform sre, enterprise infrastructure, cloud operations, mission critical systems, elastic engineer, telemetry engineer, security operations, devsecops, automation engineer, distributed architecture, operational support #LI-RE1