1

Reliability Manager Jobs in Surry, VA (NOW HIRING)

* Locations One Howmet Drive, Hampton, VA, 23661, US (On-site) * Job Schedule Full time * Export-Controlled Data This position entails access to export-controlled items and employment offers are ...

Job Info * Job Identification 118786 * Job Category Operations * Posting Date 08/13/2026, 11:40 AM * Locations One Howmet Drive, Hampton, VA, 23661, US (On-site) * Job Schedule Full time * Export ...

Senior Database Reliability Engineer: Job Type: Full-time Location: Remote Job Summary: Join our ... Key Responsibilities: • Design, implement, and manage highly available, resilient PostgreSQL ...

Senior Database Reliability Engineer: Job Type: Full-time Location: Remote Job Summary: Join our ... Key Responsibilities: • Design, implement, and manage highly available, resilient PostgreSQL ...

Senior Database Reliability Engineer: Job Type: Full-time Location: Remote Job Summary: Join our ... Key Responsibilities: • Design, implement, and manage highly available, resilient PostgreSQL ...

Showing results 21-40

Reliability Manager information

See Surry, VA salary details

$89.2K

$169K

$242.3K

How much do reliability manager jobs pay per year?

As of Sep 7, 2026, the average yearly pay for reliability manager in Surry, VA is $168,974.00, according to ZipRecruiter salary data. Most workers in this role earn between $135,900.00 and $201,400.00 per year, depending on experience, location, and employer.

What does a reliability manager do?

A Reliability Manager is responsible for ensuring that equipment, processes, and systems operate efficiently and consistently to minimize downtime and maximize performance. They develop and implement reliability strategies, conduct root cause analyses, and oversee preventive and predictive maintenance programs. Their role involves working closely with maintenance teams, engineers, and production staff to improve asset reliability and extend equipment lifespan. Additionally, they analyze failure data, recommend improvements, and help optimize operational costs through reliability-centered maintenance practices.

What are the key skills and qualifications needed to thrive as a reliability manager?

A Reliability Manager needs strong analytical skills, a solid background in engineering or maintenance, and experience with reliability-centered maintenance methodologies. Familiarity with tools like Failure Mode and Effects Analysis (FMEA), Root Cause Analysis (RCA), and certifications such as Certified Reliability Engineer (CRE) are often required. Leadership, problem-solving, and the ability to communicate complex technical information clearly are crucial soft skills for this role. These skills help ensure equipment uptime, optimize maintenance processes, and foster a culture of continuous improvement within the organization.

What job categories do people searching Reliability Manager jobs in Surry, VA look for?

The top searched job categories for Reliability Manager jobs in Surry, VA are:

What cities near Surry, VA are hiring for Reliability Manager jobs?

Cities near Surry, VA with the most Reliability Manager job openings:

Infographic showing various Reliability Manager job openings in Surry, VA as of June 2026, with employment types broken down into 1% As Needed, 91% Full Time, 6% Part Time, and 2% Contract. Highlights an 94% Physical, 2% Hybrid, and 4% Remote job distribution, with an average salary of $168,974 per year, or $81.2 per hour.

Elastic SRE - Security & Observability with Security Clearance

Zachary Piper Solutions, LLC

Hampton, VA • On-site

$180K - $200K/yr

Other

Re-posted 12 days ago


Job description

Zachary Piper Solutions is seeking an experienced Elastic Site Reliability Engineer (SRE) to support a high-visibility federal engagement focused on observability, platform reliability, and security operations across classified environments. This position will support mission-critical Elastic infrastructure deployments at Langley AFB, VA. The ideal candidate will have deep expertise supporting enterprise Elastic Stack environments, Kubernetes-based deployments, and production SRE operations within secure or regulated infrastructure environments. SECRET CLEARANCE REQUIRED Key Responsibilities * Operate, maintain, and optimize large-scale Elastic Stack environments supporting logging, search, observability, and telemetry operations. * Ensure platform reliability, uptime, scalability, and performance across production mission systems. * Manage Kubernetes-based Elastic deployments, including ECK operator environments. * Develop and maintain automation for deployment workflows, monitoring, alerting, and incident response processes. * Integrate Elastic infrastructure with SIEM and security tooling including Splunk, EDR platforms, and telemetry systems. * Troubleshoot complex issues across distributed systems, infrastructure, and application environments. * Implement and support observability frameworks including logging, metrics, tracing, and monitoring solutions. * Support CI/CD pipelines and infrastructure-as-code initiatives within DevOps environments. * Maintain operational runbooks, escalation procedures, and technical documentation. * Participate in on-call support rotations and incident response activities. Qualifications * 5+ years of experience supporting Site Reliability Engineering, DevOps, or infrastructure operations environments. * Strong hands-on experience with Elastic Stack in enterprise production environments. * Advanced Kubernetes experience, including ECK operator deployments. * Strong Linux/Unix administration and networking fundamentals. * Experience supporting observability, telemetry, logging, and monitoring platforms. * Experience working within secure, classified, federal, or highly regulated environments. * Ability to work onsite at Langley AFB (VA).
Nice-to-Haves * Elastic certifications including Elastic Engineer, Security, or Observability. * Experience with Terraform, Ansible, and CI/CD pipeline automation. * Exposure to SIEM and EDR technologies including Splunk, CrowdStrike, or Trellix. * Experience supporting GovCloud, DoD, or federal infrastructure environments. * Prior experience supporting distributed logging or telemetry platforms. Soft Skills * Strong incident response and operational troubleshooting mindset. * Ability to remain calm and effective during production outages or high-pressure situations. * Strong collaboration skills across security, infrastructure, DevOps, and operations teams. * Excellent communication skills for escalation and operational coordination environments. * Self-sufficient and capable of operating independently within classified environments. Compensation & Benefits * Compensation: $180,000 - $200,000 annually. * Long-term federal engagement supporting mission-critical infrastructure initiatives. * Opportunity to support advanced observability and security operations within classified environments. Keywords elastic sre, site reliability engineer, elastic stack, elasticsearch, kibana, logstash, beats, observability, telemetry, logging infrastructure, distributed systems, kubernetes, eck, elastic cloud on kubernetes, sre, devops, platform engineering, infrastructure engineering, production support, linux administration, unix systems, networking, monitoring, tracing, metrics, incident response, automation, ci/cd, infrastructure as code, terraform, ansible, cloud infrastructure, distributed logging, telemetry systems, siem, splunk, edr, crowdstrike, trellix, platform reliability, reliability engineering, scalability, uptime, performance tuning, root cause analysis, operational excellence, federal infrastructure, dod, govcloud, classified systems, mission systems, secret clearance, hanscom afb, langley afb, secure environments, production engineering, elastic observability, elastic security, sre engineer, kubernetes engineer, platform sre, enterprise infrastructure, cloud operations, mission critical systems, elastic engineer, telemetry engineer, security operations, devsecops, automation engineer, distributed architecture, operational support #LI-RE1