1

Reliability Manager Jobs in Chesapeake, VA (NOW HIRING)

Under the Service Management, Integration, and Transport (SMIT) program, the Leidos team delivers ... As part of the SRE organization you will develop and execute tests focused on system resilience ...

SRE Cyber Tools - Splunk Admin.

Norfolk, VA · On-site

$55.25 - $73.25/hr

Under the Service Management, Integration, and Transport (SMIT) program, the Leidos team delivers ... As part of the SRE organization you will develop and execute tests focused on system resilience ...

* Locations One Howmet Drive, Hampton, VA, 23661, US (On-site) * Job Schedule Full time * Export-Controlled Data This position entails access to export-controlled items and employment offers are ...

Job Info * Job Identification 118786 * Job Category Operations * Posting Date 08/13/2026, 11:40 AM * Locations One Howmet Drive, Hampton, VA, 23661, US (On-site) * Job Schedule Full time * Export ...

Senior Database Reliability Engineer: Job Type: Full-time Location: Remote Job Summary: Join our ... Key Responsibilities: • Design, implement, and manage highly available, resilient PostgreSQL ...

Senior Database Reliability Engineer: Job Type: Full-time Location: Remote Job Summary: Join our ... Key Responsibilities: • Design, implement, and manage highly available, resilient PostgreSQL ...

Senior Database Reliability Engineer: Job Type: Full-time Location: Remote Job Summary: Join our ... Key Responsibilities: • Design, implement, and manage highly available, resilient PostgreSQL ...

Showing results 21-40

Reliability Manager information

See Chesapeake, VA salary details

$60.2K

$114.1K

$163.6K

How much do reliability manager jobs pay per year?

As of Sep 8, 2026, the average yearly pay for reliability manager in Chesapeake, VA is $114,103.00, according to ZipRecruiter salary data. Most workers in this role earn between $91,800.00 and $136,000.00 per year, depending on experience, location, and employer.

What does a reliability manager do?

A Reliability Manager is responsible for ensuring that equipment, processes, and systems operate efficiently and consistently to minimize downtime and maximize performance. They develop and implement reliability strategies, conduct root cause analyses, and oversee preventive and predictive maintenance programs. Their role involves working closely with maintenance teams, engineers, and production staff to improve asset reliability and extend equipment lifespan. Additionally, they analyze failure data, recommend improvements, and help optimize operational costs through reliability-centered maintenance practices.

What are the key skills and qualifications needed to thrive as a reliability manager?

A Reliability Manager needs strong analytical skills, a solid background in engineering or maintenance, and experience with reliability-centered maintenance methodologies. Familiarity with tools like Failure Mode and Effects Analysis (FMEA), Root Cause Analysis (RCA), and certifications such as Certified Reliability Engineer (CRE) are often required. Leadership, problem-solving, and the ability to communicate complex technical information clearly are crucial soft skills for this role. These skills help ensure equipment uptime, optimize maintenance processes, and foster a culture of continuous improvement within the organization.

What are popular job titles related to Reliability Manager jobs in Chesapeake, VA?

For Reliability Manager jobs in Chesapeake, VA, the most frequently searched job titles are:

What job categories do people searching Reliability Manager jobs in Chesapeake, VA look for?

The top searched job categories for Reliability Manager jobs in Chesapeake, VA are:

What cities near Chesapeake, VA are hiring for Reliability Manager jobs?

Cities near Chesapeake, VA with the most Reliability Manager job openings:

Infographic showing various Reliability Manager job openings in Chesapeake, VA as of August 2026, with employment types broken down into 84% Full Time, 14% Part Time, and 2% Contract. Highlights an 79% Physical, 2% Hybrid, and 19% Remote job distribution, with an average salary of $114,103 per year, or $54.9 per hour.

SRE Cyber Tools - Splunk Admin.

Via Logic LLC

Norfolk, VA • On-site

$150 - $200/hr

Other

Posted 7 days ago


Job description

More About the Role:

Leidos is seeking a Site Reliability Engineer (SRE) Splunk Enterprise Administrator focused on Cyber support the largest IT services program for the Navy. Under the Service Management, Integration, and Transport (SMIT) program, the Leidos team delivers the core backbone of the Navy-Marine Corps Intranet, including cybersecurity services, network operations, service desk, and data transport. Leidos supports the Navy in unifying its shore-based networks and data management to improve capability and service while also saving significant dollars by focusing efforts under one enterprise network.

As part of the SRE organization you will develop and execute tests focused on system resilience, performance underload, and failure scenarios. You will also work in tandem with other Site Reliability Engineers (SREs) and development teams to create automated testing frameworks that simulate real-world conditions that validate system behavior under normal and stress conditions, ensuring our services are resilient and meet established service level objectives (SLOs). The SRE will support the operations and maintenance of the enterprise network. Your work will contribute to the development of robust and scalable services that operate reliably in production.

The Splunk Enterprise Administrator supports mission-critical cybersecurity operations by administering and maintaining distributed Splunk Enterprise platforms across hybrid cloud and on-premises environments. The role is responsible for daily platform operations, performance and ingestion monitoring, data onboarding support, incident troubleshooting, security compliance, documentation, and modernization activities in an Agile/DevOps operating culture.

Key Metrics of Success for the Team:
  • Improved system reliability, as measured by adherence to Service Level Objectives (SLOs) and reduced Mean Time to Recovery (MTTR).
  • Comprehensive and regularly updated automated test coverage for all critical systems and infrastructure components.
  • Timely identification and resolution of performance bottlenecks and failure points.
  • Integration of automated testing into the CI/CD pipeline, ensuring continuous reliability validation.
  • Increased scalability and performance of systems under high load due to effective performance testing.
What You'll Get to Do:
  • Administer, maintain, and perform daily Operations and Maintenance (O&M) for distributed Splunk Enterprise environments, including Search Heads, Indexers, Heavy Forwarders, Intermediate Forwarders, Universal Forwarders, and Deployment Servers across hybrid cloud and on-premises infrastructure.
  • Monitor platform health, data ingestion, indexing throughput, search performance, and retention utilization; troubleshoot data onboarding, parsing, field extraction, forwarding, indexing, and ingestion issues.
  • Support Splunk Cloud integrations and associated hybrid operational activities.
  • Maintain Splunk applications, dashboards, alerts, saved searches, and knowledge objects.
  • Support patching, vulnerability remediation, Security Technical Implementation Guide (STIG) compliance, and system hardening activities.
  • Support platform upgrades, migrations, infrastructure modernization, technology refresh, and automation initiatives using scripting and configuration-management tools.
  • Participate in incident response, outage troubleshooting, problem resolution, and root cause analysis.
  • Maintain operational documentation, architecture diagrams, runbooks, and standard operating procedures (SOPs), and coordinate with cybersecurity, network, server, engineering, and customer teams during operational and modernization activities.
  • Participate in after-hours support and an on‑call rotation as required.
You’ll Bring These Qualifications:
  • Requires BS degree and 5-10 years of prior relevant experience or Master’s with 4-8 years of prior relevant experience.
  • U.S. Citizen and posses an active Secret Security Clearance.
  • Minimum of DoD 8570.01 IAT Level II Certification required.
  • Experience designing, developing, and maintaining operational dashboards, visualizations, and executive reporting in Splunk Enterprise and Splunk Cloud, including IT Service Intelligence (ITSI), Service Analyzer, Glass Tables, and KPI-driven service health monitoring.
  • Experience with automated script design, coding, debugging, and maintenance skills (using bash, python, etc.) preferred.
  • Ability to work onsite at Norfolk Naval Station Monday through Friday day shift.
  • Must have a vendor certification e.g., Splunk Enterprise Certified Admin, Splunk Cloud Certified Admin, Scaled Agile Framework (SaFe).
  • Strong communication, analytical, and problem‑solving skills, ability to work in a team environment.
  • Minimum three years of experience administering Splunk Enterprise v9 and supporting distributed Splunk architectures in hybrid cloud and on‑premises production environments.
  • Strong working knowledge of Splunk data ingestion pipelines, indexing, parsing, forwarding, search optimization, and retention management.
  • Experience administering and troubleshooting RedHat Enterprise Linux (RHEL) 8 and/or RHEL9 servers.
  • Working knowledge of cybersecurity monitoring and Security Information and Event Management (SIEM) operations, with demonstrated experience troubleshooting operational incidents in complex enterprise environments.
  • Working knowledge of TCP/UDP networking, SSL certificates, Syslog, REST APIs, and authentication integrations.
  • Experience using at least one scripting or automation language: Python, Bash, or PowerShell.
  • Strong troubleshooting, analytical, communication, and documentation skills, with the ability to work independently, manage competing priorities, collaborate across technical teams, and maintain a customer‑focused approach in a high‑tempo production environment.
These Qualifications Would be Nice to Have:
  • Splunk Core Certified Power User and/or Splunk Enterprise Certified Admin certification. Experience supporting Splunk Cloud environments or integrations and both Splunk Enterprise v9 and v10.
  • Experience working in classified government or Department of Defense environments.
  • Familiarity with STIGs, the Risk Management Framework (RMF), vulnerability management, and compliance frameworks.
  • Familiarity with RHEL10, DevOps practices, and Infrastructure‑as‑Code concepts.
  • Experience with automation or configuration‑management tools such as Ansible, Jenkins, Chef, or Terraform.
NGEN

If you're looking for comfort, keep scrolling. At Leidos, we outthink, outbuild, and outpace the status quo — because the mission demands it. We're not hiring followers. We're recruiting the ones who disrupt, provoke, and refuse to fail. Step 10 is ancient history. We're already at step 30 — and moving faster than anyone else dares.

Original Posting:

August 5, 2026

For U.S. Positions: While subject to change based on business needs, Leidos reasonably anticipates that this job requisition will remain open for at least 3 days with an anticipated close date of no earlier than 3 days after the original posting date as listed above.

Pay Range:

Pay Range $107,900.00 - $195,050.00

The Leidos pay range for this job level is a general guideline only and not a guarantee of compensation or salary. Additional factors considered in extending an offer include (but are not limited to) responsibilities of the job, education, experience, knowledge, skills, and abilities, as well as internal equity, alignment with market data, applicable bargaining agreement (if any), or other law.

#J-18808-Ljbffr