1

Senior Reliability Engineer Jobs in Elkridge, MD

They are seeking an experienced Senior Site Reliability Engineer to enhance the reliability, availability, and operational performance of IT and cloud systems under the MC&FP Outreach and Digital ...

Barbaricum is seeking an experienced Senior Site Reliability Engineer to support the reliability, availability, automation, and operational performance of IT and cloud systems under the Military ...

Barbaricum is seeking an experienced Senior Site Reliability Engineer to support the reliability, availability, automation, and operational performance of IT and cloud systems under the Military ...

Barbaricum is seeking an experienced Senior Site Reliability Engineer to support the reliability, availability, automation, and operational performance of IT and cloud systems under the Military ...

Reliability Engineer

Rockville, MD ยท On-site

$104K - $131K/yr

Approximately 5+ years of reliability, mission assurance, or closely related engineering experience supporting flight hardware or similarly high-reliability systems; senior candidates typically ...

Reliability Engineer

Rockville, MD ยท On-site

$150 - $200/hr

Approximately 5+ years of reliability, mission assurance, or closely related engineering experience supporting flight hardware or similarly high-reliability systems; senior candidates typically ...

next page

Showing results 1-20

Senior Reliability Engineer information

See Elkridge, MD salary details

$21

$63

$91

How much do senior reliability engineer jobs pay per hour?

As of Sep 7, 2026, the average hourly pay for senior reliability engineer in Elkridge, MD is $63.62, according to ZipRecruiter salary data. Most workers in this role earn between $52.45 and $76.20 per hour, depending on experience, location, and employer.

What does a senior reliability engineer do?

A Senior Reliability Engineer is responsible for ensuring that systems, products, or processes operate reliably and efficiently over time. They analyze failure data, design reliability tests, develop maintenance strategies, and work with cross-functional teams to improve system performance and reduce downtime. Their expertise helps organizations minimize risk, optimize lifecycle costs, and maintain high standards of quality and safety. Senior Reliability Engineers often mentor junior team members and play a key role in developing reliability standards and best practices.

What are the key skills and qualifications needed to thrive as a senior reliability engineer?

To thrive as a Senior Reliability Engineer, you need expertise in reliability engineering principles, root cause analysis, and a relevant engineering degree such as mechanical, electrical, or industrial engineering. Familiarity with tools like FMEA, RCA software, CMMS, and certifications such as Certified Reliability Engineer (CRE) are often required. Strong analytical thinking, communication skills, and the ability to lead cross-functional teams set top performers apart. These skills are essential for minimizing downtime, improving system reliability, and ensuring safe, efficient operations.

What are some common challenges faced by senior reliability engineers, and how are they typically addressed within the team?

Senior Reliability Engineers often encounter challenges such as diagnosing complex system failures, balancing proactive maintenance with urgent reactive fixes, and ensuring consistent communication across multidisciplinary teams. These challenges are typically addressed through root cause analysis, prioritization frameworks, and fostering a culture of knowledge sharing. Regular collaboration with operations, maintenance, and engineering teams helps in developing effective solutions and continuous improvement strategies.

What is the difference between Senior Reliability Engineer vs Reliability Engineer?

AspectSenior Reliability EngineerReliability Engineer
CredentialsTypically requires 5+ years experience, certifications like CRE or Six SigmaEntry to mid-level, often with 2-4 years experience, similar certifications
Work EnvironmentDesigns and oversees reliability programs, leads projectsPerforms analysis, supports reliability improvements
Industry UsageUsed across manufacturing, energy, aerospaceCommon in same industries, often as a stepping stone to senior roles

The main difference between a Senior Reliability Engineer and a Reliability Engineer lies in experience, leadership responsibilities, and scope of work. Senior Reliability Engineers typically lead projects and develop strategies, while Reliability Engineers focus on analysis and supporting reliability initiatives. Both roles are vital in ensuring equipment and system dependability across industries.

How much do senior reliability engineers get paid?

Senior reliability engineers typically earn between $90,000 and $130,000 annually, depending on experience, industry, and location. They often have expertise in systems analysis, failure modes, and reliability tools like FMEA and RCM, which can influence compensation levels.

What are the most commonly searched types of Reliability Engineer jobs in Elkridge, MD?

The most popular types of Reliability Engineer jobs in Elkridge, MD are:

What cities near Elkridge, MD are hiring for Senior Reliability Engineer jobs?

Cities near Elkridge, MD with the most Senior Reliability Engineer job openings:

Infographic showing various Senior Reliability Engineer job openings in Elkridge, MD as of August 2026, with employment types broken down into 95% Full Time, 2% Part Time, and 3% Contract. Highlights an 86% Physical, 5% Hybrid, and 9% Remote job distribution, with an average salary of $132,323 per year, or $63.6 per hour.

Senior Reliability Engineer

Barbaricum

Washington, DC โ€ข On-site

Full-time

Re-posted 22 days ago


Job description

Job Summary:
Barbaricum is a rapidly growing government contractor providing leading-edge support to federal customers, focusing on Defense and National Security mission sets. They are seeking an experienced Senior Site Reliability Engineer to enhance the reliability, availability, and operational performance of IT and cloud systems under the MC&FP Outreach and Digital Enterprise Services contract.
Responsibilities:
โ€ข Monitor and maintain system reliability, availability, and performance across on-premises, cloud, and hybrid IT environments supporting MC&FP mission requirements.
โ€ข Implement proactive performance monitoring, automated alerting, incident response workflows, and resilience engineering practices to reduce downtime and improve operational visibility.
โ€ข Develop, maintain, and improve scalable automated infrastructure solutions that support reliable system operations and repeatable service delivery.
โ€ข Implement rollback strategies, recovery approaches, and chaos engineering practices to validate resilience, reduce operational risk, and improve system stability.
โ€ข Analyze usage patterns, capacity trends, and performance indicators to support dynamic scaling, resource optimization, and system improvement decisions.
โ€ข Develop and maintain real-time operational dashboards, reports, and metrics that enable rapid decision-making, leadership awareness, and system optimization.
โ€ข Respond to and resolve system outages, impairments, and service disruptions while coordinating with technical teams to minimize mission impact.
โ€ข Conduct post-incident reviews to identify root causes, document lessons learned, and implement preventative measures that reduce recurrence.
โ€ข Collaborate with software developers, cloud engineers, cybersecurity personnel, and operations teams to improve services, reliability patterns, deployment practices, and operational standards.
โ€ข Create and maintain system documentation, configuration standards, operational runbooks, monitoring procedures, and service reliability guidance.
โ€ข Automate common operations tasks to reduce manual workloads, improve consistency, and increase system efficiency.
โ€ข Implement security best practices across operational activities, infrastructure automation, monitoring, incident response, and system administration functions.
Qualifications:
Required:
โ€ข Bachelor's degree in Computer Science, Information Technology, Systems Engineering, Cybersecurity, or a related field.
โ€ข 10+ years of experience in site reliability engineering, systems administration, infrastructure operations, cloud operations, DevSecOps, or a similar technical role, particularly in a government, federal, defense, or secure IT setting.
โ€ข Expert knowledge of site reliability engineering practices, system monitoring, incident management, automation, performance tuning, and operational resilience.
โ€ข Strong understanding of Windows and Linux administration, infrastructure operations, system configuration, service management, and troubleshooting practices.
โ€ข Experience with automation platforms and configuration management tools such as Ansible, Puppet, Chef, or similar technologies.
โ€ข Proficiency with scripting languages such as Python, Shell, PowerShell, or similar tools used to automate operational and infrastructure tasks.
โ€ข Knowledge of cloud services and infrastructure across AWS, Microsoft Azure, Google Cloud, or comparable cloud environments.
โ€ข Strong understanding of network troubleshooting, configuration, connectivity analysis, system dependencies, and performance bottleneck identification.
โ€ข Ability to design, interpret, and maintain dashboards, alerts, metrics, logs, and operational reporting that support service health and decision-making.
โ€ข Ability to conduct root cause analysis, post-incident reviews, and corrective action planning in complex technical environments.
โ€ข Strong problem-solving skills and the ability to work under pressure during outages, impairments, and time-sensitive operational issues.
โ€ข Excellent written and verbal communication skills, with the ability to explain technical findings, incident impacts, and reliability recommendations to technical and non-technical stakeholders.
โ€ข Demonstrated experience maintaining reliable, scalable, and efficiently managed IT systems across on-premises, cloud, or hybrid environments.
โ€ข Experience developing automated infrastructure, operational scripts, monitoring solutions, dashboards, runbooks, and configuration standards.
โ€ข Experience supporting incident response, system outage resolution, post-incident reviews, root cause analysis, and operational improvement initiatives.
โ€ข Experience collaborating with development, infrastructure, cloud, cybersecurity, and program teams to improve reliability, security, and service performance.
โ€ข DoD Secret Security Clearance.
Preferred:
โ€ข Master's degree preferred.
โ€ข Certifications related to cloud computing, system administration, site reliability engineering, DevSecOps, or automation are beneficial.
Company:
Barbaricum is a government relations company that offers strategic communications, research, and analysis solutions. Founded in 2008, the company is headquartered in Washington, USA, with a team of 201-500 employees. The company is currently Growth Stage.