1

Site Reliability Engineer Manager Jobs in Virginia

Senior Site Reliability Engineer

Mclean, VA Β· On-site

$58.50 - $77.75/hr

Overview Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS ... As a Senior Site Reliability Engineer, you will play a key role in designing, operating, and ...

Site Reliability Engineer II

Mclean, VA Β· On-site

$57.50 - $76.50/hr

Overview Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS ... As an SRE II, you will help operate and improve the reliability, scalability, and performance of ...

Site Reliability Engineer II

Mclean, VA Β· On-site

$58.50 - $77.75/hr

Overview Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS ... As an SRE II, you will help operate and improve the reliability, scalability, and performance of ...

Senior Site Reliability Engineer

Mclean, VA Β· On-site

$58.50 - $77.75/hr

Overview Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS ... As a Senior Site Reliability Engineer, you will play a key role in designing, operating, and ...

Senior Site Reliability Engineer

Mclean, VA Β· On-site

$57.50 - $76.50/hr

Overview Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS ... As a Senior Site Reliability Engineer, you will play a key role in designing, operating, and ...

Site Reliability Engineer II

Mclean, VA Β· On-site

$58.50 - $77.75/hr

Overview Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS ... As an SRE II, you will help operate and improve the reliability, scalability, and performance of ...

Site Reliability Engineer II

Arlington, VA Β· On-site

$65.75 - $87.25/hr

... SRE, DevOps, systems administration, or cloud support * Solid Linux administration (RHEL, Amazon ... AWS Infrastructure Management * Kubernetes/EKS Administration * Linux Administration * CI/CD ...

Senior Site Reliability Engineer

Arlington, VA Β· On-site

$65.50 - $87.25/hr

Collaborate closely with cross-functional teams, product managers, and stakeholders to align on ... Desired Qualifications * 10+ years in a Site Reliability Engineer or DevOps role supporting a SaaS ...

Staff Site Reliability Engineer

Reston, VA Β· On-site

$59.25 - $78.75/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work on the platform. As a Staff Site ...

Infrastructure & Automation β€’ Design, deploy, and manage cloud infrastructure using ... SRE practices and tools β€’ Document systems, processes, and runbooks β€’ Drive continuous ...

Staff Site Reliability Engineer

Reston, VA Β· On-site

$59.25 - $78.75/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work on the platform. As a Staff Site ...

Senior Site Reliability Engineer

Arlington, VA Β· On-site

$65.50 - $87.25/hr

Collaborate closely with cross-functional teams, product managers, and stakeholders to align on ... Desired Qualifications * 10+ years in a Site Reliability Engineer or DevOps role supporting a SaaS ...

Infrastructure & Automation β€’ Design, deploy, and manage cloud infrastructure using ... SRE practices and tools β€’ Document systems, processes, and runbooks β€’ Drive continuous ...

Site Reliability Engineer II Metro DC Β· Hybrid Β· 24/7 FedRAMP Operations Β· Rotational Shift Β· Initial Contract till March 27. KEY REQUIREMENT This role requires US citizenship and residence on US ...

SRE focus on observability

Richmond, VA Β· On-site

$55.75 - $74.25/hr

RIT Solutions, Inc. is a company seeking Site Reliability Engineers with a focus on observability. The role involves developing and maintaining queries and dashboards for analyzing customer impact ...

Contract Site Reliability Engineer II

Reston, VA Β· On-site

$59.25 - $78.75/hr

AWS Infrastructure Management * Kubernetes/EKS Administration * Linux Administration (RHEL, Amazon ... Cloud Support * DevOps * Site Reliability Engineering (SRE) * Systems Administration Tools ...

New

Senior Site Reliability Engineer

Reston, VA Β· On-site

$59.25 - $78.75/hr

Collaborate closely with cross-functional teams, product managers, and stakeholders to align on ... Desired Qualifications * 10+ years in a Site Reliability Engineer or DevOps role supporting a SaaS ...

Showing results 41-60

Site Reliability Engineer Manager information

See Virginia salary details

$10

$63

$91

How much do site reliability engineer manager jobs pay per hour?

As of Sep 14, 2026, the average hourly pay for site reliability engineer manager in Virginia is $63.20, according to ZipRecruiter salary data. Most workers in this role earn between $54.33 and $72.21 per hour, depending on experience, location, and employer.

What is a site reliability engineer manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.

How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How much do site reliability engineer managers get paid?

Site Reliability Engineer Managers typically earn between $120,000 and $180,000 annually, depending on experience, location, and company size. They often oversee teams responsible for system reliability, incident response, and infrastructure automation, requiring strong leadership and technical skills.

Is a Site Reliability Engineer Manager a stressful job?

A Site Reliability Engineer Manager role can be stressful due to the responsibility of maintaining system uptime, managing incident responses, and ensuring reliability across complex infrastructure. The job often involves working under pressure, handling outages, and coordinating teams, but it also offers opportunities for problem-solving and leadership. Stress levels vary depending on company size, team structure, and workload management skills.

What are the most commonly searched types of Site Reliability Engineer jobs in Virginia?

The most popular types of Site Reliability Engineer jobs in Virginia are:

What cities in Virginia are hiring for Site Reliability Engineer Manager jobs?

Cities in Virginia with the most Site Reliability Engineer Manager job openings:

Infographic showing various Site Reliability Engineer Manager job openings in Virginia as of August 2026, with employment types broken down into 80% Full Time, 19% Part Time, and 1% Contract. Highlights an 81% Physical, 2% Hybrid, and 17% Remote job distribution, with an average salary of $131,446 per year, or $63.2 per hour.

Site Reliability Engineer II

Newport News, VA β€’ On-site

Phase2 Technology
Software DevelopmentΒ β€’Β 51 - 200 employees

$91K - $145K/yr

Other

Medical, Dental, Vision, Retirement, PTO

Posted 10 days ago


Job description

At Jefferson Lab,you'llchampioncutting-edgescience and operational excellence while shaping the future of discovery. Join us and make your mark - where excellence meets purpose, andgreat mindstruly matter.

The good-faith pay range for this role is $91,800 - $145,050 per year. Actual compensation may vary and may be above the posted range based on factors such as a candidate's skills, experience, education, certifications, and work location.

What your job will be like:

As a Site Reliability Engineer on the High Performance Data Facility (HPDF) team, you will help build and operate the facility's first systems on its path to operations. You will create the monitoring, alerting, and automation that the full facility will eventually run on, participate in incident response, and help define and report on the service level objectives that measure how well the facility serves its users. You will work as part of a small site reliability engineering team, with day to day direction from the team's lead, and alongside staff at both Jefferson Lab and Berkeley Lab. The users you support are research physicists and computational scientists, and helping them succeed is a core measure of this role.

In this job you will:
  • Build and operate monitoring, logging, and alerting for HPDF systems using modern observability tooling (for example Prometheus, Grafana, OpenTelemetry, ELK), within guidelines set with the lead site reliability engineer and senior staff.
  • Develop automation and tooling in Python, Go, or shell that eliminates manual operations and reduces operational risk, following standard software development practices including version control, code review, and testing.
  • Respond to incidents and author the runbooks, postmortems, and operational documentation that convert each incident into a lasting improvement to facility reliability.
  • Implement and report on Service Level Objectives (SLOs) and Service Level Indicators (SLIs) defined with the architecture team and scientific stakeholders.
  • Support expert scientific users by helping research physicists and computational scientists understand system behavior and by translating their needs into reliability requirements.
  • Collaborate with colleagues at both Jefferson Lab and Berkeley Lab, sharing tooling, reviews, and operational practice across the HPDF partnership.
  • Conduct testing and performance analysis to validate reliability and resilience decisions and to identify bottlenecks.
Additional Responsibilities
  • Contribute to evaluations of vendor and open source technologies against reliability, performance, and security requirements.
  • Participate in an on-call rotation as the facility moves toward operations.
Experience
  • Required: 3 or more years related experience in Site Reliability Engineering, DevOps, systems engineering, or software engineering with operational responsibility.
  • Preferred: 3 or more years experience supporting scientific computing, HPC, or research environments; experience operating systems in support of users outside the engineer's own team
  • Preferred: Experience with containers and Kubernetes
  • Preferred: Practical experience applying AI assisted or autonomous automation to operational tasks, and routine use of AI tools in day to day engineering work.
  • Preferred: Experience with configuration management and infrastructure as code tools (for example Ansible, Terraform, Puppet).
Education
  • Required: Bachelor's Degree Computer Science or Related Field
  • Preferred: Master's Degree Computer Science or Related Field
Experience and Education Exchange

Education above the minimum may be substituted for experience. Relevant experience may not be substituted for education.

Knowledge, Skills, and Abilities
  • Solid Linux systems skills and command line fluency, with the ability to troubleshoot across the application, operating system, and network layers.
  • Scripting and automation ability in Python, Go, or shell, including familiarity with standard software development practices.
  • Working knowledge of monitoring and observability tools (for example Prometheus, Grafana, ELK, OpenTelemetry) and strong motivation to deepen that expertise in a new facility environment.
  • Clear written and verbal communication, and the ability to work productively with a community of scientific users and with colleagues across both laboratories.
  • A self starter's approach to learning new technologies and to identifying and solving problems within an assigned scope.
  • Familiarity with public cloud environments (AWS, Azure, GCP).
  • Networking fundamentals (IPv4/IPv6, DNS, firewalls, access control lists) and security conscious operational habits.
  • Exposure to HPC or scientific computing environments.
About Jefferson Lab

Join a community with a common purpose of solving the most challenging scientific and engineering problems of our time. The Jefferson Lab campusis located insoutheasternVirginiaamidst a vibrant and growing technology community.

A career at Jefferson Lab is more than a job. You will be part of "big science" and work alongside top scientists and engineers from around the world unlocking the secrets of our visible universe. Managed by SURATech, LLC, Thomas Jefferson National Accelerator Facility is entering an exciting period of mission growth and is seeking new team members ready to apply their skills and passion to have an impact. You could call it work, or you could call it a mission. We call it a challenge. We do things that will change the world.

Total Rewards at Jefferson Lab

At Jefferson Lab, we believe that a comprehensive employee benefits program is an important and meaningful part of the compensation employees receive. Our benefits program includes, but is not limited to:

  • Medical, Dental, and Vision Care Plans
  • Flexible Spending Accounts
  • Paid Time-off and Leave Programs (Paid Parental, vacation, holidays, and sick leave)
  • 401(k) Plan - 9% Lab Contribution; 100% vested
  • Flexible Work Arrangements
  • (Remote & Alternate Work Schedules available)
  • Tuition Assistance, Training and Professional Development Programs
  • Live near the waterways of the Chesapeake Bay region with access to nearby beaches,
  • mountains, and all major metropolitan centers on the East Coast

SURATech, LLC manages and operates the Thomas Jefferson National Accelerator Facility (Jefferson Lab). SURATech is an Equal Opportunity Employer.

SURATech is committed to providing reasonable accommodation for people with disabilities (unless doing so will result in an undue hardship). If you need a reasonable accommodation for any part of the employment process, please send an e-mail to recruiting@jlab.org or contact Human Resources by calling (757) 269-7100 and selecting option 1 between 8 am - 5 pm EST to provide the nature of your request.

Employment with SURATech is conditional upon DOE approval if at any time during your employment you are participating in a Foreign Government Talent Recruitment Program or Affiliated activity. Generally, such programs/activities include any foreign-state-sponsored attempt to acquire U.S.-funded scientific research through programs run or funded by the government that target scientists, engineers, students, academics, researchers, and entrepreneurs of all nationalities working or educated in the United States. This includes positions or appointments, both domestic and foreign, titled academic, professional, or institutional appointments whether or not remuneration is received and whether full-time, part-time or voluntary.

#J-18808-Ljbffr