1

Sre Manager Jobs in Virginia (NOW HIRING)

Site Reliability Engineer (SRE)

Vienna, VA · On-site

$57.25 - $76/hr

The AWS Site Reliability Engineer (SRE) is responsible for the operational health, availability ... CloudWatch, performance tuning in cloud environments, IaC tools, Databricks management and ...

Site Reliability Engineer

Sterling, VA · On-site

$56.50 - $75/hr

The SRE executes and analyzes manual IT operations/admin tasks (log analysis, performance tuning, patch management, testing, and incident response) and converts them to automated tasks. The SRE works ...

Site Reliability Engineer

Sterling, VA · On-site

$56.50 - $75/hr

The SRE executes and analyzes manual IT operations/admin tasks (log analysis, performance tuning, patch management, testing, and incident response) and converts them to automated tasks. The SRE works ...

Site Reliability Engineer

Richmond, VA · On-site

$56.50 - $75/hr

Site Reliability Engineer One and Done Virtual Interview Needs to be onsite from day 1 in Richmond Virginia Only candidates that can convert in 12 months with no sponsorship Must haves: Log Data The ...

SRE ENGINEER/ MANAGER

Reston, VA · On-site

$59.25 - $78.75/hr

Job Summary (Sr. Manager SRE): - Design, implement, and manage scalable, secure, and fault-tolerant cloud infrastructure using AWS, Azure, or GCP. - Automate infrastructure provisioning and ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

... manage Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets for ... Site reliability engineering, monitoring, automation, incident response, performance optimization ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

... manage Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets for ... Site reliability engineering, monitoring, automation, incident response, performance optimization ...

Site Reliability Engineer (SRE)

Vienna, VA · On-site

$57.25 - $76/hr

Up to 2 years in duration MUST HAVES: • Minimum of 8 years of experience as a Site Reliability Engineerwith a strong understanding of SRE principles for highly scalable and reliable systems • ...

Site Reliability Engineer - Hybrid

Reston, VA · On-site

$59.25 - $78.75/hr

Second round would be an in-person interview Manager's call notes * This is an SRE role. SRE is under a shared services team within Fannie Mae who works with different application teams. So, multi ...

Site Reliability Engineer

Herndon, VA · On-site

$86K - $198K/yr

Site Reliability Engineer The Opportunity: Engineering to make a system more resilient and ... Experience with performing Release Management activities Clearance: Applicants selected will be ...

Site Reliability Engineer

Reston, VA · On-site

$59.25 - $78.75/hr

Site Reliability Engineer Location: Reston, VA (Hybrid) Duration: long-term Required Skills: * Strong experience with Kubernetes administration and platform engineering. * Expertise in PostgreSQL ...

next page

Showing results 1-20

Sre Manager information

See Virginia salary details

$61.5K

$116.5K

$167.1K

How much do sre manager jobs pay per year?

As of Jul 28, 2026, the average yearly pay for sre manager in Virginia is $116,480.00, according to ZipRecruiter salary data. Most workers in this role earn between $93,700.00 and $138,800.00 per year, depending on experience, location, and employer.

What are SRE Managers?

SRE Managers are leaders responsible for overseeing Site Reliability Engineering (SRE) teams. They ensure the reliability, scalability, and performance of software systems by guiding engineers in implementing best practices, automation, and monitoring processes. SRE Managers collaborate closely with development and operations teams to balance feature development with system stability. Their role also includes mentoring SREs, managing incident response, and driving improvements in system reliability and operational efficiency.

What are the key skills and qualifications needed to thrive as an SRE Manager, and why are they important?

To thrive as an SRE Manager, you need a deep understanding of site reliability engineering principles, strong experience with systems architecture, and a background in computer science or a related field. Familiarity with tools such as Kubernetes, Prometheus, cloud platforms, and CI/CD pipelines, as well as certifications like AWS Certified Solutions Architect, are commonly expected. Leadership, effective communication, and problem-solving skills are crucial for driving team performance and collaborating across departments. These skills ensure high system reliability, efficient incident management, and a culture of continuous improvement within technical organizations.

What are some common challenges SRE Managers face when leading Site Reliability Engineering teams?

SRE Managers often encounter challenges balancing reliability with rapid development, ensuring their teams have the right mix of software engineering and operations skills. They must also foster a culture of continuous improvement while managing on-call rotations and incident response without causing burnout. Additionally, collaborating effectively with development and product teams to set realistic service level objectives (SLOs) and drive adoption of SRE best practices can require strong communication and negotiation skills.

What is the difference between Sre Manager vs DevOps Engineer?

AspectSre ManagerDevOps Engineer
CredentialsTypically requires a Bachelor's/Master's in CS or related field, with certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentLeads teams, manages incident response, and oversees reliability strategiesFocuses on automation, CI/CD pipelines, and infrastructure deployment
Industry UsageCommon in large tech companies, financial services, and cloud providersWidely used across startups, tech firms, and enterprises adopting DevOps practices

The Sre Manager and DevOps Engineer roles share overlapping skills in cloud computing, automation, and infrastructure management. While the Sre Manager oversees reliability and team coordination, the DevOps Engineer focuses on implementing automation tools and deployment pipelines. Both roles are crucial for modern IT operations, but the Sre Manager typically has a broader leadership responsibility, whereas the DevOps Engineer is more hands-on with technical implementation.

What are the most commonly searched types of Sre jobs in Virginia? The most popular types of Sre jobs in Virginia are:
What are popular job titles related to Sre Manager jobs in Virginia? For Sre Manager jobs in Virginia, the most frequently searched job titles are:
What job categories do people searching Sre Manager jobs in Virginia look for? The top searched job categories for Sre Manager jobs in Virginia are:
What cities in Virginia are hiring for Sre Manager jobs? Cities in Virginia with the most Sre Manager job openings:
Infographic showing various Sre Manager job openings in Virginia as of July 2026, with employment types broken down into 84% Full Time, 14% Part Time, and 2% Contract. Highlights an 94% Physical, 2% Hybrid, and 4% Remote job distribution, with an average salary of $116,480 per year, or $56 per hour.
Site Reliability Engineer (SRE)

Site Reliability Engineer (SRE)

System One Holdings, LLC

Mclean, VA • On-site

$57.50 - $76.50/hr

Full-time

Medical, Dental, Vision, Life, Retirement

Posted 5 days ago


Job description

Site Reliability Engineer (SRE)
Remote
No sponsorship available. Must be able to obtain a Public Trust clearance.
What You Will Do
We are seeking a Site Reliability Engineer (SRE) to support the SBA Disaster Lending Platform modernization effort in a remote capacity. This role will help establish and mature SRE practices across AWS cloud environments, with a focus on reliability, automation, scalability, observability, incident response, and operational excellence.
In this role, you will work closely with engineering, DevOps, cloud, security, and product teams to improve system resilience, reduce downtime, strengthen deployment practices, and support reliable cloud-based application delivery in an Agile environment.
Responsibilities include:
• Help establish and mature SRE practices within an Agile Scrum delivery environment.
• Support system design reviews to identify reliability risks, failure points, scalability concerns, and opportunities for automation.
• Improve operational readiness by contributing to code reviews, deployment reviews, monitoring practices, and reliability-focused engineering standards.
• Support incident management activities, including troubleshooting, root-cause analysis, mitigation planning, and post-incident improvements.
• Build and maintain automation to improve reliability, reduce manual effort, and support self-healing cloud infrastructure.
• Support AWS cloud platform operations across monitoring, logging, security, scalability, and availability.
• Work with CI/CD and Infrastructure as Code tools to support repeatable, secure, and reliable deployments.
• Create and maintain clear technical documentation for systems, processes, runbooks, and operational procedures.
• Collaborate with cross-functional teams and stakeholders to promote DevOps, automation, and reliability best practices.
What You Will Need
• Minimum of four years of experience supporting the reliability, scalability, security, and operational excellence of AWS cloud platforms.
• Bachelor's degree required, or four additional years of relevant experience in lieu of a degree.
• Hands-on experience with CI/CD and Infrastructure as Code tools such as Terraform, Ansible Automation Platform, GitLab, Artifactory, and Packer.
• Strong scripting and automation experience using Python, PowerShell, and Bash; Python experience is preferred.
• Experience supporting Windows and Linux environments.
• Strong understanding of networking concepts, cloud troubleshooting, monitoring, logging, and incident response.
• Experience designing, deploying, or supporting cloud-based systems with a focus on reliability, scalability, security, and performance.
• Knowledge of source control best practices.
• Experience working in Agile delivery environments, including Scrum, Kanban, SAFe, or similar methodologies.
• Strong analytical, troubleshooting, and problem-solving skills, including the ability to resolve complex technical issues in high-pressure situations.
• Strong communication skills and the ability to collaborate effectively with technical teams, stakeholders, and cross-functional partners.
• Must be authorized to work in the United States without sponsorship and able to obtain a Public Trust clearance.
Nice to Have
• Current or prior government contracting experience.
• Red Hat, CompTIA, AWS, or related technical certifications.
• Experience mentoring technical teams or helping promote DevOps/SRE practices across engineering groups.
System One, and its subsidiaries including Joulé and Mountain Ltd., are leaders in delivering outsourced services and workforce solutions across North America. We help clients get work done more efficiently and economically, without compromising quality. System One not only serves as a valued partner for our clients, but we offer eligible employees health and welfare benefits coverage options including medical, dental, vision, spending accounts, life insurance, voluntary plans, as well as participation in a 401(k) plan.
System One is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity, age, national origin, disability, family care or medical leave status, genetic information, veteran status, marital status, or any other characteristic protected by applicable federal, state, or local law.
#M1
#LI-CS1
Ref: #851-Rockville-S1