1

Site Reliability Manager Jobs in Virginia (NOW HIRING)

Cloud SRE

Richmond, VA · On-site

$56.50 - $75/hr

Build and optimize Infrastructure as Code (IaC) using Terraform to manage AWS resources related to SRE solutions, incorporating cost-efficient design principles o Optimize Infrastructure as Code (IaC ...

New

Strong understanding of JVM fundamentals (heap/memory management, garbage collection, OOM issues, thread analysis) * Proven experience with SRE practices, including: * Incident response and on-call ...

Site Reliability Engineer - Hybrid

Reston, VA · On-site

$59.25 - $78.75/hr

Second round would be an in-person interview Manager's call notes * This is an SRE role. SRE is under a shared services team within Fannie Mae who works with different application teams. So, multi ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

... manage Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets for ... Site reliability engineering, monitoring, automation, incident response, performance optimization ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

... manage Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets for ... Site reliability engineering, monitoring, automation, incident response, performance optimization ...

Manager, Site Reliability Engineering

Reston, VA · On-site

$59.25 - $78.75/hr

... build site reliability knowledge, and provide clear data. Responsibilities Key Responsibilities ... Incident and ServiceLifecycle Management: - Advisesteam members on performing data collection ...

Site Reliability Engineer

Arlington, VA · On-site

$230K - $250K/yr

GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability ... Define and manage SLIs, SLOs, and error budgets. * Automate operational tasks and eliminate manual ...

Serve as the equipment reliability Subject Matter Expert (SME) for the site. * Assist the Maintenance Manager in mentoring and leading Reliability Engineers, Planners, and members of the mill ...

Under the Service Management, Integration, and Transport (SMIT) program, the Leidos team delivers ... As part of the SRE organization you will develop and execute tests focused on system resilience ...

Site Reliability Data Engineer

Norfolk, VA · On-site

$55.25 - $73.25/hr

Under the Service Management, Integration, and Transport (SMIT) program, the Leidos team delivers ... As part of the SRE organization you will develop and execute tests focused on system resilience ...

Senior Site Reliability Engineer

Arlington, VA · On-site

$65.50 - $87.25/hr

Collaborate closely with cross-functional teams, product managers, and stakeholders to align on ... Desired Qualifications * 10+ years in a Site Reliability Engineer or DevOps role supporting a SaaS ...

next page

Showing results 1-20

Site Reliability Manager information

See Virginia salary details

$61.5K

$116.5K

$167.1K

How much do site reliability manager jobs pay per year?

As of Aug 24, 2026, the average yearly pay for site reliability manager in Virginia is $116,480.00, according to ZipRecruiter salary data. Most workers in this role earn between $93,700.00 and $138,800.00 per year, depending on experience, location, and employer.

What is the difference between Site Reliability Manager vs DevOps Engineer?

AspectSite Reliability ManagerDevOps Engineer
CredentialsTypically requires a Bachelor's in Computer Science, certifications like SRE or Cloud certificationsOften holds a Bachelor's in Computer Science or related field, with certifications in cloud platforms or automation tools
Work EnvironmentLeads teams managing large-scale systems, focusing on reliability and uptimeWorks on automation, CI/CD pipelines, and infrastructure deployment
Industry UsageCommon in tech, cloud services, and large enterprisesWidely used in startups, tech companies, and organizations adopting DevOps practices

The main difference is that Site Reliability Managers focus on ensuring system reliability and managing SRE teams, while DevOps Engineers concentrate on automation, deployment, and continuous integration. Both roles require technical expertise but serve different strategic objectives within IT operations.

What is the role of a site reliability manager?

A site reliability manager (SRM) is responsible for ensuring the reliability, availability, and performance of a company's IT systems and services. They often oversee incident response, implement automation tools, and collaborate with development teams to improve system stability and scalability, typically requiring knowledge of monitoring tools and scripting skills.
Infographic showing various Site Reliability Manager job openings in Virginia as of August 2026, with employment types broken down into 80% Full Time, 19% Part Time, and 1% Contract. Highlights an 81% Physical, 2% Hybrid, and 17% Remote job distribution, with an average salary of $116,480 per year, or $56 per hour.

$56.50 - $75/hr

Other

Medical, Dental, Vision, Life, Retirement

Posted 2 days ago

New


Job description

Job#: 3047318
Job Description:
Cloud Engineer/Developer
Role Overview
As a Senior Cloud Engineer/Developer in the Cloud SRE team, you will be responsible for designing and developing cloud solutions and engineering reliability tools. This is a developer-focused role within the infrastructure space, where you will apply software engineering practices to build scalable and reusable solutions. This position requires working EST hours and U.S. citizenship.
Key Responsibilities
  • Design, develop, and maintain reliability solutions and SRE utilities using Python in AWS environments to reduce toil, improve cloud platform reliability, and industrialize SRE practices across the system
    • Build automation scripts, APIs, and utilities in Python to reduce toil and improve platform reliability.
    • Implement observability and monitoring solutions (Grafana, AWS CloudWatch) leveraging Python for custom metrics and dashboards.
  • Build and optimize Infrastructure as Code (IaC) using Terraform to manage AWS resources related to SRE solutions, incorporating cost-efficient design principles

o Optimize Infrastructure as Code (IaC) with Terraform for AWS resources, integrating Python-based workflows.
  • Develop CI/CD pipelines and automated testing to ensure code quality, reliability, and rapid delivery of the solutions
  • Define SRE standards, best practices, and guidelines for adoption across teams; establish SRE metrics like SLI, SLOs, etc.
  • Apply software engineering best practices including version control, code reviews, test-driven development, and documentation to all development

Participate in incident management and on-call rotation, providing technical support for SRE tools, troubleshooting production issues, and collaborating with teams to reduce incident recurrence through proactive detection and pattern analysis
Stay current with emerging AWS services, SRE methodologies, and cloud-native development technologies, and drive adoption of innovative solutions
  • Collaborate within Agile and Scaled Agile frameworks with cross-functional teams to deliver integrated cloud automation solutions
  • Produce clear, blameless postmortems with actionable items and documented failure scenarios

Qualifications
  • Must have 5+ years of advanced Python development experience, building enterprise-grade, highly available tools, APIs, and utilities for AWS.
  • Bachelor's degree in computer science, Information Systems, or equivalent background or equivalent experience
  • 7+ years of extensive experience in software development with focus on reliability and platform engineering
  • 3+ years of hands-on experience developing solutions in AWS environments with deep understanding of core services (EC2, VPC, S3, Lambda, IAM, CloudFormation, EventBridge, Step Functions etc.) and resource cost optimization
  • 3+ years of experience applying SRE principles including observability, toil automation, SLIs/SLOs and reliability engineering
  • Expert-level proficiency with Infrastructure as Code (IaC) using Terraform, including module development and state management
  • Strong experience with CI/CD pipelines, automated testing frameworks, and DevOps practices
  • Experience with observability tools and practices including Grafana, AWS CloudWatch, AWS Canary

Experience defining, implementing, and managing SLOs/SLIs and error budgets; familiarity with conducting RCAs and producing postmortem documentation
  • Working experience in Agile and Scaled Agile environments and familiarity with ITSM processes (incident, change, and problem management), resilience testing and chaos engineering practices

Experience: A minimum of 7 years of experience in software development with a focus on reliability and platform engineering is required. This includes over 3 years of hands-on experience developing solutions in AWS environments and applying SRE principles.
Technical Skills:
  • Advanced Python development skills (5+ years) are required.
  • Expert-level proficiency with Infrastructure as Code (IaC) using Terraform.
  • Strong experience with CI/CD pipelines, automated testing frameworks, and DevOps practices.
  • Experience with observability tools such as Grafana and AWS CloudWatch.
  • Deep understanding of core AWS services (e.g., EC2, VPC, S3, Lambda, IAM, CloudFormation, EventBridge, Step Functions).
  • Experience defining and managing SLOs/SLIs and familiarity with ITSM processes.
Preferred Qualifications
  • Experience with GoLang or additional programming languages is a plus.
  • Working experience in Agile and Scaled Agile environments.

Everforth Apex is a world-class IT services company that serves thousands of clients across the globe. When you join Everforth Apex, you become part of a team that values innovation, collaboration, and continuous learning. We offer quality career resources, training, certifications, development opportunities, and a comprehensive benefits package. Our commitment to excellence is reflected in many awards, including ClearlyRateds Best of Staffing in Talent Satisfaction in the United States and Great Place to Work in the United Kingdom and Mexico.
Everforth Apex uses a virtual recruiter as part of the application process. Click for more details. By applying for this job, you agree to receive calls, AI-generated calls, text messages, or emails from Everforth Apex and its affiliates, and contracted partners. Frequency varies for text messages. Message and data rates may apply. Carriers are not liable for delayed or undelivered messages. You can reply STOP to cancel and HELP for help. You can access our privacy policy at
Everforth Apex Benefits Overview: Everforth Apex offers a range of supplemental benefits, including medical, dental, vision, life, disability, and other insurance plans that offer an optional layer of financial protection. We offer an ESPP (employee stock purchase program) and a 401K program which allows you to contribute typically within 30 days of starting, with a company match after 12 months of tenure. Everforth Apex also offers a HSA (Health Savings Account on the HDHP plan), a SupportLinc Employee Assistance Program (EAP) with up to 8 free counseling sessions, a corporate discount savings program and other discounts. In terms of professional development, Everforth Apex hosts an on-demand training program, provides access to certification prep and a library of technical and leadership courses/books/seminars once you have 6+ months of tenure, and certification discounts and other perks to associations that include CompTIA and IIBA. Everforth Apex has a dedicated customer service team for our Consultants that can address questions around benefits and other resources, as well as a certified Career Coach. You can access a full list of our benefits, programs, support teams and resources within our 'Welcome Packet' as well, which an Everforth Apex team member can provide.
Everforth Apex Systems is an equal opportunity employer. We do not discriminate or allow discrimination on the basis of race, color, religion, creed, sex (including pregnancy, childbirth, breastfeeding, or related medical conditions), age, sexual orientation, gender identity, national origin, ancestry, citizenship, genetic information, registered domestic partner status, marital status, disability, status as a crime victim, protected veteran status, political affiliation, union membership, or any other characteristic protected by law. Everforth Apex will consider qualified applicants with criminal histories in a manner consistent with the requirements of applicable law.
If you require an accommodation under the Americans with Disabilities Act to participate in an interview with a virtual recruiter or to use our website for a search or application, please contact our Benefits Department at or . Please note that this contact information is strictly to be used for medical ADA accommodations and that no other inquiries will be answered.
UnitedHealthcare creates and publishes the Transparency in Coverage Machine-Readable Files on behalf of Everforth Apex Systems.