1

Azure Site Reliability Engineer Jobs in Virginia

Site Reliability Engineer

Sterling, VA · On-site

$56.50 - $75/hr

Site Reliability Engineer Location: Sterling, VA Clearance: TS/SCI Poly **This position is ... Azure Certified Solutions Architect Expert. Salary & Benefits The salary associated with this ...

Site Reliability Engineer

Sterling, VA · On-site

$56.50 - $75/hr

The Site Reliability Engineer (SRE) collaboratively works closely with the contract leadership ... Azure Certified Solutions Architect Expert. Salary & Benefits The salary associated with this ...

Site Reliability Engineer (SRE)

Vienna, VA · On-site

$57.25 - $76/hr

... SRE principles for highly scalable and reliable systems • Possess a bachelor's degree • ... AWS, Azure) • Can establish and maintain a high level of client trust and confidence with your ...

Staff Site Reliability Engineer

Fairfax, VA · On-site

$56.50 - $75.25/hr

Citizenship / No clearance needed / 100% remote within the US Staff Site Reliability Engineer ... This role requires deep expertise across multiple cloud platforms (Azure and AWS) and container ...

Staff Site Reliability Engineer

Fairfax, VA · On-site

$56.50 - $75.25/hr

Citizenship / No clearance needed / 100% remote within the US Staff Site Reliability Engineer ... This role requires deep expertise across multiple cloud platforms (Azure and AWS) and container ...

Site Reliability Engineer - Hybrid

Reston, VA · On-site

$59.25 - $78.75/hr

Site Reliability Engineer V Location: Reston, VA (Hybrid onsite - 3 days a week from day 1) ... Design, implement, and manage cloud-based infrastructure using platforms like AWS, Azure, or GCP.

Site Reliability Engineer

Arlington, VA · On-site

$230K - $250K/yr

GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability ... AWS/Azure certifications * DevOps certifications * ITIL preferred Clearance Required: Must have an ...

Site Reliability Engineer

Arlington, VA · On-site

$230K - $250K/yr

GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability ... AWS/Azure certifications * DevOps certifications * ITIL preferred Clearance Required: Must have an ...

Site Reliability Engineer (SRE)

Vienna, VA · On-site

$57.25 - $76/hr

The AWS Site Reliability Engineer (SRE) is responsible for the operational health, availability, and performance of the AWS and Databricks environments built by the Platform Engineering team. You ...

Senior Technology Site Reliability Engineer

Reston, VA · On-site

$59.25 - $78.75/hr

Senior Technology Site Reliability Engineer Cooley is seeking a Senior Site Reliability Engineer to ... Experience working with advanced ETL data workflows including technologies such as AWS EMR, Azure ...

Site Reliability Engineer

Richmond, VA · On-site

$56.50 - $75/hr

Site Reliability Engineer One and Done Virtual Interview Needs to be onsite from day 1 in Richmond Virginia Only candidates that can convert in 12 months with no sponsorship Must haves: Log Data The ...

Site Reliability Engineer - CTJ - POLY

Reston, VA · On-site

$59.25 - $78.75/hr

Azure Data Transfer enables secure access and data transfer between enclaves and supports multiple transfer and access patterns for highly regulated industries. In this role, you will apply SRE ...

Everforth ECS is seeking a Senior Site Reliability Engineer to work remotely . Everforth ECS is seeking talented professionals to join our successful and growing team in building the next-generation ...

New

Everforth ECS is seeking a Senior Site Reliability Engineer to work remotely . Everforth ECS is seeking talented professionals to join our successful and growing team in building the next-generation ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

The SRE Engineer will improve the reliability, availability, performance, and operational resilience of mission-critical systems for a federal enterprise program. Responsibilities : • Define ...

next page

Showing results 1-20

Azure Site Reliability Engineer information

What is the difference between Azure Site Reliability Engineer vs Cloud Engineer?

AspectAzure Site Reliability EngineerCloud Engineer
CertificationsAzure certifications, SRE-related skillsCloud platform certifications (Azure, AWS, GCP)
Work EnvironmentFocus on reliability, monitoring, automation in AzureDesign, implement, manage cloud infrastructure across platforms
Industry UsageTech companies using Azure for scalable, reliable servicesOrganizations adopting cloud solutions, multi-cloud environments

Azure Site Reliability Engineers specialize in maintaining and improving the reliability of Azure-based systems, focusing on automation and monitoring. Cloud Engineers have a broader scope, working across various cloud platforms to design and manage cloud infrastructure. While both roles require cloud certifications and involve cloud environments, SREs emphasize system reliability within Azure, whereas Cloud Engineers focus on overall cloud architecture and deployment.

What job categories do people searching Azure Site Reliability Engineer jobs in Virginia look for? The top searched job categories for Azure Site Reliability Engineer jobs in Virginia are:
Infographic showing various Azure Site Reliability Engineer job openings in Virginia as of August 2026, with employment types broken down into 33% Full Time, and 67% Contract. Highlights an 100% In-person job distribution.

Site Reliability Engineer

Nightwing

Sterling, VA • On-site

$56.50 - $75/hr

Full-time

Medical, Dental, Vision, Retirement, PTO

Re-posted 23 days ago


Job description

Nightwing provides technically advanced full-spectrum cyber, data operations, systems integration and intelligence mission support services to meet our customers' most demanding challenges. Our capabilities include cyber space operations, cyber defense and resiliency, vulnerability research, ubiquitous technical surveillance, data intelligence, lifecycle mission enablement, and software modernization. Nightwing brings disruptive technologies, agility, and competitive offerings to customers in the intelligence community, defense, civil, and commercial markets.
Job Title: Site Reliability Engineer
Location: Sterling, VA
Clearance: TS/SCI Poly
**This position is CONTINGENT upon contract award**
The Site Reliability Engineer (SRE) collaboratively works closely with the contract leadership, Platform teams, and Sponsor to refine the operational and technical strategy to automate key portions of IT operations and enable the Product team (Platform) to bring new software or new features to production as quickly as possible. The SRE executes and analyzes manual IT operations/admin tasks (log analysis, performance tuning, patch management, testing, and incident response) and converts them to automated tasks. The SRE works with the Platform, Network and Data Operations teams to assist in deployment planning and onboard systems. They assist with monitoring, system analysis, and IT operations support. Daily tasks include, but are not limited to:
  • Work with Sponsor, Mission partners, and technical personnel to deliver robust scalable operations architecture that meets the customer goals for the enterprise.
  • Analyze, define, and document requirements for data, workflow, logical processes, hardware and operating system environment, and network connectivity, other system interfaces, internal and external checks and controls, and outputs.
  • Monitor and track metrics, logs and traces across all services in the system/network and provide context for identifying root causes in the event of an incident, performance degradation, or availability issue.
  • Perform Network/Cloud optimization and resilience planning
  • Develop capabilities to automate hardware/software provisioning, monitoring, patching, and troubleshooting.
  • Collaborate with and assist Platform team and leadership in network and security health, intrusions or inappropriate activities.
  • Optimize business processes, workflows, and service operations by building efficient on-call processes and streamlining alerting workflows.
  • Leverage operational data to automate systems administration, operations and incident response processes to improve enterprise reliability to manage IT environment complexity.
  • Works with LSA, Lab Manager, and CM to compose technical documents including Design, Deployment, System specifications and Host Nation baselines, updates, user's manuals, training materials, installation guides, proposals, and reports.
  • Work with the OM to implement ITSM best practices for ICA/Service discrepancy and reporting, issue resolution and operations support to include Tier 2/3 escalation.

Required Skills:
  • Programming: Proficiency in at least one programming language (e.g., Python, Go, Java, or JavaScript) is essential for automating tasks and developing tools.
  • Linux/Unix Systems Administration: Strong knowledge of Linux/Unix operating systems, including command-line tools and system administration tasks.
  • Networking: Understanding of network protocols, infrastructure, and troubleshooting techniques.
  • Database Management: Familiarity with database technologies and principles.
  • Automation: Experience with automation tools and techniques, such as configuration management (e.g., Ansible, Puppet, Chef) and orchestration (e.g., Kubernetes).
  • Monitoring and Logging: Experience with monitoring tools and logging systems.
  • Problem-Solving: Strong analytical and problem-solving skills to diagnose and resolve system issues.
  • Communication: Ability to communicate technical information clearly and concisely to both technical and non-technical audiences.
  • Collaboration: Ability to work effectively with cross-functional teams, including software developers and operations personnel.

Desired Skills:
  • Cloud Technologies: Experience with cloud platforms (e.g., AWS, Google Cloud, Azure).
  • Containerization: Knowledge of containerization technologies (e.g., Docker, Kubernetes).
  • DevOps Principles: Understanding DevOps principles and practices.
  • Service Level Objectives (SLOs) and Service Level Agreements (SLAs): Experience with defining, tracking, and managing SLOs and SLAs.
  • Data Analysis: Experience with data analysis and visualization tools.

Desired Certs:
  • Global Skill Development Council (GSDC) Site Reliability Engineering (SRE) Foundation Certification (CSREF).
  • AWS Certified SysOps Administrator - Associate.
  • Google Cloud Certified Professional Cloud Architect.
  • Azure Certified Solutions Architect Expert.

Salary & Benefits
The salary associated with this position ($122,000-$253,000) is commensurate with the selected candidate's qualifications, years of relevant experience, and demonstrated level of expertise. Compensation will be determined based on these factors to ensure alignment with skills, responsibilities, and market standards.
Nightwing offers medical, vision and dental insurance coverage in addition to a 401k plan, PTO, Holidays, and additional insurances.
Nightwing is An Equal Opportunity/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or veteran status, age or any other federally protected class.