1

Senior Site Reliability Engineer Jobs in Virginia

As a senior engineer in the organization, you will also provide mentorship within the SRE team and across peer engineering teams, helping elevate operational maturity, drive adoption of SRE practices ...

Site Reliability Engineer

Richmond, VA · On-site

$56.50 - $75/hr

Site Reliability Engineer One and Done Virtual Interview Needs to be onsite from day 1 in Richmond Virginia Only candidates that can convert in 12 months with no sponsorship Must haves: Log Data The ...

Escalation points for junior Site Reliability Engineers during complex or high-impact incidents. * Manage and execute complex manual Change Management tickets, by working closely with the service ...

Escalation points for junior site reliability engineers during complex or high-impact incidents. * Manage and execute complex manual Change Management tickets, by working closely with the service ...

Escalation points for junior Site Reliability Engineers during complex or high-impact incidents. * Manage and execute complex manual Change Management tickets, by working closely with the service ...

Escalation points for junior Site Reliability Engineers during complex or high-impact incidents. * Manage and execute complex manual Change Management tickets, by working closely with the service ...

Escalation points for junior site reliability engineers during complex or high-impact incidents. * Manage and execute complex manual Change Management tickets, by working closely with the service ...

Escalation points for junior site reliability engineers during complex or high-impact incidents. * Manage and execute complex manual Change Management tickets, by working closely with the service ...

Escalation points for junior site reliability engineers during complex or high-impact incidents. * Manage and execute complex manual Change Management tickets, by working closely with the service ...

Escalation points for junior Site Reliability Engineers during complex or high-impact incidents. * Manage and execute complex manual Change Management tickets, by working closely with the service ...

Escalation points for junior Site Reliability Engineers during complex or high-impact incidents. * Manage and execute complex manual Change Management tickets, by working closely with the service ...

SRE Engineer

Arlington, VA · On-site

$65.75 - $87.25/hr

The SRE Engineer will improve the reliability, availability, performance, and operational resilience of mission-critical systems for a federal enterprise program. Responsibilities : • Define ...

Site Reliability Engineer (SRE)

Vienna, VA · On-site

$57.25 - $76/hr

Up to 2 years in duration MUST HAVES: • Minimum of 8 years of experience as a Site Reliability Engineerwith a strong understanding of SRE principles for highly scalable and reliable systems • ...

Showing results 41-60

Senior Site Reliability Engineer information

See Virginia salary details

$21

$63

$91

How much do senior site reliability engineer jobs pay per hour?

As of Aug 24, 2026, the average hourly pay for senior site reliability engineer in Virginia is $63.86, according to ZipRecruiter salary data. Most workers in this role earn between $52.69 and $76.49 per hour, depending on experience, location, and employer.

What is a senior site reliability engineer?

Senior Site Reliability Engineers (SREs) are experienced IT professionals who ensure the reliability, scalability, and performance of complex software systems. They bridge the gap between software development and operations by automating processes, monitoring system health, responding to incidents, and implementing best practices for system stability. Senior SREs often lead teams, design resilient infrastructure, and mentor junior engineers to improve overall reliability and efficiency in technology environments.

What are the key skills and qualifications needed to thrive as a senior site reliability engineer?

To thrive as a Senior Site Reliability Engineer, you need deep experience in systems administration, software engineering, automation, and cloud infrastructure, typically supported by a related degree and several years in SRE or DevOps roles. Proficiency with tools like Kubernetes, Docker, Terraform, monitoring platforms (e.g., Prometheus, Datadog), and certifications in AWS, GCP, or Azure are common requirements. Strong problem-solving skills, effective communication, and a proactive mindset help you collaborate across teams and respond quickly to incidents. These skills ensure high system reliability, scalability, and rapid recovery from failures, which are critical to business continuity.

What are some typical challenges a senior site reliability engineer faces when balancing system reliability with rapid feature releases?

Senior Site Reliability Engineers often navigate the challenge of maintaining high system uptime and stability while supporting fast-paced software development cycles. This involves implementing robust automation and monitoring, collaborating closely with development teams to ensure new features don't compromise reliability, and proactively addressing potential bottlenecks. Effective communication and prioritization are key, as SREs must advocate for best practices in reliability without slowing down innovation. Continuous learning and adaptation are vital to keep up with evolving technology stacks and complex distributed systems.

What does a senior site reliability engineer do?

A senior site reliability engineer (SRE) is responsible for maintaining and improving the reliability, availability, and performance of large-scale systems and services. They develop automation tools, monitor system health, troubleshoot issues, and implement best practices to ensure continuous operation, often using skills in coding, systems administration, and cloud platforms.

What are the most commonly searched types of Site Reliability Engineer jobs in Virginia?

The most popular types of Site Reliability Engineer jobs in Virginia are:

What job categories do people searching Senior Site Reliability Engineer jobs in Virginia look for?

The top searched job categories for Senior Site Reliability Engineer jobs in Virginia are:

What cities in Virginia are hiring for Senior Site Reliability Engineer jobs?

Cities in Virginia with the most Senior Site Reliability Engineer job openings:

Infographic showing various Senior Site Reliability Engineer job openings in Virginia as of August 2026, with employment types broken down into 1% As Needed, 77% Full Time, 20% Part Time, and 2% Contract. Highlights an 93% Physical, 3% Hybrid, and 4% Remote job distribution, with an average salary of $132,831 per year, or $63.9 per hour.

Site Reliability Engineer IV

Capitolis

Sterling, VA • On-site

$120 - $160/hr

Other

Posted 18 days ago


Job description

Candescent is the leading cloud-based digital banking solutions provider for financial institutions. We are transforming digital banking with intelligent, cloud-powered solutions that connect account opening, digital banking, and branch experiences for financial institutions. Our advanced technology and developer tools enable seamless, differentiated customer journeys that elevate trust, service, and innovation. Success here requires flexibility in a fast-paced environment, a client-first mindset, and a commitment to delivering consistent, reliable results as part of a performance-driven, values-led team. With team members around the world, Candescent is an equal opportunity employer.

Position: Site Reliability Engineer IV

Experience: 9-12 Years

Location: Bangalore (Ecospace)

Candescent Site Reliability Engineering (SRE) mission is to proactively ensure the reliability, availability and performance of our Digital First banking applications. As a member of the SRE team, you will focus on building and operating highly reliable application platforms by applying SRE principles such as automation, observability, resilience and continuous improvement.

You will partner closely with application and platform teams to define reliability standards, implement monitoring, alerting and incident response practices and embed scalability and performance considerations into application design and delivery. Through tooling, automation, and best practices, you will help development teams build and operate services that meet agreed reliability objectives.

As a senior engineer in the organization, you will also provide mentorship within the SRE team and across peer engineering teams, helping elevate operational maturity, drive adoption of SRE practices, and strengthen reliability culture across our core initiatives.

Responsibilities
  • Support and operate production applications running on Kubernetes and AWS
  • Troubleshoot application-level issues using logs, metrics, traces, and runtime signals
  • Participate in incident response, root cause analysis, and post-incident reviews
  • Work closely with development teams to understand application architecture, dependencies, and data flows
  • Improve application observability by defining meaningful alerts, dashboards, and SLOs
  • Automate repetitive operational tasks to reduce toil
  • Support application deployments, rollbacks, and runtime configuration changes
  • Identify reliability, performance, and scalability gaps in application behavior
  • Drive continuous improvements in operational readiness, runbooks, and on-call practices
  • Influence application teams to adopt shift-left reliability practices
Must-Have Skills & Experience
  • Hands-on experience supporting Java applications in production
  • Strong understanding of JVM fundamentals (heap/memory management, garbage collection, OOM issues, thread analysis)
  • Proven experience with SRE practices, including:
    • Incident response and on-call support
    • Root cause analysis and postmortems
    • SLIs, SLOs, and reliability-driven operations
  • Strong experience troubleshooting using application logs, metrics, and monitoring tools
  • Experience operating Java applications on Kubernetes (EKS) from an application/runtime perspective
  • Experience with deployment strategies (rolling, blue/green, canary)
  • Ability to write automation and scripts (Python or any) to reduce operational toil
  • Solid understanding of application architecture and service dependencies (databases, messaging systems, external APIs)
  • Strong collaboration and communication skills; ability to work closely with development teams
  • Demonstrates accountability and sound judgment when responding to high-pressure incidents
Good-to-Have Skills & Experience
  • Exposure to platform or infrastructure concepts supporting application workloads
  • Experience with AWS services such as EKS, RDS/Aurora, S3, EFS, and CloudWatch
  • CI/CD pipeline experience (GitHub Actions, Jenkins)
  • Familiarity with GitOps practices
  • Experience with cloud migrations or modernization efforts
#J-18808-Ljbffr