1

Observability Engineer Jobs in Chantilly, VA (NOW HIRING)

Sr. Software Engineer

Bethesda, MD · On-site

$131K - $172K/yr

The ideal candidate has deep technical expertise in the Open-Source Observability, Data platform domain. Position Responsibilities As a Senior Engineer, you will: * Focus on Single or multiple areas ...

Sr. Software Engineer

Bethesda, MD · On-site

$131K - $172K/yr

The ideal candidate has deep technical expertise in the Open-Source Observability, Data platform domain. Position Responsibilities As a Senior Engineer, you will: * Focus on Single or multiple areas ...

Sr. Software Engineer

Chevy Chase, MD · On-site

$130K - $172K/yr

... for the Observability Engineering domain • Accountable for the quality, usability, and performance of the solutions • Be a executor as well as an active learner, helping to coach TDPs and ...

This role is designed for a senior technologist with deep expertise in log/telemetry routing, largescale data engineering, and enterprise-grade observability architectures. You will shape pipeline ...

This role is designed for a senior technologist with deep expertise in log/telemetry routing, largescale data engineering, and enterprise-grade observability architectures. You will shape pipeline ...

AZURE DEVOPS ENGINEER

Vienna, VA · On-site

$53 - $72.50/hr

You will also play a key role in renovating and extending our core data platforms for high observability, reliability, and actionable insights. Key Responsibilities: Developer Experience & Platform ...

This role focuses on automation, observability, and incident response while upholding strict Service Level Objectives (SLOs). The SRE will help build resilient systems that scale, automate manual ...

This role focuses on automation, observability, and incident response while upholding strict Service Level Objectives (SLOs). The SRE will help build resilient systems that scale, automate manual ...

Showing results 41-60

Observability Engineer information

What does an observability engineer do?

An Observability Engineer is responsible for designing, implementing, and maintaining monitoring, logging, and tracing systems to ensure the health, performance, and reliability of applications and infrastructure. They work with tools like Prometheus, Grafana, OpenTelemetry, and ELK to collect and analyze telemetry data. Their goal is to provide visibility into system behavior, detect and diagnose issues quickly, and improve overall system observability. They collaborate with developers, SREs, and operational teams to create automated and scalable observability solutions.

What are the key skills and qualifications needed to thrive as an observability engineer?

To thrive as an Observability Engineer, you need a solid understanding of monitoring, logging, and tracing systems, as well as expertise in programming, cloud infrastructure, and incident response. Proficiency in tools like Prometheus, Grafana, ELK stack, and familiarity with cloud platforms such as AWS, Azure, or GCP are commonly required, and certifications like AWS Certified DevOps Engineer can be advantageous. Strong analytical thinking, collaborative skills, and effective communication are essential soft skills for diagnosing issues and working across development and operations teams. These competencies are vital for proactively maintaining system reliability, ensuring performance, and resolving complications before they impact business operations.

How much do observability engineers make?

Observability engineers typically earn between $90,000 and $150,000 annually, depending on experience, location, and company size. Senior roles or those with expertise in tools like Prometheus, Grafana, or cloud platforms may command higher salaries.
Infographic showing various Observability Engineer job openings in Chantilly, VA as of August 2026, with employment types broken down into 91% Full Time, 5% Part Time, and 4% Contract. Highlights an 87% Physical, 5% Hybrid, and 8% Remote job distribution.

Senior Platform Engineer -- Enterprise API Gateway

BaseCamp Consulting & Solutions

Reston, VA • On-site

$140 - $190/hr

Other

Posted 13 days ago


Job description

POSITION OVERVIEW

The Senior Platform Engineer stands up, automates, and sustains the infrastructure the Customer's enterprise API gateway runs on. The role covers container platform engineering, infrastructure-as-code, deployment automation, and the observability and high-availability work that keeps the gateway serving production traffic through upgrades, patch cycles, and failures. It is the platform counterpart to the security and application seats on the same team. This position requires an active Moderate Background Investigation (MBI) suitability determination at start.

RESPONSIBILITIES
  • Deploy and operate the gateway control and data planes on Kubernetes or Red Hat OpenShift - namespaces, RBAC, quotas, ingress and routes, rolling updates
  • Automate installation, configuration, upgrade, and patching with Helm, Terraform, and Ansible, delivered as reusable modules rather than one-off scripts
  • Build and maintain CI/CD pipelines in Jenkins or GitHub Actions with artifact management, image scanning, and approval gates under Government change control
  • Manage declarative gateway configuration as code and promote it across environment tiers with tested rollback
  • Architect multi-AZ high availability and disaster recovery for the gateway tier and its datastores
  • Provision and maintain supporting cloud infrastructure - VPC design, subnets and routing, security groups, managed databases, backups, encryption
  • Automate certificate lifecycle, mutual TLS, and secrets handling for platform components
  • Build observability across metrics, logs, and traces with dashboards and alerting tied to service objectives
  • Establish performance and capacity baselines, load test the gateway tier, and tune for throughput and latency
  • Maintain runbooks, SOPs, and design artifacts sufficient for Government and follow-on staff to sustain the platform
REQUIRED QUALIFICATIONS
  • Active MBI suitability determination, current and transferable as of your start date
  • U.S. citizenship, as required for Customer contractor personnel
  • Eight years of hands-on platform, DevOps, or cloud infrastructure engineering in production, including three years supporting a Federal system subject to FISMA
  • Production operation of Kubernetes or Red Hat OpenShift, including cluster resources, ingress controllers, operators, and Helm
  • Infrastructure-as-code at scale with Terraform and Ansible, including reusable modules and version-controlled state
  • Ownership of enterprise CI/CD pipelines in Jenkins, GitHub Actions, or GitLab, with artifact repositories and security scanning integrated
  • AWS-native architecture across VPC, EC2 or EKS, IAM, S3, RDS, Lambda, and CloudWatch, or equivalent depth in another major cloud
  • Demonstrated design and operation of multi-AZ, high-availability services meeting a stated availability commitment
  • Observability engineering with tools such as Prometheus, Grafana, Splunk, AppDynamics, OpenTelemetry, or the ELK stack
  • Linux administration on RHEL, shell or Python scripting, and troubleshooting at the OS and network layer
  • Certificate and secrets handling for platform services - TLS and mutual TLS, key rotation, and a secrets management platform
  • Experience promoting configuration through Government-controlled environments under formal change control
  • Technical documentation - design artifacts, runbooks, and SDLC deliverables that hold up under Government review
  • Agile or Scrum delivery alongside Government product owners and multiple contractor teams
  • Bachelor's degree in Computer Science, Engineering, or Information Systems, or equivalent hands-on experience
PREFERRED QUALIFICATIONS
  • An active vendor certification on the program's API gateway platform, or willingness to earn it within 90 days of start with training provided
  • An active platform certification such as CKA, Red Hat OpenShift, AWS Solutions Architect, Terraform Associate, or ITIL v4
  • Prior support to a Federal financial or tax administration program
  • GitOps delivery with Argo CD, and service mesh exposure such as Istio
  • Messaging and integration middleware such as Kafka, IBM MQ, ActiveMQ, or webMethods
  • Zero Trust implementation - mTLS, RBAC/ABAC, MFA, PIV, and LDAP/AD or Kerberos integration
  • Supporting ATO and continuous monitoring from the platform side - scan remediation and evidence collection
#J-18808-Ljbffr