2

Remote Observability Engineer Jobs in Maryland (NOW HIRING)

Site Reliability Engineer

Columbia, MD ยท On-site +1

$55.50 - $73.75/hr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Hybrid Columbia MD 3 times per week OR Remote (as applicable to role) Work Authorization ... This role is responsible for implementing observability and automation practices, supporting ...

QA Platform Engineer

Bethesda, MD ยท On-site +1

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

This opportunity is full time and onsite/remote at the NCBI in Bethesda, MD and/or remote. NCBI is ... Develops and continuously improves DevSecOps, DataOps, and Observability platforms. * Develops and ...

Senior Mobile Engineer

Baltimore, MD ยท Remote

$140K - $150K/yr

This role is not eligible for relocation, visa sponsorship, or fully remote work.*** What You'll ... Supporting experimentation, feature flagging, observability, and release management. * Creating ...

Senior DevOps Engineer (US REMOTE)

Beltsville, MD ยท Remote

$140K - $170K/yr

  • Medical

  • Dental

  • Retirement

... s Full-Stack Engineer with expertise in IaC (Terraform), Helm, MySQL, Kubernetes, and CI/CD ... Familiarity with observability tooling (Prometheus/Grafana, ELK Stack or similar) * Knowledge of ...

Lead Agentic AI Engineer - Remote US

Lanham, MD ยท Remote

$165K - $175K/yr

  • Medical

  • PTO

Create evaluation and observability systems so we can measure agent performance, catch failures ... engineering, with meaningful time building AI/ML or agentic systems * Location: Remote (US)

AI Software Engineer

Baltimore, MD ยท Remote

$100K - $135K/yr

This is a fully remote, full-time position offering the opportunity to work on emerging AI ... Familiarity with prompt engineering, AI evaluation, and model observability. * Experience with ...

Sr. Data Analytics Engineer

Baltimore, MD ยท On-site +1

$125K - $165K/yr

  • Medical

  • Retirement

Drive engineering best practices include testing, observability, modular modeling, and ... Location: Remote (East Coast strongly preferred to optimize collaboration with HQ and cross ...

Sr. Data Analytics Engineer

Baltimore, MD ยท On-site +1

$125K - $165K/yr

  • Medical

  • Retirement

Drive engineering best practices include testing, observability, modular modeling, and ... Location: Remote (East Coast strongly preferred to optimize collaboration with HQ and cross ...

Platform DevEx Engineer

Bethesda, MD ยท On-site +1

$56.50 - $77.25/hr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

This opportunity is full time and onsite/remote at the NCBI in Bethesda, MD and/or remote. NCBI is ... Familiarity with observability tools like Prometheus, the EFK (ElasticSearch, fluentd, Kibana) or ...

Platform Engineer (DataOps)

Bethesda, MD ยท On-site +1

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

This opportunity is full time and onsite/remote at the NCBI in Bethesda, MD and/or remote. NCBI is ... Modern observability and logging tools: Prometheus, EFK (ElasticSearch, fluentd, Kibana), TIGK ...

AI Platform Engineer, Senior

Laurel, MD ยท On-site +1

$86K - $198K/yr

  • Medical

  • Life

  • Retirement

  • PTO

Remote Work: No Job Number: R0234072 Location: Laurel,MD,US Share job via: Share AI Platform ... Experience implementing and scaling observability solutions, including APM, OpenTelemetry, Grafana ...

Passionate about educating customers on observability risks that are meaningful to their business ... Remote

Senior Platform Engineer (DataOps)

Bethesda, MD ยท On-site +1

$111K - $153K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

This opportunity is full time and onsite/remote at the NCBI in Bethesda, MD and/or remote. NCBI is ... Modern observability and logging tools: Prometheus, EFK (ElasticSearch, fluentd, Kibana), TIGK ...

next page

Showing results 1-20

Remote Observability Engineer information

What is a remote observability engineer?

A Remote Observability Engineer is a professional responsible for designing, implementing, and maintaining systems that monitor the health, performance, and reliability of software applications and infrastructure from a remote location. They use observability tools to collect and analyze logs, metrics, and traces, helping organizations quickly detect and resolve issues. Their work ensures that distributed systems are transparent, reliable, and efficient, often collaborating with development, operations, and security teams. Remote Observability Engineers often work from anywhere, leveraging cloud-based tools and platforms to manage complex IT environments.

What are the typical collaboration patterns for a remote observability engineer working with distributed teams?

Remote Observability Engineers frequently collaborate with software developers, DevOps teams, and IT operations to ensure systems are monitored effectively and issues are detected early. Working remotely, you'll often use communication tools like Slack, Jira, and video conferencing to coordinate incident response, discuss monitoring strategies, and review system health dashboards. Regular sync meetings and asynchronous updates are common, and you'll likely contribute to documentation and knowledge sharing to keep all stakeholders informed. Building strong communication habits is important, as much of the troubleshooting and improvement work hinges on clear coordination with multiple teams.

What are the key skills and qualifications needed to thrive as a remote observability engineer, and why are they important?

To thrive as a Remote Observability Engineer, you need strong expertise in monitoring, logging, and tracing systems, along with a background in computer science or related technical fields. Familiarity with tools like Prometheus, Grafana, ELK Stack, Datadog, and cloud platforms is typically required, as well as relevant certifications such as AWS Certified Cloud Practitioner or Google Cloud Professional DevOps Engineer. Excellent problem-solving abilities, communication skills, and a proactive mindset help you detect and resolve issues before they impact users. These competencies ensure system reliability, enable rapid incident response, and support seamless collaboration in distributed environments.

What is the difference between Remote Observability Engineer vs Site Reliability Engineer?

AspectRemote Observability EngineerSite Reliability Engineer
CredentialsKnowledge of monitoring tools, scripting, cloud platformsSame as Observability Engineer, plus SRE certifications often preferred
Work EnvironmentFocus on monitoring, logging, and tracing systems remotelyBroader scope including system reliability, incident response, and automation
Industry UsagePrimarily in tech, SaaS, cloud servicesWidely in tech, finance, and large-scale online services

The Remote Observability Engineer specializes in monitoring and analyzing system performance remotely, focusing on tools like logs and metrics. In contrast, the Site Reliability Engineer has a broader role, ensuring overall system reliability, automation, and incident management. While both roles require similar technical skills, SREs often have additional responsibilities related to system resilience and scalability.

What are the most commonly searched types of Observability Engineer jobs in Maryland?

The most popular types of Observability Engineer jobs in Maryland are:

What are popular job titles related to Remote Observability Engineer jobs in Maryland?

For Remote Observability Engineer jobs in Maryland, the most frequently searched job titles are:

What job categories do people searching Remote Observability Engineer jobs in Maryland look for?

The top searched job categories for Remote Observability Engineer jobs in Maryland are:

What cities in Maryland are hiring for Remote Observability Engineer jobs?

Cities in Maryland with the most Remote Observability Engineer job openings:

Site Reliability Engineer

Cogent People Inc

Columbia, MD โ€ข On-site, Remote

$55.50 - $73.75/hr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Re-posted 11 days ago


Job description

Description

Employment Type: Full-time, W2 position with Cogent People Inc. This is a direct hire position with full benefits.ย 


Location: Hybrid Columbia MD 3 times per week OR Remote (as applicable to role)ย 


Work Authorization Requirementsย 


To comply with government contracting requirements, candidates must meet all of the following:ย 

  • Must be a U.S. Citizen, Permanent Resident, or valid EAD holderย ย 
  • Must have lived in the United States for at least 3 of the past 5 yearsย ย 
  • Must be currently authorized to work in the U.S. without sponsorshipย ย 

Sponsorship (H-1B) is not available for this position (now or in the future).ย 


Candidates who do not meet these requirements will not be considered.ย 


Clearance Requirementย 


Public Trust required or ability to obtain, depending on assignment.ย 


About Cogent People Inc.ย 


Cogent People Inc. is a government consulting and technology services firm supporting mission-critical federal and commercial programs. We deliver secure, scalable, and modern digital solutions across complex IT environments.ย 

Our teams thrive at the intersection of engineering excellence and mission impact, building systems that matter.ย 


Job Overview ย 

Cogent People Inc. is seeking a Site Reliability to support system reliability, monitoring, and operational stability across environments.


This role is responsible for implementing observability and automation practices, supporting production systems, and ensuring system performance and availability. The position plays a key role in incident response, root cause analysis, and ongoing system optimization in collaboration with DevOps and development teams.


The ideal candidate will bring experience in system monitoring, DevOps practices, and production support, along with the ability to collaborate across cross-functional engineering teams in a fast-paced environment.

This position may be contingent upon contract award.

Requirements

What You'll Doย 


System Reliability & Observabilityย 

  • Support system reliability, monitoring, and operational stability across environments ย 
  • Implement and maintain observability practices, including monitoring, logging, and alerting ย 
  • Contribute to automation efforts that improve system reliability and operational efficiency ย 

Incident Response & Performance Optimizationย 

  • Participate in incident response activities and production support ย 
  • Perform root cause analysis for system issues and outages ย 
  • Support performance optimization and tuning of applications and infrastructure ย 

DevOps & Collaborationย 

  • Work with DevOps and development teams to maintain production readiness ย 
  • Contribute to continuous improvement of deployment and operational processes ย 
  • Collaborate across engineering teams to support stable and scalable systems ย 

What We're Looking Forย 

  • Bachelor's degree in Computer Science, Information Systems, or a related field, or an equivalent combination of education and experience ย 
  • Experience in system reliability, DevOps, or production support roles ย 
  • Experience with monitoring, logging, and observability tools ย 
  • Understanding of incident management and root cause analysis processes ย 
  • Familiarity with cloud environments and infrastructure concepts ย 
  • Experience supporting automated deployment or operational workflows ย 
  • Strong problem-solving and troubleshooting skills ย 
  • Excellent written and verbal communication skills ย 
  • Ability to work effectively in fast-paced, production-critical environments ย 
  • Strong collaboration skills across development and operations teams ย 

What Will Set You Apartย 

  • Experience with AWS or other cloud platforms ย 
  • Familiarity with infrastructure-as-code tools (e.g., Terraform or similar) ย 
  • Experience with tools such as Splunk, Datadog, Prometheus, or similar observability platforms ย 
  • Experience with CI/CD pipelines and DevOps automation tools ย 
  • Prior experience supporting enterprise-scale or regulated environments ย 
  • Knowledge of application performance tuning and distributed systems behaviorย 

Why Cogent People Inc.?ย 


At Cogent People, we combine technical excellence with a mission-driven culture. Our teams work on meaningful, high-impact projects that support government and enterprise transformation initiatives.ย 


We offer:ย 

  • Competitive compensationย ย 
  • Career growth and professional development opportunitiesย ย 
  • Exposure to complex, mission-critical systemsย ย 
  • A collaborative and supportive team environmentย ย 
  • Long-term client engagements with stability and continuityย ย 

We are a Certified Great Place to Work, committed to building an inclusive and high-performance culture.ย 


Benefitsย 

  • Medical, Dental, and Vision Insurance (comprehensive coverage)ย ย 
  • 401(k) with company matchย ย 
  • Company-paid life insuranceย ย 
  • Short-term and long-term disability coverageย ย 
  • Paid Time Off: 3 weeks annually + 10 paid holidaysย ย 
  • Employee assistance and wellness resources (as applicable)ย ย 

Compliance Noticeย 


Cogent People Inc. conducts employment verification for all candidates. Misrepresentation of work authorization, residency history, or professional experience will result in disqualification.ย 


We are an Equal Opportunity Employer (EEO) and evaluate all applicants based on qualifications, experience, and role requirements.ย 


We do not engage third-party recruiters for this role unless explicitly stated.