2

Remote Observability Engineer Jobs in Atlanta, GA

Senior Agentic (AI) Engineer

Atlanta, GA ยท On-site +1

$100K - $138K/yr

Evals & Observability: LangSmith / Langfuse / Braintrust-style tooling, DataDog Requirements * 5+ ... All Remote Hires will be required to travel to Orlando, Florida at least twice per year for Town ...

Full Stack Web Engineer

Atlanta, GA ยท Remote

$110K - $150K/yr

Remote | Full-time About Sched Sched powers thousands of events worldwide professional development ... Contribute to monitoring, alerting, and observability across the platform * Help maintain and ...

Full Stack Web Engineer

Atlanta, GA ยท On-site +1

$110K - $150K/yr

Remote | Full-time About Sched Sched powers thousands of events worldwide professional development ... Contribute to monitoring, alerting, and observability across the platform * Help maintain and ...

Principal Engineer

Atlanta, GA ยท Remote

$50K - $60K/yr

Evaluate, recommend, and implement architectural improvements to enhance scalability, observability ... System Engineering & Optimization * Build and maintain distributed systems using Spring Boot ...

Senior Site Reliability Engineer (SRE)

Atlanta, GA ยท On-site +1

$120K - $175K/yr

Create and support observability/monitoring tools and vendor integrations. * Drive the growth of a ... S. and are willing to consider remote candidates. #LI-Remote Working at PrizePicks: The typical ...

Principal Backend Engineer

Atlanta, GA ยท Remote

$50K - $60K/yr

Evaluate, recommend, and implement architectural improvements to enhance scalability, observability ... System Engineering & Optimization * Build and maintain distributed systems using Spring Boot ...

Senior AI Engineer (Remote)

Atlanta, GA ยท On-site +1

$99K - $136K/yr

The Senior AI Engineer is responsible for designing, building, scaling, and optimizing production ... end observability (logging, metrics, distributed tracing, and automated alerting) across models ...

Backend Engineer

Atlanta, GA ยท Remote

$90/hr

Clerkie is a remote-first company, with over 40 employees spanning 4 time zones across the United ... Improved system observability and production reliability * Supported high-priority product ...

Senior Director Technology

Atlanta, GA ยท Remote

$240K/yr

SENIOR DIRECTOR - DEVOPS | REMOTE | US Travel obsessed? Big tech fan? Hey, you're in good company ... Establish scalable standards for release management, observability, testing automation, deployment ...

Infrastructure Engineer

Atlanta, GA ยท On-site +1

$150K - $195K/yr

Collaborate with engineering teams to improve system observability and incident response ... Remote flexibility: Work from anywhere in the U.S., or join our collaborative HQ team in Atlanta ...

Full-Stack Engineer

Atlanta, GA ยท Remote

$90/hr

Clerkie is a remote-first company, with over 40 employees spanning 4 time zones across the United ... Improved system observability and production reliability * Supported high-priority product ...

Senior Cloud Devops Engineer

Atlanta, GA ยท On-site +1

$125K - $160K/yr

Implement and maintain observability stacks including New Relic, CloudWatch, and Elastic ... Strong Terraform skills with production-grade IaC, modules, and remote state management. * Hands-on ...

Atlanta / Remote What You'll Do * Design, build, and maintain infrastructure-as-code (Terraform ... Datadog or equivalent observability platform experience * Container hardening / security scanning ...

RAG, context engineering, structured outputs, and evaluation/observability for LLM systems. Solid ... Must be authorized to work in the United States; role is US-remote. Nice to Have Experience with ...

Machine Learning Platform Engineer

Atlanta, GA ยท On-site +1

$155K - $185K/yr

You will implement automated retraining pipelines and observability for ML systems to ensure data ... S. and are willing to consider remote candidates. #LI-Remote Working at PrizePicks: The typical ...

Staff Machine Learning Engineer

Atlanta, GA ยท On-site +1

$220K - $280K/yr

You will implement automated retraining pipelines and observability tools to ensure data drift and ... S. and are willing to consider remote candidates. #LI-Remote Working at PrizePicks: The typical ...

Data Engineer, Senior

Atlanta, GA ยท On-site +1

$77K - $176K/yr

Remote Work: Hybrid Job Number: R0247682 Location: Atlanta,GA,US Share job via: Share Data Engineer ... Experience implementing data quality, observability, monitoring, data security, and enterprise data ...

Showing results 41-60

Remote Observability Engineer information

See Atlanta, GA salary details

$36.5K

$111.4K

$184.2K

How much do remote observability engineer jobs pay per year?

As of Sep 6, 2026, the average yearly pay for remote observability engineer in Atlanta, GA is $111,422.00, according to ZipRecruiter salary data. Most workers in this role earn between $79,800.00 and $145,700.00 per year, depending on experience, location, and employer.

What is a remote observability engineer?

A Remote Observability Engineer is a professional responsible for designing, implementing, and maintaining systems that monitor the health, performance, and reliability of software applications and infrastructure from a remote location. They use observability tools to collect and analyze logs, metrics, and traces, helping organizations quickly detect and resolve issues. Their work ensures that distributed systems are transparent, reliable, and efficient, often collaborating with development, operations, and security teams. Remote Observability Engineers often work from anywhere, leveraging cloud-based tools and platforms to manage complex IT environments.

What are the typical collaboration patterns for a remote observability engineer working with distributed teams?

Remote Observability Engineers frequently collaborate with software developers, DevOps teams, and IT operations to ensure systems are monitored effectively and issues are detected early. Working remotely, you'll often use communication tools like Slack, Jira, and video conferencing to coordinate incident response, discuss monitoring strategies, and review system health dashboards. Regular sync meetings and asynchronous updates are common, and you'll likely contribute to documentation and knowledge sharing to keep all stakeholders informed. Building strong communication habits is important, as much of the troubleshooting and improvement work hinges on clear coordination with multiple teams.

What are the key skills and qualifications needed to thrive as a remote observability engineer, and why are they important?

To thrive as a Remote Observability Engineer, you need strong expertise in monitoring, logging, and tracing systems, along with a background in computer science or related technical fields. Familiarity with tools like Prometheus, Grafana, ELK Stack, Datadog, and cloud platforms is typically required, as well as relevant certifications such as AWS Certified Cloud Practitioner or Google Cloud Professional DevOps Engineer. Excellent problem-solving abilities, communication skills, and a proactive mindset help you detect and resolve issues before they impact users. These competencies ensure system reliability, enable rapid incident response, and support seamless collaboration in distributed environments.

What is the difference between Remote Observability Engineer vs Site Reliability Engineer?

AspectRemote Observability EngineerSite Reliability Engineer
CredentialsKnowledge of monitoring tools, scripting, cloud platformsSame as Observability Engineer, plus SRE certifications often preferred
Work EnvironmentFocus on monitoring, logging, and tracing systems remotelyBroader scope including system reliability, incident response, and automation
Industry UsagePrimarily in tech, SaaS, cloud servicesWidely in tech, finance, and large-scale online services

The Remote Observability Engineer specializes in monitoring and analyzing system performance remotely, focusing on tools like logs and metrics. In contrast, the Site Reliability Engineer has a broader role, ensuring overall system reliability, automation, and incident management. While both roles require similar technical skills, SREs often have additional responsibilities related to system resilience and scalability.

What are the most commonly searched types of Observability Engineer jobs in Atlanta, GA?

The most popular types of Observability Engineer jobs in Atlanta, GA are:

What are popular job titles related to Remote Observability Engineer jobs in Atlanta, GA?

For Remote Observability Engineer jobs in Atlanta, GA, the most frequently searched job titles are:

What job categories do people searching Remote Observability Engineer jobs in Atlanta, GA look for?

The top searched job categories for Remote Observability Engineer jobs in Atlanta, GA are:

What cities near Atlanta, GA are hiring for Remote Observability Engineer jobs?

Cities near Atlanta, GA with the most Remote Observability Engineer job openings:

Infographic showing various Remote Observability Engineer job openings in Atlanta, GA as of August 2026, with employment types broken down into 90% Full Time, 5% Part Time, and 5% Contract. Highlights an 87% Physical, 5% Hybrid, and 8% Remote job distribution, with an average salary of $111,422 per year, or $53.6 per hour.

Senior Agentic (AI) Engineer

Worth AI

Atlanta, GA โ€ข On-site, Remote

$100K - $138K/yr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Re-posted 23 days ago


Job description

Worth AI is hiring a Senior Agentic AI Engineer to design and ship production agent systems that automate KYB, underwriting, and risk decisions on regulated financial data. You'll own agents end-to-end architecture, retrieval, tools, evals, and production deployment and partner closely with our Chief AI Officer, applied scientists, and platform teams.

Responsibilities
  • Design and ship multi-step agentic systems (planner/executor, tool-using, multi-agent, human-in-the-loop) for onboarding, underwriting, case review, and continuous monitoring.
  • Architect agent graphs in LangGraph (or comparable - CrewAI, AutoGen, Claude Agent SDK) with explicit state, durable execution, retries, and safe fallbacks.
  • Build the retrieval layer powering our agents - chunking, hybrid search, reranking, and grounded citation.
  • Own the eval stack: golden sets, offline regression suites, LLM-as-judge, online A/B and shadow evals, and red-teaming for jailbreaks, prompt injection, and PII leakage.
  • Expose agents to production systems via well-typed tools and MCP servers. Treat tool surface area as a product.
  • Drive production MLOps: deployment, versioning, traffic shaping, cost/latency budgets, tracing, and on-call playbooks for agent incidents.
  • Partner with security and compliance to keep agents inside SOC 2, GDPR, CCPA, and fair-lending posture - auditability and explainability built in, not bolted on.
  • Mentor engineers on agent patterns, prompt hygiene, eval discipline, and LLM failure modes.
  • Technology Stack
    • Languages: Python, Node.js, TypeScript
    • Agent / LLM frameworks: LangGraph, LangChain, Claude Agent SDK, MCP, OpenAI SDK
    • Models: Anthropic Claude, OpenAI, open-weight where appropriate
    • Retrieval & Data: PostgreSQL, pgvector, OpenSearch, Kafka, Redshift, Redis
    • Infra: AWS, Kubernetes (EKS), ArgoCD, Terraform
    • Evals & Observability: LangSmith / Langfuse / Braintrust-style tooling, DataDog

Requirements

  • 5+ years of software engineering experience, with 2+ years building production LLM or agentic systems (not just notebooks or demos).
  • Hands-on experience with a modern agent framework (LangGraph strongly preferred) and a track record of shipping agents that run, fail gracefully, and recover.
  • Strong RAG fundamentals chunking, embeddings, hybrid retrieval, reranking, grounding - and judgment about when RAG isn't the right answer.
  • Real eval experience golden sets, offline and online evaluations, used to make ship/no-ship calls.
  • Production MLOps fluency: deployed LLM workloads under real latency, cost, and reliability constraints.
  • Strong Python; comfortable in TypeScript / Node.js.
  • Solid systems engineering instincts APIs, async patterns, queues, databases, distributed system failure modes.
  • Calibrated communicator; thrives in ambiguous, fast-moving environments.
  • Prior experience in fintech, lending, payments, KYB/KYC, fraud, or AML.
  • Experience building MCP servers or other structured tool interfaces for LLMs.
  • Background in classical ML (ranking, scoring, calibration).
  • Experience designing explainable / auditable AI workflows for regulated environments.
  • Open-source contributions to agent frameworks, eval tooling, or retrieval libraries.
  • AWS depth (EKS, MSK, RDS, S3, Lambda) and IaC with Terraform.
Success Metrics
  • Agent Quality: Measurable improvements in task success rate, grounding accuracy, and hallucination rate on our eval suites.
  • Production Reliability: Agents you own meet defined SLOs for latency (P90/P99), tool-call success, and cost per task.
  • Velocity: New agent capabilities go from prototype to production in weeks, without skipping evals or guardrails.
  • Risk Posture: Zero material incidents tied to prompt injection, PII leakage, or unsafe tool use on agents you own.
  • Force Multiplier: Patterns, tools, and eval scaffolding you build get adopted across engineering.

All Remote Hires will be required to travel to Orlando, Florida at least twice per year for Town Halls and team collaboration, in addition to orientation in Orlando.

Benefits

  • Health Care Plan (Medical, Dental & Vision)
  • Retirement Plan (401k, IRA)
  • Life Insurance
  • Flexible Paid Time Off
  • 9 paid Holidays
  • Family Leave
  • Remote
  • Hybrid work (for Orlando Associates)
  • Free Food & Snacks (Orlando)
  • Wellness Resources