1

Director Observability Jobs (NOW HIRING)

Observability Tech Lead

Somerville, MA · On-site

$160K - $200K/yr

Direct experience developing and distributing Claude Skills, Gemini Gems, or any other generic AI ... Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across ...

Observability Engineer (Splunk) Locations: Jacksonville, FL; Boston, MA; Kanas City, MO | Hybrid 6x ... Communication that is direct, respectful, and documented. This Role May Not Be a Fit If * You ...

Showing results 41-60

Director Observability information

What is the difference between Director Observability vs Site Reliability Engineer?

AspectDirector ObservabilitySite Reliability Engineer
Primary FocusOversees observability strategies, tools, and teams to ensure system visibility and performanceBuilds and maintains reliable systems, automates deployment, and manages incident response
CredentialsTypically requires advanced knowledge of monitoring, cloud platforms, and leadership experienceOften has software engineering background, with skills in scripting, automation, and systems engineering
Work EnvironmentLeads teams in tech companies, focusing on monitoring and analytics toolsWorks closely with development and operations teams to ensure system reliability

While both roles focus on system performance and reliability, the Director Observability primarily manages observability strategies and teams, whereas the Site Reliability Engineer is hands-on, building and maintaining reliable systems. The roles complement each other in ensuring optimal system performance and uptime.

What are the key skills and qualifications needed to thrive as a director of observability, and why are they important?

To thrive as a Director of Observability, you need deep expertise in monitoring, logging, and distributed systems, typically backed by a degree in computer science or a related field and extensive experience in IT or DevOps leadership roles. Proficiency with observability tools such as Prometheus, Grafana, Datadog, Splunk, and APM solutions, along with knowledge of cloud platforms and relevant certifications, is essential. Strong leadership, strategic thinking, and communication skills help drive cross-functional initiatives and foster a culture of reliability. These skills and qualities are crucial for ensuring system health, rapid incident response, and alignment between technical teams and organizational objectives.

How does a director of observability typically collaborate with engineering and operations teams to drive organizational goals?

A Director of Observability works closely with engineering and operations teams to ensure that systems are monitored effectively and issues are identified and resolved quickly. This collaboration often involves developing unified monitoring strategies, aligning observability tools and processes, and facilitating incident response post-mortems. The Director also leads cross-functional meetings to establish best practices, set key performance indicators (KPIs), and ensure observability is integrated into the software development lifecycle. By acting as a bridge between technical teams, they help foster a culture of transparency, reliability, and continuous improvement.

What does a director of observability do?

A Director of Observability leads the strategy and implementation of monitoring, logging, and tracing systems to ensure the health and performance of technical infrastructure. They work with engineering and operations teams to develop best practices, select appropriate tools, and set standards for observability across the organization. Their goal is to provide visibility into system behavior, quickly identify and resolve incidents, and support continuous improvement in system reliability and performance.
More about Director Observability jobs

What cities are hiring for Director Observability jobs?

Cities with the most Director Observability job openings:

What are the most commonly searched types of Observability jobs?

The most popular types of Observability jobs are:

What states have the most Director Observability jobs?

States with the most job openings for Director Observability jobs include:

Infographic showing various Director Observability job openings in the United States as of August 2026, with employment types broken down into 95% Full Time, and 5% Contract. Highlights an 77% In-person, and 23% Remote job distribution.

Staff Software Engineer - Observability

Intuit

San Diego, CA • On-site

Full-time

Posted 25 days ago


Intuit rating

8.2

Company rating: 8.2 out of 10

Based on 92 frontline employees who took The Breakroom Quiz

106th of 244 rated software companies


Job description

Intuit is a leading software provider of business and financial management solutions for small and mid-sized businesses, consumers, financial institutions and accounting professionals. You probably know us by our flagship products, QuickBooks, Quicken and TurboTax, but that's just the start. we're taking on exciting challenges, such as SaaS and mobile applications. Over 50 million users, seven million small businesses and 1,600 financial institutions depend on Intuit because we innovate at the crossroads of real customer problems and breakthrough technology. Join us and let your ingenious ideas be heard.

Interested in creating and leading the platforms that are high scale and mission critical? Want to solve large scale and highly availability platform challenges for on premise and public cloud deployments?  Intuit is seeking Staff Software Engineer, who is characterized by progressive technical experience and has demonstrated progression in technical prowess, to join PDX Observability Engineering team.

The Staff Software Engineer will join the Core PDX Observability Engineering team at Intuit to design and deliver the next-generation logging and observability platform. This role focuses on creating and leading high-scale, mission-critical pipelines and solving large-scale, high-availability challenges across on-premise and public cloud (AWS, GCP) deployments. You will own the architecture and evolution of Intuit's One Logging system - spanning ingestion, routing, cost optimization, and MCP Server-driven automation - building platform capabilities that maximize velocity for thousands of Intuit developers.


Responsibilities

  • Architect, build, and evolve the One Intuit Logging system end-to-end from log generation at the edge through ingestion, routing, and storage.
  • Own and drive pipeline and cost optimization initiatives across the logging stack, reducing ingestion volume and infrastructure spend without losing signal fidelity.
  • Lead design and development of core logging components: Front End Logging Service (FELS), S3 Log Writer, Kinesis/CloudWatch Log Writer, Log Router, Asterias Splunk, GCP Logs Processor, Index Controller, and Asset-to-Log DB.
  • Drive the Automation Revamp/Rewrite initiative, modernizing legacy tooling into scalable, maintainable services.
  • Design and maintain edge/collection agents - Fluent Bit DaemonSet, OIL sidecar (Fluent Bit), EC2 Logger Agent - and integrate Kubernetes metadata enrichment into the pipeline.
  • Build and extend the FELS Onboarding Plugin to streamline developer onboarding to the logging platform.
  • Leverage MCP Server capabilities to enable AI-assisted authoring, automation, and operational tooling across the observability platform.
  • Build observability into the platform itself - Grafana dashboards, metrics, and alerting for pipeline health, throughput, and cost.
  • Partner with platform governance efforts (e.g., SplunkCraft) to enforce ingestion quality, policy, and guardrails upstream in the pipeline.
  • Provide technical leadership and mentorship, setting engineering standards and design direction across the team.
  • Collaborate cross-functionally with SRE, platform, and product engineering teams to align logging platform capabilities with organization-wide needs.

Qualifications

  • 8-10+ years of experience in software engineering, with significant experience designing and operating large-scale distributed systems, logging/data pipelines, or observability platforms
  • Bachelor's degree (BE/BTech/MS/MTech) in Computer Science or related field required;
  • Deep expertise in building and operating high-throughput data pipelines (log/event ingestion, streaming, routing) at multi-TB/PB daily scale.
  • Strong hands-on experience with Kubernetes, containerized workloads, and sidecar/daemonset architectures (e.g., Fluent Bit)
  • Proficiency with public cloud platforms (AWS - S3, Kinesis, CloudWatch, EC2; GCP - logging/monitoring services)
  • Experience with Splunk or equivalent log management/observability platforms at scale
  • Strong programming skills in Go, Java, Python, or similar languages used in infrastructure/platform engineering
  • Demonstrated ability to lead architecture and design for mission-critical, high-availability systems
  • Experience driving cost-optimization initiatives for large-scale infrastructure
  • Strong track record of technical leadership, mentorship, and cross-team influence without direct reporting authority
  • Excellent communication skills - able to translate complex technical tradeoffs for both engineering and leadership audiences

Footer

Intuit provides a competitive compensation package with a strong pay for performance rewards approach. This position may be eligible for a cash bonus, equity rewards and benefits, in accordance with our applicable plans and programs (see more about our compensation and benefits at Intuit: Careers | Benefits). Pay offered is based on factors such as job-related knowledge, skills, experience, and work location. To drive ongoing fair pay for employees, Intuit conducts regular comparisons across categories of ethnicity and gender. The expected base pay range for this position is: 


Southern, CA: $188,500 - $255,000



Employment Type: Full-Time

What Intuit employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom