1

Director Observability Jobs in San Ramon, CA (NOW HIRING)

Senior Director, AI Engineering

South San Francisco, CA ยท On-site

$303K/yr

Own the technical integration between AWS Bedrock AgentCore and our orchestration and observability ... Direct experience shipping multi-agent or RAG-based systems into enterprise B2B environments.

Sr. Director, Strategy & Planning

San Jose, CA ยท On-site

$253K - $324K/yr

Develop the Cisco and Splunk "better together" narrative, connecting networking, security, observability, and data capabilities to high-value customer outcomes. * Establish a portfolio of repeatable ...

Director, Product Management (AI Defense)

San Mateo, CA ยท On-site

$265K - $277K/yr

Meet the Team Cisco's AI Software & Platform Group builds enterprise AI platforms and developer tooling, integrating our networking, security, and observability portfolio with modern AI to power ...

Director, Product Management (AI Defense)

Milpitas, CA ยท On-site

$271K - $284K/yr

Meet the Team Cisco's AI Software & Platform Group builds enterprise AI platforms and developer tooling, integrating our networking, security, and observability portfolio with modern AI to power ...

Director, Product Management (AI Defense)

Milpitas, CA ยท On-site

$271K - $284K/yr

Meet the Team Cisco's AI Software & Platform Group builds enterprise AI platforms and developer tooling, integrating our networking, security, and observability portfolio with modern AI to power ...

Showing results 41-60

Director Observability information

What is the difference between Director Observability vs Site Reliability Engineer?

AspectDirector ObservabilitySite Reliability Engineer
Primary FocusOversees observability strategies, tools, and teams to ensure system visibility and performanceBuilds and maintains reliable systems, automates deployment, and manages incident response
CredentialsTypically requires advanced knowledge of monitoring, cloud platforms, and leadership experienceOften has software engineering background, with skills in scripting, automation, and systems engineering
Work EnvironmentLeads teams in tech companies, focusing on monitoring and analytics toolsWorks closely with development and operations teams to ensure system reliability

While both roles focus on system performance and reliability, the Director Observability primarily manages observability strategies and teams, whereas the Site Reliability Engineer is hands-on, building and maintaining reliable systems. The roles complement each other in ensuring optimal system performance and uptime.

What are the key skills and qualifications needed to thrive as a director of observability, and why are they important?

To thrive as a Director of Observability, you need deep expertise in monitoring, logging, and distributed systems, typically backed by a degree in computer science or a related field and extensive experience in IT or DevOps leadership roles. Proficiency with observability tools such as Prometheus, Grafana, Datadog, Splunk, and APM solutions, along with knowledge of cloud platforms and relevant certifications, is essential. Strong leadership, strategic thinking, and communication skills help drive cross-functional initiatives and foster a culture of reliability. These skills and qualities are crucial for ensuring system health, rapid incident response, and alignment between technical teams and organizational objectives.

How does a director of observability typically collaborate with engineering and operations teams to drive organizational goals?

A Director of Observability works closely with engineering and operations teams to ensure that systems are monitored effectively and issues are identified and resolved quickly. This collaboration often involves developing unified monitoring strategies, aligning observability tools and processes, and facilitating incident response post-mortems. The Director also leads cross-functional meetings to establish best practices, set key performance indicators (KPIs), and ensure observability is integrated into the software development lifecycle. By acting as a bridge between technical teams, they help foster a culture of transparency, reliability, and continuous improvement.

What does a director of observability do?

A Director of Observability leads the strategy and implementation of monitoring, logging, and tracing systems to ensure the health and performance of technical infrastructure. They work with engineering and operations teams to develop best practices, select appropriate tools, and set standards for observability across the organization. Their goal is to provide visibility into system behavior, quickly identify and resolve incidents, and support continuous improvement in system reliability and performance.
What are popular job titles related to Director Observability jobs in San Ramon, CA? For Director Observability jobs in San Ramon, CA, the most frequently searched job titles are:
What job categories do people searching Director Observability jobs in San Ramon, CA look for? The top searched job categories for Director Observability jobs in San Ramon, CA are:
What cities near San Ramon, CA are hiring for Director Observability jobs? Cities near San Ramon, CA with the most Director Observability job openings:
Infographic showing various Director Observability job openings in San Ramon, CA as of June 2026, with employment types broken down into 1% As Needed, 85% Full Time, 12% Part Time, and 2% Contract. Highlights an 78% Physical, 6% Hybrid, and 16% Remote job distribution.

Director of Engineering, Lakehouse Platform

TetraScience

San Francisco, CA โ€ข On-site

$298K/yr

Full-time

Life, Retirement, PTO

Re-posted 17 days ago


Job description

ABOUT TETRASCIENCE

TetraScience is the Scientific Data and AI Cloud purpose-built for biopharma. Our platform - Tetra OS - harmonizes instrument, lab, and enterprise data into a unified scientific lakehouse, enabling AI-driven drug discovery and development at the world's leading pharmaceutical companies. We are mission-driven and technically ambitious, united by the belief that better data infrastructure accelerates science.

THE ROLE

TetraScience is hiring a Director of Engineering to lead the Lakehouse Data Products Platform team. The Lakehouse Platform is the foundational data infrastructure and platform layer that powers all scientific data analysis and AI products on Tetra OS.

You will own the architecture, drive technical and operational strategy aligned with product and customer growth. As an AI-forward, hands-on leader, you with develop internal tools and systems alongside your team. As a deeply technical domain expert, you with lead, manage and grow a team 8+ engineers spanning Data Platforms and Infrastructure engineering.

This is a player-coach role. You will be deeply fluent in the modern data and AI stack and build internal tools, lead architecture and code reviews. At TetraScience, everyone builds. Engineers, managers, and leadership alike. A pure management background without active technical contribution will not succeed here.

WHAT YOU WILL DO

Architecture and Technical Strategy

  • Own and evolve the Lakehouse Platform and Infrastructure foundations:
    • Storage layer, catalog, Spark query engine, Databricks IaC, Platform Data Pipelines and releases.
    • Lakehouse Developer Platform and DX that supports Tetraflows, Semantic services and Data Resources built on the Lakehouse Platform
    • Lead the evolution of TetraScience's Lakehouse architecture from first principles across open table formats, partition strategies, schema evolution, governance, and API contracts.
  • Design for durability: version coupling, artifact deployment safety, and protocol compatibility across major platform releases.
  • Collaborate with other platform and infrastructure teams to define integration contracts, data pipelines, shared execution roadmaps and operational excellence standards.

Engineering Execution

  • Own technical prioritization and delivery across a team of 8+ engineers; drive sprint-level execution and quarterly delivery commitments.
  • Lead incident response and root-cause analysis for production issues; build systemic fixes, not one-off patches.
  • Make and defend build-vs-buy decisions for Lakehouse components based on strategic fit and engineering cost.
  • Establish and maintain engineering standards: testing practices, observability instrumentation, and release safety specific to data infrastructure.

People and Team Development

  • Coach engineers at all levels - technical mentorship, growth plans, and direct performance feedback delivered consistently.
  • Hire and develop senior ICs and tech leads; build team depth to reduce knowledge concentration risk.
  • Model the builder culture: write code, ship internal tools, and set the bar for technical craft on your team.

AI-Forward Development

  • Champion AI-assisted development practices across the team - your own workflow should demonstrate what good looks like.
  • Identify opportunities to apply AI to data quality, schema inference, anomaly detection, and platform observability on the Lakehouse layer.
  • Contribute to TetraScience's broader AI platform strategy from the Lakehouse data infrastructure layer up.

Requirements

WHAT YOU WILL BRING

Required

  • 10+ years of engineering experience at top tier technology organizations, with at least 4 years focused on data engineering or distributed systems at production scale.
  • Demonstrable expertise in Lakehouse technologies: Apache Spark, Delta Lake or Apache Iceberg, Databricks or an equivalent distributed compute platform.
  • Experience leading Data Engineering teams (6+ engineers) with direct accountability for strategy, execution, operational excellence and people development.
  • Strong command of distributed systems, cloud and modern data stack: columnar and open table formats, query execution engines, partitioning strategies, metadata catalogs and decoupled storage and compute platforms
  • Track record of building, shipping and operating Tier 1 production data platforms and products.
  • Cloud-native fluency: AWS, GCP, or Azure, containerization, and Infrastructure-as-Code (Terraform or equivalent).
  • Clear written and verbal communication; can present architectural decisions to both technical and non-technical stakeholders.

Strong Advantage

  • Experience in regulated industries (life sciences, healthcare, financial services) with data governance and auditability requirements.
  • Familiarity with scientific data pipelines, analysis or instrument data pipelines.
  • Real-time or near-real-time streaming experience (Kafka, Apache Flink, or equivalent).
  • Exposure to AI/ML workflows over Lakehouse data: feature stores, ML pipelines, or vector search infrastructure.
  • Prior career arc as a senior IC (Staff+ / Principal Engineer equivalent) before moving into engineering leadership.

WHY TETRASCIENCE

  • Mission-driven: your infrastructure directly enables scientific drug development processes that save and extend lives.
  • Builder culture: everyone ships - engineers, managers, and leadership. No passengers.
  • Technically ambitious: we are rebuilding how science-grade data infrastructure works, not deploying off-the-shelf solutions.
  • AI-first: we apply AI across every layer of the platform today, not as a future roadmap item.
  • Competitive compensation, meaningful equity, and strong benefits.

Benefits

  • Competitive compensation
  • Stock options in a VC-backed Series C company
  • 100% employer-paid benefits for all eligible employees and immediate family members
  • 401(k)
  • Unlimited paid time off (PTO)
  • Flexible working arrangements
  • Company-paid Life Insurance, LTD/STD

The salary range for this position is $180,000-$270,000. The salary range posted reflects our target baseline for this role. Final compensation is determined by a thorough evaluation of factors including the candidate's specific experience, localized market data, and internal team equity.