1

Director Observability Jobs in Arizona (NOW HIRING)

We are seeking a Director of Observability to stand up a brand-new observability and reliability practice from the ground up at Kestra Holdings. This is a working leadership role - the Director will ...

The Director, Site Reliability Engineering plays a critical role in ensuring the continuous and ... Partner with Engineering leadership to design and implement observability tools and processes that ...

Director, DevOps

Mesa, AZ · On-site

$52.25 - $71.75/hr

The Director leads a team of senior DevOps and SRE engineers directly - setting direction ... Own the observability practice for delivery: metrics, logs, traces, and synthetic monitoring that ...

Director, DevOps

Mesa, AZ · Hybrid

$52.25 - $71.75/hr

The Director leads a team of senior DevOps and SRE engineers directly - setting direction ... Own the observability practice for delivery: metrics, logs, traces, and synthetic monitoring that ...

Director - Technologies

Phoenix, AZ · On-site

$144K - $256K/yr

Responsibilities As the Director Software Engineering - AWS Public Cloud (Platform Services), you ... Experience with observability tools such as Prometheus, Splunk, ELK, and Dynatrace. * Strong ...

Director, Site Reliability Engineering

Phoenix, AZ · On-site

$56.50 - $75.25/hr

The Director, Site Reliability Engineering plays a critical role in ensuring the continuous and ... Partner with Engineering leadership to design and implement observability tools and processes that ...

Director, Site Reliability Engineering

Phoenix, AZ · On-site

$56.50 - $75.25/hr

The Director, Site Reliability Engineering plays a critical role in ensuring the continuous and ... Partner with Engineering leadership to design and implement observability tools and processes that ...

Summary The Director, Engineering is responsible for leading the design, development, delivery ... Establish and promote engineering best practices including automation, CI/CD, observability ...

Key Responsibilities • Design and manage AWS network architectures (VPCs, Transit Gateway, Direct ... observability tools (CloudWatch, Datadog, Splunk, OpenTelemetry) for network monitoring and ...

Key Responsibilities • Design and manage AWS network architectures (VPCs, Transit Gateway, Direct ... observability tools (CloudWatch, Datadog, Splunk, OpenTelemetry) for network monitoring and ...

Key Responsibilities • Design and manage AWS network architectures (VPCs, Transit Gateway, Direct ... observability tools (CloudWatch, Datadog, Splunk, OpenTelemetry) for network monitoring and ...

They are seeking a Solution Lead / Director to shape and win complex services opportunities and ... SLAs, observability, lifecycle management, and continuous optimization. • Commercial fluency in ...

next page

Showing results 1-20

Director Observability information

What is the difference between Director Observability vs Site Reliability Engineer?

AspectDirector ObservabilitySite Reliability Engineer
Primary FocusOversees observability strategies, tools, and teams to ensure system visibility and performanceBuilds and maintains reliable systems, automates deployment, and manages incident response
CredentialsTypically requires advanced knowledge of monitoring, cloud platforms, and leadership experienceOften has software engineering background, with skills in scripting, automation, and systems engineering
Work EnvironmentLeads teams in tech companies, focusing on monitoring and analytics toolsWorks closely with development and operations teams to ensure system reliability

While both roles focus on system performance and reliability, the Director Observability primarily manages observability strategies and teams, whereas the Site Reliability Engineer is hands-on, building and maintaining reliable systems. The roles complement each other in ensuring optimal system performance and uptime.

What are the key skills and qualifications needed to thrive as a Director of Observability, and why are they important?

To thrive as a Director of Observability, you need deep expertise in monitoring, logging, and distributed systems, typically backed by a degree in computer science or a related field and extensive experience in IT or DevOps leadership roles. Proficiency with observability tools such as Prometheus, Grafana, Datadog, Splunk, and APM solutions, along with knowledge of cloud platforms and relevant certifications, is essential. Strong leadership, strategic thinking, and communication skills help drive cross-functional initiatives and foster a culture of reliability. These skills and qualities are crucial for ensuring system health, rapid incident response, and alignment between technical teams and organizational objectives.

How does a Director of Observability typically collaborate with engineering and operations teams to drive organizational goals?

A Director of Observability works closely with engineering and operations teams to ensure that systems are monitored effectively and issues are identified and resolved quickly. This collaboration often involves developing unified monitoring strategies, aligning observability tools and processes, and facilitating incident response post-mortems. The Director also leads cross-functional meetings to establish best practices, set key performance indicators (KPIs), and ensure observability is integrated into the software development lifecycle. By acting as a bridge between technical teams, they help foster a culture of transparency, reliability, and continuous improvement.

What does a Director of Observability do?

A Director of Observability leads the strategy and implementation of monitoring, logging, and tracing systems to ensure the health and performance of technical infrastructure. They work with engineering and operations teams to develop best practices, select appropriate tools, and set standards for observability across the organization. Their goal is to provide visibility into system behavior, quickly identify and resolve incidents, and support continuous improvement in system reliability and performance.
What are the most commonly searched types of Observability jobs in Arizona? The most popular types of Observability jobs in Arizona are:
What are popular job titles related to Director Observability jobs in Arizona? For Director Observability jobs in Arizona, the most frequently searched job titles are:
What cities in Arizona are hiring for Director Observability jobs? Cities in Arizona with the most Director Observability job openings:
Director of Observability

Director of Observability

Kestra Holdings

Tempe, AZ • On-site

Full-time

Medical, Retirement

Posted 12 days ago


Job description

Kestra Holdings offers industry-leading wealth management platforms for independent wealth management professionals nationwide. Kestra is dedicated to empowering independent financial professionals-including traditional and hybrid RIAs-to grow their businesses and deliver exceptional client service. We combine advanced business management technology with personalized consulting to provide unmatched scale, efficiency, and support. Our advisor-focused culture is built on innovation and advocacy, enabling advisors to offer comprehensive securities and investment advisory solutions to their clients.
Lead with Purpose. Partner with Impact.
We are seeking a Director of Observability to stand up a brand-new observability and reliability practice from the ground up at Kestra Holdings. This is a working leadership role - the Director will be expected to be hands-on in the early phase: selecting and configuring tooling, writing instrumentation standards, building the first dashboards and alerting pipelines, and personally running incident command for major events while the team and platform mature.
This is a newly created leadership role reporting directly to the Head of IT Infrastructure & Cybersecurity. The Director will start with two direct reports - a to be hired Senior Observability Architect (India-based) and a future US-based Observability/Reliability Engineer - and will be expected to scale the team over time as the practice and service catalog grow.
What you'll Do:
  • Observability Strategy & Platform Hands-On Build-Out.
  • Define and execute the observability strategy for Kestra Holdings, aligned with business objectives, regulatory requirements, and the enterprise technology roadmap.
  • Personally lead the initial build-out of the observability platform across metrics, logs, traces, profiles, and alerting - including tool evaluation, POCs, architecture, deployment, and configuration (e.g., Azure Monitor/Log Analytics, Datadog, Grafana, OpenTelemetry, Elastic/Splunk).
  • Work with and enforce existing instrumentation standards (OpenTelemetry, structured logging, distributed tracing) across infrastructure and application teams.
  • Build the first generation of dashboards, SLO scorecards, and a single pane of glass for Tier-1 service health - rolling up sleeves alongside the Sr. Architect and engineer.
  • Operate the firm's end-to-end incident management lifecycle - detection, response, escalation, communication, and blameless post-incident review.
  • Stand up on-call schedules, escalation policies, and runbook-driven triage for Sev1-Sev4 incidents via PagerDuty/xMatters or equivalent.
  • Serve as primary incident commander for major incidents during the initial build phase, transitioning command responsibilities to senior team members as the practice matures.
  • Integrate the incident lifecycle with Jira / Jira Service Management (JSM) for ticketing, change correlation, and remediation tracking; partner with Cybersecurity so incidents run once, not separately by Infra and Cyber
  • Facilitate post-incident reviews (PIRs), track remediation items in Jira, and report trends to leadership.
  • Drive adoption of SRE principles across the firm: SLI/SLO definition, error budget policy and enforcement, toil identification and automation, and operational readiness reviews.
  • Establish release of reliability gates and embed reliability into the service lifecycle from design through production.
  • Partner with Cloud & Platform Engineering, Cybersecurity, and application teams to ensure all services are fully instrumented, measurable, and integrated into the firm's SLO and incident frameworks.
  • Ensure observability and incident management practices align with NIST CSF 2.0 maturity targets and support the firm's cybersecurity roadmap.
  • Partner with the Cybersecurity team to integrate observability data with Jira/JSM, CMDB, and SIEM for enriched context during incidents.
  • Support regulatory and audit requirements appropriate for a SEC-regulated financial services firm (e.g., logging retention, evidentiary integrity, access controls on telemetry data).
  • Directly lead and mentor a small, high-leverage team of two to start: a Senior Observability Architect (India-based) and a US-based Observability/Reliability Engineer.
  • Operate as a player-coach - splitting time between strategic leadership, hands-on engineering, and direct mentorship of the senior architect.
  • Build a multi-year workforce plan and talent pipeline to scale the team as the platform, service catalog, and 24/7 coverage needs grow.
  • Foster a culture of blameless learning, operational excellence, and engineering-led reliability across US and India hours of coverage.
  • Serve as the primary technical liaison for observability and incident management vendors (e.g., Datadog, PagerDuty/xMatters, Grafana Labs, Elastic/Splunk, Atlassian).
  • Represent Observability & Reliability in the Architecture Review Board, IT Change Management Board, and incident command forums.
  • Provide regular reporting to SVP and executive leadership on reliability KPIs (MTTD, MTTA, MTTR, SLO compliance, alert signal-to-noise), incident trends, and strategic initiatives.

What You Bring:
  • 10+ years in observability, SRE, platform engineering, or infrastructure operations roles, with 3+ years in a people leadership capacity (Director or Sr. Manager level).
  • Demonstrated experience building an observability or SRE practice from scratch - tool selection, instrumentation rollout, first SLOs, and standing up incident command.
  • Strong hands-on technical depths must be willing and able to write code/IaC, configure platforms, build dashboards, and run incidents personally, not just delegate.
  • Deep expertise across the observability stack: metrics (Prometheus, Datadog, Azure Monitor), log aggregation (Elastic/OpenSearch, Log Analytics, Splunk), distributed tracing (OpenTelemetry, Jaeger, Datadog APM), and profiling.
  • Proven experience defining SLIs/SLOs, error budgets, and toil reduction programs.
  • Hands-on experience with incident management platforms (PagerDuty, xMatters) and Jira / Jira Service Management integration for ticketing and workflow.
  • Experience leading distributed teams across US and India time zones.
  • Experience operating in a regulated industry (financial services, healthcare, or similar) with familiarity with compliance frameworks (NIST CSF, SOC 2, SEC, FINRA).
  • Excellent communication skills - able to present reliability posture, incident retrospectives, and risk to executive and board-level audiences.
  • Experience with IaC (Terraform, Bicep, ARM), CI/CD pipelines, and embedding observability-as-code into modern DevOps practices.

Internal Application Policy:
Internal applicants must be in good standing and have a minimum of 1 year of service with Kestra. Internal applicants must also have a minimum of 1 year service in current role unless approved by EVP.
Benefits to support you:
  • Competitive pay and benefits with a large employer (over 1600 employees nationwide)
  • 401(k), health insurance, and a competitive benefits package
  • Work in a supportive, collaborative environment committed to professional excellence
  • Help clients navigate meaningful financial decisions with confidence
  • Opportunities for training, development, and long-term growth within the firm
  • Tuition reimbursement for qualified expenses

Kestra Values:
Our Mission is Powering Financial Independence, enabling the growth and success of investing clients and the advisors who serve them. We do that by living our values: Serve, Make it Happen, and One team.
Explore Life at Kestra
Kestra Holdings Website: https://www.kestrafinancial.com/
Careers Portal: https://jobs.dayforcehcm.com/en-US/kestra/KESTRACAREERSITE
LinkedIn: https://www.linkedin.com/company/kestra-financial
Apply Today
Lead with purpose. Apply now and help shape the future of Kestra.
DisclosureBy applying to a job at Kestra Financial, Inc., you are agreeing to the following statements:
  • You acknowledge that if hired, Kestra Financial, Inc. may, obtain and use background information concerning your credit, character, general reputation, personal characteristics, work habits, performance and experience for evaluation for your potential employment.
  • It is the policy of Kestra Financial to ensure equal employment opportunity without discrimination or harassment on the basis of race, color, religion, sex, sexual orientation, gender, identity or expression, age, disability, marital status, citizenship, national origin, genetic information, or any other characteristic protected by law. Kestra Financial prohibits any such discrimination or harassment.