1

Director Observability Jobs in Oregon (NOW HIRING)

Senior AI Platform Engineer, Infrastructure Services

OR ยท On-site +1

$108K - $147K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

Hands-on experience with API gateway technologies (Kong, Envoy, Apigee, or similar); direct experience with AI/LLM gateway patterns (rate limiting, semantic caching, prompt/response observability) is ...

Staff AI Platform Engineer, Infrastructure Services

OR ยท On-site +1

$107K - $140K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

Hands-on experience with API gateway technologies (Kong, Envoy, Apigee, or similar); direct experience with AI/LLM gateway patterns (rate limiting, semantic caching, prompt/response observability) is ...

Sr. Data Governance Lead

Lake Oswego, OR ยท On-site

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

... Monte Carlo observability, and dbt. * Exceptional cross-functional leadership skills with a direct, concise, and metric-driven communication style. Bonus Points For: * Hands-on experience ...

Sr. Data Governance Lead

Lake Oswego, OR

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

... Monte Carlo observability, and dbt. * Exceptional cross-functional leadership skills with a direct, concise, and metric-driven communication style. Bonus Points For: * Hands-on experience ...

Staff AI Platform Engineer

OR ยท On-site +1

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

... observability, or audit logging; direct AI and LLM platform experience is a strong plus but not required. * Practical experience with agent orchestration frameworks (LangGraph or equivalent) and ...

Senior Staff AI Platform Engineer

OR ยท On-site +1

$104K - $143K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

... observability, or audit logging; direct AI and LLM platform experience is a strong plus but not required. * Practical experience with agent orchestration frameworks (LangGraph or equivalent) and ...

Senior Software Engineer

OR ยท On-site +1

$165K - $215K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

Prescryptive is the healthcare technology company enabling the direct access marketplace for ... CI/CD, environment management, and observability * Raise the bar through code reviews, design input ...

Experience with cloud platforms, infrastructure automation, CI/CD, automated testing, observability ... Director. You will work closely with Product Managers, Principal Engineers, Architects, Data ...

Network Automation & AI Business Consultant

OR ยท On-site +1

  • Medical

  • Dental

  • Retirement

Advise on integrations with ITSM, monitoring, and observability platforms; stay current on emerging ... direct experience in a large enterprise environment. * Demonstrated ability to lead customer ...

Lead, Product Manager - Platform Services

OR ยท On-site +1

$125K - $165K/yr

  • Medical

  • Retirement

The platform portfolio includes shared commerce services such as APIs, integrations, performance, observability, and foundational capabilities that support our direct-to-consumer digital experience.

Senior Software Engineer, Infrastructure

OR ยท On-site +1

$180K - $250K/yr

  • Medical

  • Dental

  • Vision

  • PTO

We deliver direct customer value through our cloud capabilities and security. We impact internal ... Improve observability (logging, metrics, tracing), and automated remediation to increase ...

Senior Software Engineer (Authentication)

OR ยท On-site +1

$122K - $161K/yr

  • Dental

  • Vision

  • PTO

Use observability and diagnostics tooling (metrics, logging, tracing, dashboards) to troubleshoot ... What we offer As a part of Pindrop, you'll have a direct impact on our growing list of products and ...

OR ยท On-site

$52.75 - $72.25/hr

... direct impact on product velocity and operational excellence. Key Responsibilities * Design ... Implement robust observability across systems using Google Cloud Operations Suite, Cloud Logging ...

Azure DevOps Engineer

OR ยท On-site +1

$52.75 - $72.25/hr

... direct impact on product velocity and operational excellence. Key Responsibilities * Design ... Implement robust observability across systems using Azure Monitor, Log Analytics, Application ...

OR ยท On-site

... and observability tools to make secure-by-default software the path of least resistance for ... Direct experience working within a technology alliances, ISV partnerships, or partner engineering ...

Staff Software Engineer - Data Platform

OR ยท On-site +1

$114K - $137K/yr

  • Medical

  • Dental

  • Vision

Data governance & observability: experience with data governance frameworks, data catalogs, and ... Customer/product focus: an understanding of the direct and indirect business value of your work ...

Showing results 21-40

Director Observability information

What does a director of observability do?

A Director of Observability leads the strategy and implementation of monitoring, logging, and tracing systems to ensure the health and performance of technical infrastructure. They work with engineering and operations teams to develop best practices, select appropriate tools, and set standards for observability across the organization. Their goal is to provide visibility into system behavior, quickly identify and resolve incidents, and support continuous improvement in system reliability and performance.

How does a director of observability typically collaborate with engineering and operations teams to drive organizational goals?

A Director of Observability works closely with engineering and operations teams to ensure that systems are monitored effectively and issues are identified and resolved quickly. This collaboration often involves developing unified monitoring strategies, aligning observability tools and processes, and facilitating incident response post-mortems. The Director also leads cross-functional meetings to establish best practices, set key performance indicators (KPIs), and ensure observability is integrated into the software development lifecycle. By acting as a bridge between technical teams, they help foster a culture of transparency, reliability, and continuous improvement.

What are the key skills and qualifications needed to thrive as a director of observability, and why are they important?

To thrive as a Director of Observability, you need deep expertise in monitoring, logging, and distributed systems, typically backed by a degree in computer science or a related field and extensive experience in IT or DevOps leadership roles. Proficiency with observability tools such as Prometheus, Grafana, Datadog, Splunk, and APM solutions, along with knowledge of cloud platforms and relevant certifications, is essential. Strong leadership, strategic thinking, and communication skills help drive cross-functional initiatives and foster a culture of reliability. These skills and qualities are crucial for ensuring system health, rapid incident response, and alignment between technical teams and organizational objectives.

What is the difference between Director Observability vs Site Reliability Engineer?

AspectDirector ObservabilitySite Reliability Engineer
Primary FocusOversees observability strategies, tools, and teams to ensure system visibility and performanceBuilds and maintains reliable systems, automates deployment, and manages incident response
CredentialsTypically requires advanced knowledge of monitoring, cloud platforms, and leadership experienceOften has software engineering background, with skills in scripting, automation, and systems engineering
Work EnvironmentLeads teams in tech companies, focusing on monitoring and analytics toolsWorks closely with development and operations teams to ensure system reliability

While both roles focus on system performance and reliability, the Director Observability primarily manages observability strategies and teams, whereas the Site Reliability Engineer is hands-on, building and maintaining reliable systems. The roles complement each other in ensuring optimal system performance and uptime.

What are the most commonly searched types of Observability jobs in Oregon?

The most popular types of Observability jobs in Oregon are:

What are popular job titles related to Director Observability jobs in Oregon?

For Director Observability jobs in Oregon, the most frequently searched job titles are:

What job categories do people searching Director Observability jobs in Oregon look for?

The top searched job categories for Director Observability jobs in Oregon are:

What cities in Oregon are hiring for Director Observability jobs?

Cities in Oregon with the most Director Observability job openings:

Infographic showing various Director Observability job openings in Oregon as of June 2026, with employment types broken down into 1% As Needed, 98% Full Time, and 1% Part Time. Highlights an 77% Physical, 5% Hybrid, and 18% Remote job distribution.

Senior AI Platform Engineer, Infrastructure Services

SentinelOne

OR โ€ข On-site, Remote

$108K - $147K/yr

Full-time

Medical, Dental, Vision, Life, Retirement

Posted 12 days ago


Job description

As a Senior AI Platform Engineer, Infrastructure Services, you will be tasked with taking ownership of our AI Gateway infrastructure (built on Kong AI Gateway), the system that authenticates, routes, rate-limits, and monitors AI coding assistant traffic org-wide, while also being fluent enough across our broader platform stack to design solutions that span the two. This is a high-autonomy, high-scope role: you will set technical direction for AI infrastructure, drive incident response and reliability work, and partner closely with the engineers who own our CI/CD, GitOps, and artifact systems rather than working in isolation from them.

What Will You Do?

Primary responsibilities include:

  • Work on the AI Gateway platform: architect, harden, and scale our Kong AI Gateway deployment (Konnect Hybrid on KCP/EKS), including auth (Okta/OIDC), consumer tiers and budgets, rate limiting, semantic caching, and observability.
  • Lead reliability and incident response: drive root-cause analysis and remediation for gateway issues (timeouts, latency, capacity, failover) and build the monitoring/alerting needed to catch them before users do.
  • Design across the platform, not just the gateway: work fluently with our CI/CD (Jenkins, JPAAS), GitOps and Kubernetes deployment tooling (ArgoCD across dev/gov/prod), artifact management (Artifactory/Xray), GitHub Enterprise administration, and GitHub Actions runner fleet, so that AI infrastructure decisions account for how the rest of the platform actually works.
  • Evaluate and roll out AI developer tooling: run structured pilots and adoption efforts for tools like AI-assisted PR review (Qodo) and engineering metrics platforms (LinearB), and make clear build-vs-buy recommendations.
  • Set technical direction and mentor: define architecture and standards for AI infrastructure, review designs across the team, and raise the bar for other engineers working in this space.
  • Partner cross-functionally: work directly with security, DevEx, and product engineering teams consuming the gateway to translate their needs into platform capabilities.
  • Host and serve local models: stand up and operate self-hosted/open-weight model serving infrastructure (e.g. vLLM, NVIDIA Triton/NIM, TGI, Ollama) for workloads where routing to an external provider isn't the right fit, including GPU capacity planning, autoscaling, and cost/performance tuning.
  • Support the broader model lifecycle: help build LLMOps practices such as model versioning, evaluation, and safe rollout, plus supporting infrastructure for retrieval-augmented generation (vector stores, embedding pipelines) as use cases mature.
  • Track usage and cost: build observability into token usage, latency, and spend across both API-based and self-hosted models so the business can see what AI infrastructure actually costs.
What Skills and Knowledge Will You Bring?

Ideal candidates will have:

  • 5 or more years of experience in platform, infrastructure, or DevOps engineering, with a track record of owning systems end-to-end in production.
  • Hands-on experience with API gateway technologies (Kong, Envoy, Apigee, or similar); direct experience with AI/LLM gateway patterns (rate limiting, semantic caching, prompt/response observability) is a strong plus.
  • Strong Kubernetes and GitOps experience (ArgoCD or comparable), and comfort operating across multiple environments (dev, gov, prod).
  • Solid CI/CD background: Jenkins pipeline design and administration, build infrastructure, and runner/agent fleet management (GitHub Actions runners or equivalent).
  • Experience with artifact and package management systems (Artifactory, Xray, or similar) and source control platform administration (GitHub Enterprise).
  • Working knowledge of infrastructure-as-code (Terraform) and cloud platforms (AWS/EKS).
  • Experience deploying and operating self-hosted LLM inference stacks (vLLM, NVIDIA Triton/NIM, TGI, Ollama, or similar) and GPU-backed infrastructure, including Kubernetes GPU scheduling and autoscaling.
  • Familiarity with LLMOps practices: model versioning, evaluation harnesses, and usage/cost observability across API-based and self-hosted models.
  • Track record of setting technical direction, driving cross-team initiatives, and mentoring other engineers; this role has significant scope and minimal day-to-day oversight.
  • Clear, proactive communicator who can explain infrastructure trade-offs to both engineers and non-technical stakeholders.
  • Experience operating LLM/AI-assisted developer tooling at scale (Claude Code, Copilot, or similar) inside an enterprise is preferred.
  • Familiarity with Okta/OIDC and enterprise auth patterns for internal platforms is preferred.
  • Experience with engineering productivity metrics tooling (LinearB or similar) and AI-based code review tooling (Qodo or similar) is preferred.
  • Experience with vector databases and RAG pipelines (e.g. Milvus, Pinecone, pgvector, or similar) in a production setting is preferred.
  • Exposure to model fine-tuning or lightweight training pipelines (LoRA/QLoRA or similar) for domain-specific model adaptation is preferred.
Why SentinelOne?

AI is redefining how the world operates and rewriting the rules of security in real time, and SentinelOne was built for this moment. From day one, we architected an AI-native platform designed to operate at machine speed, not as an add-on to legacy systems but as the foundation itself. If you want to build where innovation and impact move together, this is that place.

We invest in our Sentinels with comprehensive, competitive benefits designed to support you and your family:

Equity & Rewards

  • Restricted Stock Units (RSUs)
  • Employee Stock Purchase Plan (ESPP)

Time Off & Wellbeing

  • Flexible time off
  • Paid company holidays and paid sick time
  • Gender-neutral parental leave
  • Grandparent leave

Insurance & Financial Security

  • Medical, dental, and vision coverage
  • 401(k) retirement plan with company match
  • Life and disability insurance
  • Health and dependent care FSA
  • Voluntary benefits (hospital, accident, critical illness)
  • Employee Assistance Program (EAP)
  • ARAG pre-paid legal
  • Nationwide pet insurance
  • Cancer Care program
  • Global business travel medical insurance

Work Perks & Flexibility

  • Home office allowance
  • Mobile phone reimbursement

Wellness & Lifestyle

  • Wellness coach
  • Wellness/gym reimbursement
  • Fertility coverage
  • Adoption & surrogacy reimbursement