2

Director Site Reliability Engineer Remote Jobs in California

Site Reliability Engineer

San Francisco, CA · Remote

$67.25 - $89.25/hr

Director, Site Reliability Location: Remote (US) Department: Cloud Platform Engineering / SRE/Reliability Position summary The Site Reliability Engineer (SRE) owns reliability, observability, and ...

Site Reliability Engineer Location: 100% Remote Pay Range: $60-68/hr Must have Networking & Enterprise Infrastructure experience What's the Job? * Develop and maintain automation solutions to enhance ...

New

Senior Site Reliability Engineer

San Diego, CA · Remote

$60.50 - $80.50/hr

This is a remote, contract opportunity for a project Arctiq is delivering for a client. Candidates ... directing the Root Cause Analysis (RCA) process. * Security & Compliance: Lead the integration of ...

Site Reliability Engineer

San Francisco, CA · On-site +1

$155K - $222K/yr

Meet the Team The SRE Fleet team is responsible for maintaining the stability, scalability, and efficiency of the infrastructure that powers our global cloud platform. As a team of six engineers ...

Site Reliability Engineer

Palo Alto, CA · On-site +1

$165K - $190K/yr

About the DevOps / SRE Team The DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable, scalable, and high-performing production systems. We work closely with ...

Site Reliability Engineer

San Mateo, CA · Remote

$65 - $86.25/hr

About the job We\'re looking for a Site Reliability Engineer (E3) to help build and operate the ... This position is remote and requires working East Coast business hours (EST). What you\'ll do

$98K - $138K/yr

The Site Reliability Engineer II will be responsible for supporting, enhancing, and maintaining ... Wellness initiatives #BI-Remote Internal Employees - R365 is committed to growing talent from ...

$98K - $138K/yr

The Site Reliability Engineer II will be responsible for supporting, enhancing, and maintaining ... Wellness initiatives #BI-Remote Internal Employees - R365 is committed to growing talent from ...

Site Reliability Engineer - Networking

San Francisco, CA · On-site +1

$67.25 - $89.25/hr

As a Site Reliability Engineer, you will be focused on supporting a specific, highly available, and very secure production environment. You will analyze the reliability of our environments, use your ...

Senior Site Reliability Engineer

Glendale, CA · On-site +1

$60.50 - $80.25/hr

Our Site Reliability and Infrastructure Engineering team centralizes the concerns of measurement and guidance so every engineer can improve availability and efficiency in their own area of the ...

next page

Showing results 1-20

Director Site Reliability Engineer Remote information

What is the difference between Director Site Reliability Engineer Remote vs Site Reliability Engineer?

AspectDirector Site Reliability Engineer RemoteSite Reliability Engineer
CredentialsTypically requires 8+ years of experience, advanced certifications (e.g., AWS, Google Cloud), leadership skillsUsually requires 3-5 years of experience, relevant certifications, strong technical skills
Work EnvironmentRemote leadership role overseeing teams, strategic planning, cross-team collaborationPrimarily technical role, often remote or on-site, focused on system reliability and automation
Employer & Industry UsageUsed in large tech companies, cloud providers, and enterprises with complex infrastructureCommon in tech, cloud, and SaaS companies focusing on system stability

The main difference is that the Director Site Reliability Engineer Remote focuses on leadership, strategy, and overseeing teams, while the Site Reliability Engineer is more hands-on, technical, and focused on system reliability tasks. Both roles may be remote, but their responsibilities and experience levels differ significantly.

What are the most commonly searched types of Site Reliability Engineer Remote jobs in California? The most popular types of Site Reliability Engineer Remote jobs in California are:
What are popular job titles related to Director Site Reliability Engineer Remote jobs in California? For Director Site Reliability Engineer Remote jobs in California, the most frequently searched job titles are:
What job categories do people searching Director Site Reliability Engineer Remote jobs in California look for? The top searched job categories for Director Site Reliability Engineer Remote jobs in California are:
What cities in California are hiring for Director Site Reliability Engineer Remote jobs? Cities in California with the most Director Site Reliability Engineer Remote job openings:
Infographic showing various Director Site Reliability Engineer Remote job openings in California as of July 2026, with employment types broken down into 1% As Needed, 82% Full Time, 13% Part Time, 1% Temporary, and 3% Contract. Highlights an 92% Physical, 3% Hybrid, and 5% Remote job distribution.

Site Reliability Engineer

STN Inc

San Francisco, CA • Remote

$67.25 - $89.25/hr

Full-time

Re-posted 16 days ago


Job description

Site Reliability Engineer

Platform and software · shared across customers

Reports to: Director, Site Reliability

Location: Remote (US)

Department: Cloud Platform Engineering / SRE/Reliability

Position summary

The Site Reliability Engineer (SRE) owns reliability, observability, and incident response for the GPU One (GPUaaS) platform. The SRE defines and enforces SLOs aligned with contractual SLAs, builds the observability stack, and leads major incidents to resolution.

Key responsibilities
  • Define and operate Service Level Objectives (SLOs) aligned with customer SLAs

  • Build and maintain the observability stack including metrics, logs, traces, and alerting

  • Lead incident response and chair post-incident reviews

  • Drive automation to reduce toil and improve mean-time-to-recover (MTTR)

  • Author and maintain operational runbooks alongside the NOC

  • Manage on-call rotation, escalation paths, and incident-management tooling

  • Coordinate cross-functionally with NOC, Platform Engineering, and Network Engineering

  • Drive chaos engineering, game days, and reliability testing programs

  • Produce SLA performance reports in coordination with the SLA Manager

  • Mentor junior engineers and contribute to engineering culture

Required qualifications
  • 5+ years in SRE, DevOps, or production engineering roles

  • Strong programming skills in Go, Python, or both

  • Hands-on experience operating Kubernetes-based platforms at scale

  • Deep familiarity with observability tooling (Prometheus, Grafana, Datadog, OpenTelemetry)

  • Strong incident management experience including major-incident command

Preferred qualifications
  • GPU or HPC platform operational experience

  • Familiarity with SLA-driven customer environments and credit calculations

  • Experience with chaos engineering tools (Gremlin, Litmus, or similar)

  • Published SRE content or contributions