1

Site Reliability Engineer Jobs in California (NOW HIRING)

Site Reliability Engineer

San Francisco, CA · Remote

$67.25 - $89.25/hr

Site Reliability Engineer Platform and software · shared across customers Reports to: Director, Site Reliability Location: Remote (US) Department: Cloud Platform Engineering / SRE/Reliability ...

Site Reliability Engineer (SRE)

San Diego, CA · On-site

$60.50 - $80.50/hr

We are looking for the right Site Reliability Engineer to help us take our efforts to the next level. In this role, you will help lead our cloud based infrastructure team for Apple's Video Computer ...

As our first or early SRE hire you will be building the reliability engineering function at Forward - defining how we think about availability, observability, incident response, and operational ...

Site Reliability Engineer

Santa Clara, CA · On-site

$230K - $250K/yr

As our first or early SRE hire you will be building the reliability engineering function at Forward - defining how we think about availability, observability, incident response, and operational ...

Site Reliability Engineer (SRE)

San Francisco, CA · On-site

$67.25 - $89.25/hr

Methodic is seeking a Site Reliability Engineer (SRE) to focus on the stability and efficiency of their platform in production. The role involves applying software engineering principles to ensure ...

Site Reliability Engineer (SRE)

San Diego, CA

$142K - $263K/yr

  • Medical

  • Dental

  • Retirement

As a main contributor to our SRE team you will develop and maintain infrastructure, tooling, and engineering services for cloud based applications. You will be responsible for system bringup ...

Site Reliability Engineer (SRE)

Palo Alto, CA · On-site

$67 - $89.25/hr

The Site Reliability Engineer (SRE) will contribute to the stability and performance of Mithril's global GPU orchestration platform, building automation and observability tools to ensure reliable ...

Site Reliability Engineer (SRE)

San Francisco, CA · On-site

$67.25 - $89.25/hr

The Site Reliability Engineer (SRE) will contribute to the stability and performance of Mithril's global GPU orchestration platform, building automation and observability tools to ensure reliable ...

Site Reliability Engineer

Palo Alto, CA · On-site

$67 - $89.25/hr

They are seeking a Site Reliability Engineer to support and maintain the service quality of their SaaS security platform while addressing challenges around scalability and reliability.

Site Reliability Engineer (SRE)

San Diego, CA · On-site

$142.30 - $263.30/hr

  • Medical

  • Dental

  • Retirement

As a main contributor to our SRE team you will develop and maintain infrastructure, tooling, and engineering services for cloud based applications. You will be responsible for system bringup ...

Site Reliability Engineer

Newport Beach, CA · On-site

$61.25 - $81.50/hr

They are seeking a Site Reliability Engineer to support and maintain the service quality of their customer-facing SaaS security platform, address complex challenges, and collaborate closely with ...

Site Reliability Engineer

San Francisco, CA · On-site

$67.25 - $89.25/hr

TextNow is the largest provider of free phone service in the nation, and they are seeking a motivated Site Reliability Engineer to own infrastructure and reliability. This role involves designing ...

As a Site Reliability Engineer, you will strengthen infrastructure, optimize tooling, deepen observability, streamline incident response, and elevate reliability standards. These actions empower ...

SRE Engineer

San Jose, CA · On-site

$66.75 - $88.75/hr

Job Title: SRE Engineer Location: San Jose, CA / RTP, NC(Onsite) Job Type: Full Time Must Have Technical/Functional Skills: * SRE, NetApp Storage, Linux Certified, Kubernetes Certified, DevOps, ...

Java SRE Engineer Onsite San Francisco Bay Area Infrastructure Engineer (2 Positions) We are looking for an experienced Java SRE / Platform Engineer to support large-scale cloud migrations and ...

Site Reliability Engineer

Mountain View, CA · Hybrid

$67.25 - $89.25/hr

Site Reliability Engineer Onsite- Bay Area, CA Skills Relevant Skills and Experience What You'll Do (Day-to-Day) * Own and manage our cloud infrastructure (GCP or AWS, on-prem). * Build, maintain ...

next page

Showing results 1-20

Site Reliability Engineer information

See California salary details

$10

$62

$90

How much do site reliability engineer jobs pay per hour?

As of Aug 17, 2026, the average hourly pay for site reliability engineer in California is $62.91, according to ZipRecruiter salary data. Most workers in this role earn between $54.09 and $71.88 per hour, depending on experience, location, and employer.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often involves on-call duties, troubleshooting complex issues, and working with automation tools, which can contribute to work-related stress but also offers opportunities for skill development and problem-solving.

What are the most commonly searched types of Site Reliability Engineer jobs in California?

The most popular types of Site Reliability Engineer jobs in California are:

What are popular job titles related to Site Reliability Engineer jobs in California?

For Site Reliability Engineer jobs in California, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer jobs in California look for?

The top searched job categories for Site Reliability Engineer jobs in California are:

What cities in California are hiring for Site Reliability Engineer jobs?

Cities in California with the most Site Reliability Engineer job openings:

What are popular job titles related to Site Reliability Engineer jobs in CA?

For Site Reliability Engineer jobs in CA, the most frequently searched job titles are:

Infographic showing various Site Reliability Engineer job openings in California as of August 2026, with employment types broken down into 1% As Needed, 82% Full Time, 13% Part Time, 2% Temporary, and 2% Contract. Highlights an 94% Physical, 3% Hybrid, and 3% Remote job distribution, with an average salary of $130,847 per year, or $62.9 per hour.

Site Reliability Engineer

STN Inc

San Francisco, CA • Remote

$67.25 - $89.25/hr

Full-time

This job post has expired today. Applications are no longer accepted.


Job description

Site Reliability Engineer

Platform and software · shared across customers

Reports to: Director, Site Reliability

Location: Remote (US)

Department: Cloud Platform Engineering / SRE/Reliability

Position summary

The Site Reliability Engineer (SRE) owns reliability, observability, and incident response for the GPU One (GPUaaS) platform. The SRE defines and enforces SLOs aligned with contractual SLAs, builds the observability stack, and leads major incidents to resolution.

Key responsibilities
  • Define and operate Service Level Objectives (SLOs) aligned with customer SLAs

  • Build and maintain the observability stack including metrics, logs, traces, and alerting

  • Lead incident response and chair post-incident reviews

  • Drive automation to reduce toil and improve mean-time-to-recover (MTTR)

  • Author and maintain operational runbooks alongside the NOC

  • Manage on-call rotation, escalation paths, and incident-management tooling

  • Coordinate cross-functionally with NOC, Platform Engineering, and Network Engineering

  • Drive chaos engineering, game days, and reliability testing programs

  • Produce SLA performance reports in coordination with the SLA Manager

  • Mentor junior engineers and contribute to engineering culture

Required qualifications
  • 5+ years in SRE, DevOps, or production engineering roles

  • Strong programming skills in Go, Python, or both

  • Hands-on experience operating Kubernetes-based platforms at scale

  • Deep familiarity with observability tooling (Prometheus, Grafana, Datadog, OpenTelemetry)

  • Strong incident management experience including major-incident command

Preferred qualifications
  • GPU or HPC platform operational experience

  • Familiarity with SLA-driven customer environments and credit calculations

  • Experience with chaos engineering tools (Gremlin, Litmus, or similar)

  • Published SRE content or contributions