1

Site Reliability Engineer Jobs in California (NOW HIRING)

Site Reliability Engineer

San Francisco, CA · Remote

$67.25 - $89.25/hr

Site Reliability Engineer Platform and software · shared across customers Reports to: Director, Site Reliability Location: Remote (US) Department: Cloud Platform Engineering / SRE/Reliability ...

As our first or early SRE hire you will be building the reliability engineering function at Forward - defining how we think about availability, observability, incident response, and operational ...

Site Reliability Engineer (SRE)

San Diego, CA · On-site

$60.50 - $80.50/hr

We are looking for the right Site Reliability Engineer to help us take our efforts to the next level. In this role, you will help lead our cloud based infrastructure team for Apple's Video Computer ...

Site Reliability Engineer

Santa Clara, CA · On-site

$230K - $250K/yr

As our first or early SRE hire you will be building the reliability engineering function at Forward - defining how we think about availability, observability, incident response, and operational ...

The Role As a Site Reliability Engineer (SRE) at Air Apps, you will be responsible for ensuring the reliability, availability, and scalability of our systems. You will work at the intersection of ...

Site Reliability Engineer (SRE)

San Francisco, CA · On-site

$67.25 - $89.25/hr

Methodic is seeking a Site Reliability Engineer (SRE) to focus on the stability and efficiency of their platform in production. The role involves applying software engineering principles to ensure ...

As a main contributor to our SRE team you will develop and maintain infrastructure, tooling, and engineering services for cloud based applications. You will be responsible for system bringup ...

Site Reliability Engineer (SRE)

Palo Alto, CA · On-site

$67 - $89.25/hr

The Site Reliability Engineer (SRE) will contribute to the stability and performance of Mithril's global GPU orchestration platform, building automation and observability tools to ensure reliable ...

Site Reliability Engineer (SRE)

San Francisco, CA · On-site

$67.25 - $89.25/hr

The Site Reliability Engineer (SRE) will contribute to the stability and performance of Mithril's global GPU orchestration platform, building automation and observability tools to ensure reliable ...

Site Reliability Engineer

Palo Alto, CA · On-site

$67 - $89.25/hr

They are seeking a Site Reliability Engineer to support and maintain the service quality of their SaaS security platform while addressing challenges around scalability and reliability.

Site Reliability Engineer

Newport Beach, CA · On-site

$61.25 - $81.50/hr

They are seeking a Site Reliability Engineer to support and maintain the service quality of their customer-facing SaaS security platform, address complex challenges, and collaborate closely with ...

Site Reliability Engineer

San Francisco, CA · On-site

$67.25 - $89.25/hr

TextNow is the largest provider of free phone service in the nation, and they are seeking a motivated Site Reliability Engineer to own infrastructure and reliability. This role involves designing ...

Site Reliability Engineer

Mountain View, CA · Hybrid

$189K - $232K/yr

As a Site Reliability Engineer, you will strengthen infrastructure, optimize tooling, deepen observability, streamline incident response, and elevate reliability standards. These actions empower ...

Site Reliability Engineer

Irvine, CA · On-site

$60.75 - $80.75/hr

Title: Site Reliability Engineer - Product Support Analyst Location: Irvine, CA (Hybrid) Note: This is a W2 contract position - C2C, 1099, & 3rd party candidates WILL NOT be considered The Site ...

Java SRE Engineer Onsite San Francisco Bay Area Infrastructure Engineer (2 Positions) We are looking for an experienced Java SRE / Platform Engineer to support large-scale cloud migrations and ...

next page

Showing results 1-20

Site Reliability Engineer information

See California salary details

$10

$62

$90

How much do site reliability engineer jobs pay per hour?

As of Jul 27, 2026, the average hourly pay for site reliability engineer in California is $62.91, according to ZipRecruiter salary data. Most workers in this role earn between $54.09 and $71.88 per hour, depending on experience, location, and employer.

Will SRE be replaced by ai?

Site Reliability Engineers (SREs) focus on maintaining system reliability, automation, and incident response. While AI tools can assist with monitoring and automating routine tasks, SREs' expertise in system design, troubleshooting, and decision-making remains essential, making complete replacement unlikely in the near future.

What Is a Site Reliability Engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What are the key skills and qualifications needed to thrive as a Site Reliability Engineer, and why are they important?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What engineers make $500,000 a year?

Senior-level engineers in high-demand fields such as software engineering, especially those specializing in cloud infrastructure, distributed systems, or machine learning, can earn $500,000 or more annually. These roles often require extensive experience, advanced skills, and sometimes stock options or bonuses in addition to base salary.

Is SRE a stressful job?

Site Reliability Engineers (SREs) often work in high-pressure environments to ensure system stability and uptime, which can lead to stressful situations during outages or incidents. The role requires strong problem-solving skills, familiarity with monitoring tools, and sometimes on-call responsibilities, but organizations often implement practices to manage workload and reduce stress.

What are some of the most common challenges Site Reliability Engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What does a site reliability engineer do?

A site reliability engineer (SRE) is responsible for maintaining and improving the reliability, availability, and performance of software systems. They use automation, monitoring tools, and scripting to prevent outages, troubleshoot issues, and ensure systems run smoothly at scale. SREs often collaborate with development teams and may hold certifications in cloud platforms or scripting languages.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

What is a Site Reliability Engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.
What are the most commonly searched types of Site Reliability Engineer jobs in California? The most popular types of Site Reliability Engineer jobs in California are:
What job categories do people searching Site Reliability Engineer jobs in California look for? The top searched job categories for Site Reliability Engineer jobs in California are:
What cities in California are hiring for Site Reliability Engineer jobs? Cities in California with the most Site Reliability Engineer job openings:
What are popular job titles related to Site Reliability Engineer jobs in CA? For Site Reliability Engineer jobs in CA, the most frequently searched job titles are:
Infographic showing various Site Reliability Engineer job openings in California as of July 2026, with employment types broken down into 93% Full Time, 4% Part Time, and 3% Contract. Highlights an 89% Physical, 4% Hybrid, and 7% Remote job distribution, with an average salary of $130,847 per year, or $62.9 per hour.
Site Reliability Engineer

Site Reliability Engineer

STN Inc

San Francisco, CA • Remote

$67.25 - $89.25/hr

Full-time

Posted 10 days ago


Job description

Site Reliability Engineer

Platform and software · shared across customers

Reports to: Director, Site Reliability

Location: Remote (US)

Department: Cloud Platform Engineering / SRE/Reliability

Position summary

The Site Reliability Engineer (SRE) owns reliability, observability, and incident response for the GPU One (GPUaaS) platform. The SRE defines and enforces SLOs aligned with contractual SLAs, builds the observability stack, and leads major incidents to resolution.

Key responsibilities
  • Define and operate Service Level Objectives (SLOs) aligned with customer SLAs

  • Build and maintain the observability stack including metrics, logs, traces, and alerting

  • Lead incident response and chair post-incident reviews

  • Drive automation to reduce toil and improve mean-time-to-recover (MTTR)

  • Author and maintain operational runbooks alongside the NOC

  • Manage on-call rotation, escalation paths, and incident-management tooling

  • Coordinate cross-functionally with NOC, Platform Engineering, and Network Engineering

  • Drive chaos engineering, game days, and reliability testing programs

  • Produce SLA performance reports in coordination with the SLA Manager

  • Mentor junior engineers and contribute to engineering culture

Required qualifications
  • 5+ years in SRE, DevOps, or production engineering roles

  • Strong programming skills in Go, Python, or both

  • Hands-on experience operating Kubernetes-based platforms at scale

  • Deep familiarity with observability tooling (Prometheus, Grafana, Datadog, OpenTelemetry)

  • Strong incident management experience including major-incident command

Preferred qualifications
  • GPU or HPC platform operational experience

  • Familiarity with SLA-driven customer environments and credit calculations

  • Experience with chaos engineering tools (Gremlin, Litmus, or similar)

  • Published SRE content or contributions