1

Sre Intern Jobs in California (NOW HIRING)

Site Reliability Engineer

San Francisco, CA · On-site +1

$150K - $250K/yr

Site Reliability Engineer role USC or GC only are considered at this time. San Francisco - Local to Bay area only but role is remote and occasion meeting required Latest update, 03/31/2026: The Site ...

Site Reliability Engineer

San Francisco, CA · On-site +1

$150K - $250K/yr

Site Reliability Engineer role USC or GC only are considered at this time. San Francisco - Local to Bay area only but role is remote and occasion meeting required Latest update, 03/31/2026: The Site ...

SRE

San Jose, CA · On-site

$66.75 - $88.75/hr

Application Production Support (SRE - Site Reliability Engineering) with 3+ years - Preferably in ecommerce domain * Hands on experience in any of the UI Frameworks(AngularJS, VueJS etc) - 1+ years

Site Reliability Engineer

San Francisco, CA · On-site

$67.25 - $89.25/hr

TextNow is the largest provider of free phone service in the nation, and they are seeking a motivated Site Reliability Engineer to own infrastructure and reliability. This role involves designing ...

SRE Engineer

Santa Clara, CA · On-site

$67 - $89/hr

Omega Solutions, Inc. is seeking an SRE Engineer to join their team. The role involves operating infrastructure on AWS, developing self-service tooling for product engineering teams, and ensuring ...

Site Reliability Engineer

Mountain View, CA · Hybrid

$67.25 - $89.25/hr

Site Reliability Engineer Onsite- Bay Area, CA Skills Relevant Skills and Experience What You'll Do (Day-to-Day) * Own and manage our cloud infrastructure (GCP or AWS, on-prem). * Build, maintain ...

SRE Engineer

San Jose, CA · On-site

$66.75 - $88.75/hr

Job Title: SRE Engineer Location: San Jose, CA / RTP, NC(Onsite) Job Type: Full Time Must Have Technical/Functional Skills: * SRE, NetApp Storage, Linux Certified, Kubernetes Certified, DevOps, ...

Java SRE Engineer Onsite San Francisco Bay Area Infrastructure Engineer (2 Positions) We are looking for an experienced Java SRE / Platform Engineer to support large-scale cloud migrations and ...

Site Reliability Engineer Intern 2027

San Jose, CA · On-site

$65.25 - $86.75/hr

Job Title Site Reliability Engineer Intern 2027 Date posted 11-Aug-2026 Job ID 128513 City / Township / Village LOWELL, DURHAM, San Jose, Austin State / Province Texas, North Carolina, Massachusetts ...

We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster ...

Java SRE Engineer

Santa Clara, CA · On-site

$120 - $160/hr

Java SRE Engineer Onsite San Francisco Bay Area Infrastructure Engineer (2 Positions) We are looking for an experienced Java SRE / Platform Engineer to support large-scale cloud migrations and ...

Showing results 21-40

Sre Intern information

What does an SRE intern do?

An SRE (Site Reliability Engineering) Intern assists with maintaining and improving the reliability, scalability, and performance of software systems. They typically work on automating operational tasks, monitoring system health, responding to incidents, and helping implement best practices for deployment and infrastructure. SRE Interns collaborate with engineering teams to ensure services are robust and can recover quickly from failures, often writing scripts and using tools to optimize processes. This role is a great opportunity to gain hands-on experience with cloud platforms, automation, and large-scale system operations.

What are the key skills and qualifications needed to thrive as an SRE intern, and why are they important?

To thrive as an SRE Intern, you need a solid understanding of computer science fundamentals, basic scripting or programming skills (such as Python or Bash), and familiarity with Linux systems. Exposure to monitoring tools (like Prometheus or Grafana), cloud platforms, and version control systems (e.g., Git) is often expected. Strong problem-solving abilities, attention to detail, and effective communication skills make an intern stand out. These skills are crucial for maintaining system reliability, quickly addressing incidents, and collaborating with engineering teams to ensure high system availability.

What are some typical challenges SRE interns face when learning to manage large-scale systems?

SRE Interns often encounter the challenge of quickly adapting to complex, large-scale infrastructure and understanding the interplay between reliability, scalability, and performance. Learning to troubleshoot incidents, interpret monitoring dashboards, and participate in on-call rotations can be overwhelming at first. However, most teams provide mentorship, documentation, and structured onboarding to help interns ramp up successfully. Collaboration with experienced engineers is key, and asking questions is encouraged to build confidence and expertise.

What is the difference between Sre Intern vs DevOps Intern?

AspectSre InternDevOps Intern
Required CredentialsBasic knowledge of Linux, scripting, cloud platformsSimilar skills, often with some scripting and cloud familiarity
Work EnvironmentFocus on system reliability, monitoring, and automationFocus on deployment, CI/CD pipelines, and infrastructure
Employer & Industry UsageTech companies, cloud providers, large enterprisesTech startups, software companies, cloud services

Both Sre Intern and DevOps Intern roles involve working with cloud platforms, scripting, and automation. However, Sre Interns typically focus more on system reliability, monitoring, and incident response, while DevOps Interns concentrate on deployment pipelines, infrastructure automation, and continuous integration. The roles often overlap but serve different core functions within IT and software development teams.

What are the most commonly searched types of Sre jobs in California?

The most popular types of Sre jobs in California are:

What cities in California are hiring for Sre Intern jobs?

Cities in California with the most Sre Intern job openings:

Infographic showing various Sre Intern job openings in California as of August 2026, with employment types broken down into 13% Internship, 1% As Needed, 62% Full Time, 21% Part Time, 1% Temporary, and 2% Contract. Highlights an 88% Physical, 3% Hybrid, and 9% Remote job distribution.

Site Reliability Engineer

3B Staffing LLC

San Francisco, CA • On-site, Remote

$150K - $250K/yr

Full-time

This job post has expired today. Applications are no longer accepted.


Job description

Site Reliability Engineer role

USC or GC only are considered at this time.

San Francisco - Local to Bay area only but role is remote and occasion meeting required

Latest update, 03/31/2026:

The Site Reliability Engineer role is critical for us right now - we have enterprise customers with urgent reliability issues that need immediate attention. We're excited to see candidates who can jump in and own this piece of our infrastructure!

What our team says about this role

Client is looking for 2 SREs. From a financial perspective, the business has hit $7M+ in ARR, are meaningfully profitable, and are growing exponentially.

They're looking for 2 SREs with strong programming expertise and experience with large-scale systems to own reliability and performance for enterprise customers including Nvidia, Samsara, Zapier and PwC.

Avoid candidates that are too CI/CD or DevOps focused - they need people with genuine debugging experience in production environments

The role has a base salary of $150K - $250K + equity and they have a preference for on-site in San Francisco but the search is also open to remote for strong candidates that are not based in the SF Area."

"We are looking for a Site Reliability Engineer with strong programming expertise and experience with large-scale systems to own the reliability and performance for our enterprise customers including Nvidia, Samsara, Zapier and PwC.

You will work closely with our Co-founders and have a massive impact on our product and customer satisfaction."

Tech stack

Python, C, Rust, Kubernetes, FastAPI, Redis, Postgres, Prisma

Seniority

4-8+ years of experience in production or reliability engineering, with a focus on debugging and fixing system-level issues.

Work experience

  • Experience debugging memory leaks in a production environment.
  • Experience as a production engineer or reliability engineer with direct experience fixing issues rather than only reporting them.
  • Experience working with large scale systems (at least 1k+ RPS).

Hard skills

  • Strong programming ability in C and Rust
  • Experience with PostgreSQL, Redis, Kubernetes or Prometheus/Grafana

Soft skills

  • Excited to work at an early-stage startup and willing to work ~60 hours / week.

What you will do:

  • Work directly with enterprise customers to debug and resolve production issues.
  • Own the reliability and performance, with a focus on debugging memory leaks, connection pool issues, and other critical bugs.
  • Proactively improve the overall reliability of the system to prevent future issues.
  • Profile systems, run benchmarks, and work to improve latency and throughput.
  • Collaborate in a fast-paced startup environment.

Role Details

  • Title: Site Reliability Engineer
  • Core responsibilities include owning product reliability and performance, debugging memory leaks, and working closely with enterprise customers.
  • Reports to the co-founder and collaborate with the entire 5-person company.

Candidate Requirements

  • Must have experience with large-scale systems, ideally from big tech companies like Meta, Amazon, Microsoft, etc.
  • Strong debugging skills, particularly with memory leaks, are essential.
  • Programming proficiency required; C, Rust, and Python are required.
  • Looking for candidates with 4+ years of experience, but open to more senior candidates if they fit the role.

Company Context

  • Client is a profitable AI company with a $7 million ARR, used by companies like Netflix, NASA, and Nvidia.
  • The company is a small, dynamic team focusing on open-source AI gateways.

Compensation and Logistics

  • Salary range set at $200,000 to $250,000.
  • Remote work is possible, but Bay Area candidates preferred for occasional on-site work.

Timeline and Urgency

  • Hiring is urgent due to customer issues and potential churn.
  • Interview process includes recruiter screen, 30-minute call, technical round, and on-site session.

Pain Points

  • Current customer issues need immediate attention to prevent churn.
  • Lack of dedicated personnel focusing solely on product reliability.

Ideal Candidate Profile

  • Preferred from big tech companies with experience in high-traffic environments.
  • Hands-on, eager to work in a startup environment, willing to handle long hours and high-pressure situations.