1

Site Reliability Engineer Intern Jobs in Rodeo, CA

Site Reliability Engineer

San Francisco, CA ยท On-site

$67.25 - $89.25/hr

Job Summary : 1872 Consulting is seeking a dynamic Site Reliability Engineer to join their rapidly growing SRE team. The role involves operating a high-performance, scalable platform and working with ...

Site Reliability Engineer

San Francisco, CA

$67.25 - $89.25/hr

The role We're looking for a world-class Site Reliability Engineer to ensure the reliability, performance, and scalability of our AI infrastructure platform. You'll be building and operating the core ...

Site Reliability Engineer

San Francisco, CA ยท On-site

$175K - $250K/yr

The role We're looking for a world-class Site Reliability Engineer to ensure the reliability, performance, and scalability of our AI infrastructure platform. You'll be building and operating the core ...

Senior Site Reliability Engineer

San Francisco, CA ยท On-site

$67.25 - $89.25/hr

About the Role Plenful is hiring a Senior Site Reliability Engineer (SRE) to keep our production systems reliable, performant, and scalable as we grow. This role is centered on operating real systems ...

About the Role We're looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You'll partner ...

Showing results 41-60

Site Reliability Engineer Intern information

See Rodeo, CA salary details

$11

$70

$101

How much do site reliability engineer intern jobs pay per hour?

As of Sep 3, 2026, the average hourly pay for site reliability engineer intern in Rodeo, CA is $70.39, according to ZipRecruiter salary data. Most workers in this role earn between $60.53 and $80.43 per hour, depending on experience, location, and employer.

What is a site reliability engineer intern?

A Site Reliability Engineer (SRE) Intern supports the reliability, scalability, and performance of software systems by assisting with automation, monitoring, and incident response. They work closely with development and operations teams to improve system resilience and efficiency. Typical tasks include writing scripts, analyzing system logs, and contributing to documentation or tooling improvements. This role helps interns gain hands-on experience in infrastructure management, cloud services, and DevOps practices.

What types of projects or tasks does a site reliability engineer intern typically work on?

As a Site Reliability Engineer Intern, you may work on projects like automating operational processes, building monitoring dashboards, troubleshooting incidents, and helping improve infrastructure reliability. Interns often collaborate closely with full-time engineers on tasks such as writing scripts to automate deployments, responding to on-call alerts, or optimizing system performance. You'll gain practical experience using industry-standard tools and practices while learning about uptime, scalability, and reliability workflows. It's a great opportunity to develop hands-on skills and understand how large-scale systems are maintained in a professional environment.

What are the key skills and qualifications needed to thrive as a site reliability engineer intern, and why are they important?

To thrive as a Site Reliability Engineer Intern, you should have a solid understanding of computer science fundamentals, programming (such as Python or Go), and basic networking concepts, often supported by relevant coursework or prior technical internships. Familiarity with cloud platforms (like AWS or GCP), containerization tools (e.g., Docker, Kubernetes), and version control systems (Git) is highly valuable. Strong problem-solving abilities, effective communication, and a collaborative mindset help interns stand out in team settings. These skills are crucial for learning quickly, contributing to the team's reliability initiatives, and ensuring highly available and scalable system operations.

What cities near Rodeo, CA are hiring for Site Reliability Engineer Intern jobs?

Cities near Rodeo, CA with the most Site Reliability Engineer Intern job openings:

Infographic showing various Site Reliability Engineer Intern job openings in Rodeo, CA as of August 2026, with employment types broken down into 1% As Needed, 75% Full Time, 19% Part Time, 2% Temporary, and 3% Contract. Highlights an 94% Physical, 3% Hybrid, and 3% Remote job distribution, with an average salary of $146,406 per year, or $70.4 per hour.

Hiring: Site Reliability Engineer (SRE) - Microsoft Hyper-V & Private Cloud

Realtech Services

San Francisco, CA โ€ข On-site

$67.25 - $89.25/hr

Contractor

Re-posted 3 days ago


Job description


Position: Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud

Location: Jersey City, NJ; Houston, TX; San Francisco, CA (Need Local Candidates only)

Duration: Long-Term Contract

Overview

We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.

The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.

Key Responsibilities

  • Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
  • Optimize, and support highly available VDI environments on Hyper-V.
  • Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
  • Disaster recovery, backup, patch management, and business continuity strategies.
  • Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
  • Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
  • Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
  • Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
  • Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
  • Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
  • Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
  • Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
  • Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
  • Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.

Required Technical Skills

Site Reliability Engineering

  • Strong understanding of Site Reliability Engineering principles and operational excellence.
  • Experience with infrastructure reliability, service availability, resiliency, and performance optimization.
  • Storage Space Direct and failover clustering technical expertise. (Storage Spaces Direct enables you to build highly available, software-defined storage by pooling local disks (SSDs, NVMe drives, and HDDs) across multiple Windows Server nodes in a cluster. Instead of relying on an external SAN, S2D uses the servers' local storage to create a resilient shared storage pool)
  • Experience managing production-critical infrastructure environments with high availability requirements.
  • Experience with incident management, problem management, RCA, and continuous operational improvement.
  • Knowledge of monitoring, observability, alerting, and performance management.

Microsoft Hyper-V (Core Expertise)

  • Deep hands-on expertise in Microsoft Hyper-V architecture, deployment, administration, troubleshooting, and optimization.
  • Extensive experience in operating enterprise private cloud environments on Hyper-V.
  • Strong experience supporting enterprise-scale VDI deployments on Hyper-V.
  • Hyper-V Failover Clustering and high-availability architecture.
  • Storage integration including SAN, NAS, Storage Spaces Direct (S2D), Cluster Shared Volumes (CSV), and storage optimization.
  • Networking within Hyper-V environments including virtual switches, VLANs, NIC Teaming, QoS, and network performance tuning.
  • System Center Virtual Machine Manager (SCVMM).

Automation & Platform Engineering

  • Strong PowerShell scripting and automation experience.
  • Experience automating infrastructure deployment, operational tasks, health checks, and reporting.
  • Familiarity with Infrastructure as Code concepts and configuration management.
  • Experience developing reusable operational tooling to improve reliability and reduce manual effort.

Preferred Skills

  • Windows Server 2016/2019/2022 administration.
  • Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
  • Exposure to hybrid cloud and private cloud platforms.
  • Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
  • Experience supporting enterprise VDI environments.
  • Understanding of ITIL Incident, Problem, Change, and Release Management.
  • Experience working in regulated industries such as Banking or Financial Services.

Experience & Qualifications

  • 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
  • Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
  • Proven experience implementing automation to reduce operational overhead and improve service reliability.
  • Experience supporting enterprise private cloud and VDI environments.
  • Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
  • Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
  • Experience in Banking or Financial Services environments is advantageous.

Success Measures

  • Improved platform availability and reliability.
  • Reduced infrastructure incidents through automation and proactive monitoring.
  • Improved Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).
  • Increased infrastructure automation and operational efficiency.
  • Consistent achievement of service reliability objectives and operational KPIs.