1

Site Reliability Engineer Sre Jobs in Riverside, CA

Reliability Engineer

Irvine, CA · On-site

$110K - $138K/yr

The Reliability Engineer will define, design, develop, and monitor the site's critical asset register to provide standardized processes and procedures for the proper operation and care of key ...

Reliability Engineer

Irvine, CA

$110K - $138K/yr

The Reliability Engineer will define, design, develop, and monitor the site's critical asset register to provide standardized processes and procedures for the proper operation and care of key ...

The Reliability Engineer will define, design, develop, and monitor the site's critical asset register to provide standardized processes and procedures for the proper operation and care of key ...

Reliability Engineer

Irvine, CA · On-site

$110 - $140/hr

The Reliability Engineer will define, design, develop, and monitor the site's critical asset register to provide standardized processes and procedures for the proper operation and care of key ...

Showing results 21-40

Site Reliability Engineer Sre information

See Riverside, CA salary details

$11

$66

$95

How much do site reliability engineer sre jobs pay per hour?

As of Aug 16, 2026, the average hourly pay for site reliability engineer sre in Riverside, CA is $66.50, according to ZipRecruiter salary data. Most workers in this role earn between $57.16 and $76.01 per hour, depending on experience, location, and employer.

What is the difference between Site Reliability Engineer Sre vs DevOps Engineer?

AspectSite Reliability Engineer SreDevOps Engineer
CredentialsTypically requires experience in systems engineering, scripting, and monitoring toolsOften has certifications in cloud platforms, automation, and CI/CD tools
Work EnvironmentFocuses on maintaining system reliability, scalability, and incident responseEmphasizes automation, deployment, and continuous integration/delivery
Industry UsageCommon in tech, finance, and large-scale cloud servicesWidely used across startups, tech companies, and enterprises

While both roles aim to improve system performance and automation, Site Reliability Engineers Sre primarily focus on reliability and incident management, whereas DevOps Engineers concentrate on deployment automation and continuous integration. The roles often overlap but serve distinct core functions within IT and software development teams.

What are some of the most common challenges faced by site reliability engineers (SREs), and how can they be addressed?

Site Reliability Engineers often face challenges such as managing the balance between reliability and rapid feature deployment, handling on-call responsibilities, and automating manual processes. To address these, SREs work closely with development and operations teams to implement strong monitoring, establish clear service level objectives (SLOs), and continually improve incident response procedures. Building a culture of blameless postmortems and investing in automation can also help reduce repetitive work and improve overall system reliability.

What is a site reliability engineer (SRE)?

A Site Reliability Engineer (SRE) is a professional who combines software engineering and IT operations skills to ensure reliable, scalable, and efficient systems. SREs automate system operations, monitor system health, and manage incident response to minimize downtime and ensure high availability. Their work bridges the gap between development and operations teams, focusing on building robust infrastructure and processes. SREs often use tools and practices like automation, monitoring, and performance tuning to maintain service reliability. They also set and measure service level objectives (SLOs) to align system performance with business goals.

What are the key skills and qualifications needed to thrive as a site reliability engineer (SRE), and why are they important?

To thrive as a Site Reliability Engineer (SRE), you need a solid background in software engineering, systems administration, and automation, often supported by a degree in computer science or a related field. Familiarity with cloud platforms (like AWS, GCP, or Azure), containerization (Docker, Kubernetes), and monitoring tools (Prometheus, Grafana), as well as scripting languages such as Python or Bash, is typically required. Strong problem-solving, communication, and collaboration skills help SREs manage incidents and work effectively with development and operations teams. These competencies are essential to ensure system reliability, optimize performance, and maintain seamless service availability.

What are popular job titles related to Site Reliability Engineer Sre jobs in Riverside, CA?

For Site Reliability Engineer Sre jobs in Riverside, CA, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer Sre jobs in Riverside, CA look for?

The top searched job categories for Site Reliability Engineer Sre jobs in Riverside, CA are:

What cities near Riverside, CA are hiring for Site Reliability Engineer Sre jobs?

Cities near Riverside, CA with the most Site Reliability Engineer Sre job openings:

Infographic showing various Site Reliability Engineer Sre job openings in Riverside, CA as of August 2026, with employment types broken down into 57% Full Time, and 43% Contract. Highlights an 100% In-person job distribution, with an average salary of $138,320 per year, or $66.5 per hour.

Splunk Administrator (Site Reliability Engineer)

Schoolsfirstfcu

Tustin, CA

$59.75 - $79.50/hr

Full-time

Re-posted 6 days ago


Job description

We're always looking for diverse, talented, service-oriented people to join our exceptional team.

Splunk Administrator (Site Reliability Engineer)

The pay range for this position is listed below. Our pay ranges are built to allow for candidates with various levels of skill and experience to be considered, as well as for room for growth and tenure achieved in a role over time. Typical new hire salary offers fall within the minimum to midpoint of a pay range for many candidates. Any offer extended to a candidate will be based upon their unique set of knowledge, skills, education, and experience as well as internal equity.

Pay Range:

$46.90 - $75.04

Scheduled Weekly Hours:

40
What You'll Be Doing

Responsible for deploying, managing and optimizing. Onboard machine data, build queries using Search Processing Language (SPL) and design dashboards to power IT Enterprise Applications, Security Operations and observability practices. Proactively monitor and report on the environmental health of applications, as well as engage in critical system outages (triage) resolution efficiently to identify, review, analyze, debug and resolve issues with a team of Site Reliability Engineers. Serve as a subject matter expert for the Splunk Platform. .

  • Ensure enterprise applications and supporting platforms are fully functional, reliable, available, performant, and secure through frequent monitoring, operational checks, and timely restoration of service in the event of outages, ingestion issues, or system failures.

  • Ingest and onboard new operational data sources, including server, application, infrastructure, network, security, API, and system logs; ensure data is reliable, appropriately normalized, accurately timestamped, correctly categorized, and aligned to the Splunk Common Information Model where applicable.

  • Monitor and maintain Splunk platform health, including indexers, search heads, data flows, storage utilization, clustered or distributed components, internal logs, capacity trends, and service availability; resolve ingestion bottlenecks, performance issues, and platform outages.

  • Create, maintain, and improve real-time Splunk dashboards, visualizations, alerts, reports, metrics, logs, traces, and event views to provide actionable insight into system health, service levels, user experience, operational risk, network activity, security events, and threat indicators.

  • Craft, tune, and optimize Splunk Processing Language (SPL) searches, scheduled searches, alert logic, reports, dashboards, and resource-intensive queries to improve performance, reduce noise, support accurate KPI/SLA reporting, and enable analysis across large datasets.

  • Configure automated alerts and triggers for anomaly detection, system downtime, performance degradation, ingestion issues, and cyber or operational threats so critical issues are flagged immediately and routed for response.

  • Alert and accurately report KPIs on systems status with tuning recommendations at regular intervals; provide detailed analysis using APM, Splunk, and other monitoring or observability solutions.

  • Support a 24x7 production environment with a team of experienced engineers, including on-call rotation, deployment support, and timely response to critical incidents as needed.

  • Respond quickly to incidents, investigate triggered alerts, isolate performance bottlenecks, fix issues, and work with other engineers to ensure enterprise applications and platforms remain fully functional with a Member-first focus.

  • Serve as an escalation point for complex application, infrastructure, observability, security, and platform issues; coordinate resolution across infrastructure engineering, software engineering, software quality assurance, cybersecurity, operations, vendors, and other organizational teams.

  • Support incident triage, root cause analysis, post-incident reviews, postmortems, and corrective action planning; identify recurring issues and recommend preventive improvements.

  • Develop solutions to meet technical and business requirements; create detailed design documents and associated solutions built around current and new technologies.

  • Demonstrate application environment tuning abilities and provide solutions to capacity requirements, including JVM, web containers, database connections, HTTP servers, and related enterprise application components.

  • Contribute to automation and efficiency improvements for application and infrastructure buildout components, operational tasks, monitoring coverage, and deployment readiness.

  • Ensure effective release management, change management, quality assurance, rollback planning, peer review, and deployment readiness in project lifecycles using an ITIL-centric approach.

  • Support production and non-production environments, account for differences in each environment for deployments and testing, and peer review changes within the Enterprise Applications team.

  • Maintain and improve configuration documentation, diagrams, SOPs, runbooks, work/incident ticketing systems, knowledge articles, training materials, and other supporting documentation.

  • Ensure adherence to configuration, operating, security, compliance, and governance best practices and standards, including approved capacity, licensing, access control, and retention requirements.

  • Deploy and troubleshoot applications on both Java and .NET platforms as an expert for infrastructure and development teams.

  • Ensure the highest levels of service for Members and Team Members.

Additional Job Functions
  • Performs other duties as assigned
  • Complies with regulatory compliance and assigned training requirements including but not limited to BSA regulations corresponding to their specific job duties. Failure to do so may result in disciplinary and other employment related actions.
    Qualifications
    • High School Diploma or GED required
    • Bachelor's Degree in a related field or equivalent years of experience required
    • 5-7 years of prior relevant experience required
      Knowledge, Skills, and Abilities
      • Knowledge of the following:
      • Java Enterprise Edition and .NET frameworks.
      • Standard java/.NET tools to debug application and system performance issues.
      • Tuning application serving environments.
      • Software release management and the SDLC.
      • Redhat Linux and Microsoft Windows Server 2008/2012.
      • Security best practices for multi-tiered operating systems as well as application security practices.
      • Disaster recovery experience Preferred
      • Application Performance Management, such as Application Dynamics, Splunk and/or SNMP monitoring tools
      • IBM Security Access Manager (ISAM), IBM/Apache HTTP Server, IBM Secure Directory Server Preferred
      • Scripting skills for automating tasks
      • Experience with Docker, Kubernetes, Openshift, Python, Jenkins, Git/Gitflow, Ansible, Quay and Artifactory Preferred
      • Managing IBM AIX, Redhat, Linux and/or Microsoft Windows Servers, process development, support and implementation of systems architecture
      • Must possess solid oral and written communication, project management, and interpersonal skills.
        SchoolsFirst FCU is committed to Diverse, Equitable, and Inclusive Hiring
        At SchoolsFirst FCU we are dedicated to building and growing a diverse, inclusive, and authentic Dream Team, so if you're excited about a position or wanting to make a career change but your past experience doesn't align perfectly with every qualification in the job description, we encourage you to apply anyway. Many skills are transferrable and you may be just the right candidate for the position, or for other roles we are working on.

        SchoolsFirst Federal Credit Union is committed to fostering, cultivating, and preserving a culture of diversity and inclusion. SchoolsFirst FCU is an equal opportunity employer and prohibits discrimination against qualified individuals based on their status as protected veterans or individuals with disabilities and prohibits discrimination against all individuals based on their race, color, religion, sex, national origin, age, sexual orientation, gender identity or expression, political affiliation, or genetic information.

        This organization participates in E-Verify.