1

Site Reliability Engineer Manager Jobs in San Ramon, CA

Senior Site Reliability Engineer

San Francisco, CA · On-site

$67.25 - $89.25/hr

Qualifications BS/MS in Computer Science or Equivalent 6+ years of experience in Site Reliability Engineering, managing infrastructure and services at scale History of end-to-end project delivery ...

Site Reliability Engineer (SRE)

San Francisco, CA · On-site

$67.25 - $89.25/hr

Methodic is seeking a Site Reliability Engineer (SRE) to focus on the stability and efficiency of ... incident management and a track record of improving systems based on lessons learned. • ...

Site Reliability Engineer (SRE)

San Francisco, CA · On-site

$67.25 - $89.25/hr

... to manage Mithril's growing multi-cloud provider footprint. • Write clean, maintainable Python (or Go) to automate repetitive operational tasks -- from provider API reconciliation to automated ...

Site Reliability Engineer

San Francisco, CA · On-site

$67.25 - $89.25/hr

... Site Reliability Engineer to own infrastructure and reliability. This role involves designing ... and manage cloud infrastructure efficiently. • Comfortable handling production incidents ...

Site Reliability Engineer

San Francisco, CA · On-site

$67.25 - $89.25/hr

Senior Site Reliability Engineer (SRE) Work Location : San Francisco, California 94105 (Hybrid) Contract Key Responsibilities: Strong hands-on experience on AWS cloud and infrastructure as code ...

New

Site Reliability Engineer

Fremont, CA · On-site

$62.50 - $83/hr

Senior Site Reliability Engineer (SRE) Work Location : San Francisco, California 94105 (Hybrid) Contract Key Responsibilities: Strong hands-on experience on AWS cloud and infrastructure as code ...

New

Site Reliability Engineer

Hayward, CA · On-site

$65.25 - $86.75/hr

Senior Site Reliability Engineer (SRE) Work Location : San Francisco, California 94105 (Hybrid) Contract Key Responsibilities: Strong hands-on experience on AWS cloud and infrastructure as code ...

New

Site Reliability Engineer

San Jose, CA · On-site

$66.75 - $88.75/hr

Senior Site Reliability Engineer (SRE) Work Location : San Francisco, California 94105 (Hybrid) Contract Key Responsibilities: Strong hands-on experience on AWS cloud and infrastructure as code ...

New

Site Reliability Engineer

Alameda, CA · On-site

$66 - $87.75/hr

Senior Site Reliability Engineer (SRE) Work Location : San Francisco, California 94105 (Hybrid) Contract Key Responsibilities: Strong hands-on experience on AWS cloud and infrastructure as code ...

New

Site Reliability Engineer

San Mateo, CA · On-site

$65 - $86.25/hr

Senior Site Reliability Engineer (SRE) Work Location : San Francisco, California 94105 (Hybrid) Contract Key Responsibilities: Strong hands-on experience on AWS cloud and infrastructure as code ...

New

Staff Site Reliability Engineer

San Francisco, CA · On-site +1

$67.25 - $89.25/hr

Operate and scale production blockchain infrastructure, managing full nodes across networks such as ... Staff Site Reliability Engineer (IV) Senior Site Reliability Engineer (III) What you'll bring to ...

Staff Site Reliability Engineer

San Francisco, CA · On-site +1

$67.25 - $89.25/hr

Operate and scale production blockchain infrastructure, managing full nodes across networks such as ... Staff Site Reliability Engineer (IV) Senior Site Reliability Engineer (III) What you'll bring to ...

SRE ARCHITECT

Fremont, CA · On-site

$62.50 - $83.25/hr

Info Way Solutions is seeking a highly experienced Site Reliability Engineering (SRE) Architect to ... management skills Company : Founded and incorporated in 2012 , Info Way Solutions is an IT services ...

SRE Engineer

San Jose, CA · On-site

$66.75 - $88.75/hr

Job Title: SRE Engineer Location: San Jose, CA / RTP, NC(Onsite) Job Type: Full Time Must Have Technical/Functional Skills: * SRE, NetApp Storage, Linux Certified, Kubernetes Certified, DevOps, ...

next page

Showing results 1-20

Site Reliability Engineer Manager information

See San Ramon, CA salary details

$12

$71

$102

How much do site reliability engineer manager jobs pay per hour?

As of Aug 25, 2026, the average hourly pay for site reliability engineer manager in San Ramon, CA is $71.23, according to ZipRecruiter salary data. Most workers in this role earn between $61.25 and $81.39 per hour, depending on experience, location, and employer.

What is a site reliability engineer manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.

How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How much do site reliability engineer managers get paid?

Site Reliability Engineer Managers typically earn between $120,000 and $180,000 annually, depending on experience, location, and company size. They often oversee teams responsible for system reliability, incident response, and infrastructure automation, requiring strong leadership and technical skills.

Is a Site Reliability Engineer Manager a stressful job?

A Site Reliability Engineer Manager role can be stressful due to the responsibility of maintaining system uptime, managing incident responses, and ensuring reliability across complex infrastructure. The job often involves working under pressure, handling outages, and coordinating teams, but it also offers opportunities for problem-solving and leadership. Stress levels vary depending on company size, team structure, and workload management skills.

What are the most commonly searched types of Site Reliability Engineer jobs in San Ramon, CA?

The most popular types of Site Reliability Engineer jobs in San Ramon, CA are:

What cities near San Ramon, CA are hiring for Site Reliability Engineer Manager jobs?

Cities near San Ramon, CA with the most Site Reliability Engineer Manager job openings:

Infographic showing various Site Reliability Engineer Manager job openings in San Ramon, CA as of August 2026, with employment types broken down into 83% Full Time, 16% Part Time, and 1% Contract. Highlights an 79% Physical, 2% Hybrid, and 19% Remote job distribution, with an average salary of $148,164 per year, or $71.2 per hour.

Manager, Engineering - Dev Ops/SRE (Hybrid)

CrowdStrike Holdings, Inc.

Sunnyvale, CA • On-site

$67 - $89/hr

Full-time

Medical, Retirement, PTO

Re-posted 15 days ago


Job description

As a global leader in cybersecurity, CrowdStrike protects the people, processes and technologies that drive modern organizations. Since 2011, our mission hasn't changed - we're here to stop breaches, and we've redefined modern security with the world's most advanced AI-native platform. We work on large scale distributed systems, processing almost 3 trillion events per day and this traffic is growing daily. Our customers span all industries, and they count on CrowdStrike to keep their businesses running, their communities safe and their lives moving forward. We're proud to work for a mission-driven company leveraging AI to transform the way we work. CrowdStrikers drive their careers through flexibility and autonomy while also being expected to contribute to a culture of responsible AI adoption, experimentation, and innovation. We use an AI-first mindset as a force multiplier to proactively and continuously accelerate execution, build expertise, uncover insights, and solve complex problems. We're always looking to add talented CrowdStrikers to the team who have limitless passion, a relentless focus on innovation and a fanatical commitment to our customers, our community and each other. Ready to join a mission that matters? The future of cybersecurity starts with you.
About the Role:
At CrowdStrike, Site Reliability Engineering (SRE) is at the forefront of ensuring the reliability and scalability of our cloud-native security platform. In this role, you'll manage a team of talented engineers, providing technical leadership on key projects and empowering them to excel in their roles.
As an SRE Manager, you will lead a team of SRE engineers ensuring the reliability, scalability, and performance of CrowdStrike's cloud-native security platform. You'll provide technical leadership and mentorship, owning both reliability engineering and software delivery pipelines - driving engineering velocity while maintaining zero tolerance for downtime in security-critical infrastructure.
What You'll Do:
  • Define and enforce SLOs, SLIs, and error budgets across distributed systems processing millions of events per second
  • Drive system reliability by blending software engineering principles with AI-driven automation, moving from reactive firefighting to proactive, automated operations
  • Lead major incident response and facilitate blameless postmortems, driving systemic reliability improvements
  • Own capacity planning, traffic management, and load shedding strategies for high-throughput distributed systems
  • Own the end-to-end software delivery pipeline strategy - designing, building, and maintaining scalable, reliable pipelines using Jenkins, GitLab CI, and Bitbucket Pipelines
  • Build and maintain observability frameworks including metrics, distributed tracing, and log aggregation across the full stack
  • Champion chaos engineering and resilience validation practices for security-critical systems
  • Lead and grow a high-performing SRE team, mentoring engineers and fostering a culture of continuous learning and operational excellence
  • Partner with cross-functional engineering teams to embed reliability practices early in the software development lifecycle

What You'll Need:
Experience & Leadership
  • Proven track record of building, growing, and retaining high-performing SRE/DevOps engineering teams in a fast-paced, high-growth environment
  • 10+ years of software engineering experience with significant focus on reliability engineering, platform infrastructure, and production operations at scale
  • 3+ years of hands-on management experience overseeing SRE/DevOps engineering teams, including incident command and reliability ownership
  • Bachelor's degree in Computer Science or related field, or equivalent work experience

Reliability Engineering
  • Deep understanding of SRE principles including SLOs, SLAs, SLIs, and error budgeting strategies applied to large-scale distributed systems
  • Proven experience owning reliability for high-throughput distributed systems processing millions of events per second, including capacity planning, traffic management, and load shedding strategies
  • Strong incident management facilitating blameless postmortems, and driving system reliability improvements
  • Demonstrated ability to build, operationalize, and maintain highly scalable, security-critical microservices-based distributed systems with zero tolerance for data loss or downtime.
  • Advanced observability experience including Prometheus, Grafana, distributed tracing (Jaeger/OpenTelemetry), and large-scale log aggregation (ELK/Splunk) with a focus on building custom SLO dashboards and reliability scorecards.
  • Experience owning disaster recovery strategies including backup automation, failover testing, and business continuity planning for stateful distributed systems

Platform and Delivery Engineering
  • Proficiency in Python and/or Golang for automation, tooling, and platform services
  • Hands-on experience designing and managing scalable software delivery pipelines using Jenkins, GitLab CI, Bitbucket Pipelines, or equivalent
  • Strong proficiency in Infrastructure as Code (IaC) - Terraform, Ansible, Pulumi, or equivalent
  • Familiarity with GitOps workflows using ArgoCD or Flux for managing infrastructure deployments at scale

Cloud and Big Data Exposure
  • Proficiency in at least one cloud environment (AWS, Azure, GCP) with emphasis on multi-region architecture, cloud-native reliability patterns, and security-first cloud design
  • Strong experience with Kubernetes at scale - managing large cluster fleets, workload orchestration, and container lifecycle management
  • Familiarity with distributed data systems including relational databases (PostgreSQL), NoSQL (Cassandra), OLAP (Pinot), Indexing(OpenSearch) and real-time streaming platforms (Kafka, Flink)
  • Exposure to Big Data and analytics technologies like Spark,Storm.

#LI-AP1
Benefits of Working at CrowdStrike:
  • Market leader in compensation and equity awards
  • Comprehensive physical and mental wellness programs
  • Competitive vacation and holidays for recharge
  • Paid parental and adoption leaves
  • Professional development opportunities for all employees regardless of level or role
  • Employee Networks, geographic neighborhood groups, and volunteer opportunities to build connections
  • Vibrant office culture with world class amenities
  • Great Place to Work Certified™ across the globe

CrowdStrike is proud to be an equal opportunity employer. We are committed to fostering a culture of belonging where everyone is valued for who they are and empowered to succeed. We support veterans and individuals with disabilities through our affirmative action program.
CrowdStrike is committed to providing equal employment opportunity for all employees and applicants for employment. The Company does not discriminate in employment opportunities or practices on the basis of race, color, creed, ethnicity, religion, sex (including pregnancy or pregnancy-related medical conditions), sexual orientation, gender identity, marital or family status, veteran status, age, national origin, ancestry, physical disability (including HIV and AIDS), mental disability, medical condition, genetic information, membership or activity in a local human rights commission, status with regard to public assistance, or any other characteristic protected by law. We base all employment decisions--including recruitment, selection, training, compensation, benefits, discipline, promotions, transfers, lay-offs, return from lay-off, terminations and social/recreational programs--on valid job requirements.
If you need assistance accessing or reviewing the information on this website or need help submitting an application for employment or requesting an accommodation, please contact us at recruiting@crowdstrike.com for further assistance.
Find out more about your rights as an applicant.
CrowdStrike participates in the E-Verify program.
Notice of E-Verify Participation
Right to Work
CrowdStrike, Inc. is committed to fair and equitable compensation practices. Placement within the pay range is dependent on a variety of factors including, but not limited to, relevant work experience, skills, certifications, job level, supervisory status, and location. The base salary range for this position for all U.S. candidates is $140,000 - $215,000 per year, with eligibility for bonuses, equity grants and a comprehensive benefits package that includes health insurance, 401k and paid time off.
For detailed information about the U.S. benefits package, please click here.

CrowdStrike logo

About CrowdStrike

Sourced by ZipRecruiter

#WeAreCrowdStrike and our mission is to stop breaches. As a global leader in cybersecurity, our team changed the game. Since our inception, our market leading cloud-native platform has offered unparalleled protection against the most sophisticated cyberattacks. We're looking for people with limitless passion, a relentless focus on innovation and a fanatical commitment to the customer to join us in shaping the future of cybersecurity. Consistently recognized as a top workplace, CrowdStrike is committed to cultivating an inclusive, remote-first culture that offers people the autonomy and flexibility to balance the needs of work and life while taking their career to the next level. Interested in working for a company that sets the standard and leads with integrity? Join us on a mission that matters - one team, one fight.

Industry

It services

Company size

1,001 - 5,000 Employees

Headquarters location

Sunnyvale, CA, US

Year founded

2012