1

Reliability Engineer Jobs in Sunnyvale, CA (NOW HIRING)

Site Reliability Engineer

Santa Clara, CA ยท On-site

$230K - $250K/yr

As our first or early SRE hire you will be building the reliability engineering function at Forward - defining how we think about availability, observability, incident response, and operational ...

Site Reliability Engineer

Santa Clara, CA ยท On-site

$230K - $250K/yr

As our first or early SRE hire you will be building the reliability engineering function at Forward - defining how we think about availability, observability, incident response, and operational ...

Site Reliability Engineer

Palo Alto, CA ยท On-site

$67 - $89/hr

About the DevOps / SRE Team The DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable, scalable, and high-performing production systems. We work closely with ...

SRE

Sunnyvale, CA ยท On-site

$66.50 - $88.50/hr

Vantage Point Consulting Inc. is seeking a Site Reliability Engineer (SRE) to lead and manage SRE projects. The role involves collaborating with cross-functional teams to implement reliability ...

As our first reliability engineer, you will build the analytical foundation that ensures our autonomous tractor fleet meets the uptime and durability requirements of 24/7 airport operations. You will ...

SRE

Fremont, CA ยท On-site

$62.50 - $83.25/hr

Info Way Solutions is seeking an experienced SRE consultant to set up an SRE Platform and define SLI/SLO. The role requires hands-on experience with multiple observability, monitoring, and logging ...

Reliability Engineer (Hardware)

Mountain View, CA ยท On-site

$120K - $152K/yr

Bachelor's degree in Electrical Engineering, Computer Engineering or a similar field * 5+ years of industry experience working as a Reliability Engineer within the semiconductor and/or photonics ...

Reliability Engineer

San Carlos, CA ยท On-site

$110 - $150/hr

A day in the life As a Reliability Engineer at Swift, you'll drive cell-package interactions and module-level reliability, addressing the unique challenges of a first-generation technology moving ...

Site Reliability Engineer

Mountain View, CA ยท Hybrid

$189K - $232K/yr

As a Site Reliability Engineer, you will strengthen infrastructure, optimize tooling, deepen observability, streamline incident response, and elevate reliability standards. These actions empower ...

Principal Reliability Engineer

San Jose, CA ยท On-site

$120K - $151K/yr

We are looking for a Principal Reliability Engineer to join our team located in Tucson, AZ. What You Will Do: * Develop and implement reliability strategy across multiple product phases including ...

Site Reliability Engineer

Santa Clara, CA ยท On-site

$67.25 - $89.50/hr

We are looking for a Site Reliability Engineer to join our growing team in Riverwoods, IL United States! Responsibilities: * Partner with Application Development teams to build resiliency for Payment ...

The Role As a Site Reliability Engineer (SRE) at Air Apps, you will be responsible for ensuring the reliability, availability, and scalability of our systems. You will work at the intersection of ...

SRE Engineer

Santa Clara, CA ยท On-site

$67 - $89/hr

Omega Solutions, Inc. is seeking an SRE Engineer to join their team. The role involves operating infrastructure on AWS, developing self-service tooling for product engineering teams, and ensuring ...

Reliability Engineer

Santa Clara, CA ยท On-site

$130K - $155K/yr

The primary purpose of the Reliability Engineer (focused on L10/L11 execution) is to own the physical, environmental, and mechanical stress-testing validation of integrated AI rack systems. This ...

Java SRE Engineer Onsite San Francisco Bay Area Infrastructure Engineer (2 Positions) We are looking for an experienced Java SRE / Platform Engineer to support large-scale cloud migrations and ...

Showing results 21-40

Reliability Engineer information

See Sunnyvale, CA salary details

$72.6K

$140.4K

$167.7K

How much do reliability engineer jobs pay per year?

As of Sep 3, 2026, the average yearly pay for reliability engineer in Sunnyvale, CA is $140,353.00, according to ZipRecruiter salary data. Most workers in this role earn between $121,900.00 and $153,500.00 per year, depending on experience, location, and employer.

What is a reliability engineer?

Reliability Engineers are professionals responsible for ensuring that systems, equipment, or processes function consistently and efficiently over time. They analyze data, identify potential points of failure, and develop maintenance strategies to improve system reliability and minimize downtime. Their work spans various industries, including manufacturing, energy, and technology, and often involves collaborating with design, operations, and maintenance teams. By implementing reliability-centered maintenance and predictive analysis, they help organizations save costs and increase safety.

What does a reliability engineer do?

As a reliability engineer, your duties are to test and evaluate the manufacturing of products and components and ensure that the procedures are efficient and do not lead to abnormally high maintenance or operational costs. Your other responsibilities are to find solutions to product reliability risks. You may manage risk in a supply chain, develop loss prevention strategies, and track the entire lifecycle of product development, from building prototypes to moving a product into full-scale production. You analyze information from department heads and recommend strategies to reduce risk and ensure that the product works reliably.

What are the key skills and qualifications needed to thrive as a reliability engineer, and why are they important?

To thrive as a Reliability Engineer, you need a solid background in engineering principles, failure analysis, and reliability modeling, typically with a degree in engineering or a related field. Familiarity with tools such as FMEA, Root Cause Analysis (RCA), reliability-centered maintenance (RCM) software, and certifications like Certified Reliability Engineer (CRE) are highly valued. Strong problem-solving abilities, attention to detail, and effective communication are crucial soft skills in this role. These skills ensure systems are dependable, downtime is minimized, and organizational performance and safety are optimized.

What are some typical challenges reliability engineers face when implementing preventive maintenance strategies?

Reliability Engineers often encounter challenges such as balancing preventive maintenance schedules with production demands, ensuring buy-in from operations teams, and accurately predicting equipment failures. They must analyze large sets of historical data to identify trends and root causes, which can be complex in facilities with diverse machinery. Collaboration with maintenance, operations, and engineering teams is essential to develop effective strategies that minimize downtime while optimizing resources.

What is the difference between Reliability Engineer vs Maintenance Engineer?

AspectReliability EngineerMaintenance Engineer
CredentialsTypically requires engineering degree, certifications in reliability or asset managementOften requires engineering or technical diploma, certifications in maintenance or equipment repair
Work EnvironmentFocuses on analysis, design, and improvement of systems for reliabilityHands-on maintenance, repair, and troubleshooting of equipment
Industry UsageCommon in manufacturing, energy, aerospace, and industrial sectorsPrevalent in manufacturing, facilities, and industrial plants

Reliability Engineers focus on designing and improving systems to prevent failures, using data analysis and modeling. Maintenance Engineers perform hands-on repairs and upkeep of equipment to ensure operational continuity. While both roles aim to optimize equipment performance, Reliability Engineers work proactively on system reliability, whereas Maintenance Engineers handle reactive and scheduled maintenance tasks.

Are reliability engineers in demand?

Reliability engineers are in high demand across industries such as manufacturing, energy, and aerospace due to their role in improving system performance and reducing downtime. Employers seek professionals with skills in data analysis, failure modes, and maintenance strategies, often requiring certifications like Certified Reliability Engineer (CRE). The job outlook is positive, with steady growth expected as companies prioritize operational efficiency and risk management.

How much do reliability engineers get paid?

Reliability engineers typically earn a median annual salary ranging from $70,000 to $110,000, depending on experience, location, and industry. Senior or specialized reliability engineers with certifications and advanced skills can earn higher salaries, often exceeding $120,000 annually.

What are the most commonly searched types of Reliability Engineer jobs in Sunnyvale, CA?

The most popular types of Reliability Engineer jobs in Sunnyvale, CA are:

What job categories do people searching Reliability Engineer jobs in Sunnyvale, CA look for?

The top searched job categories for Reliability Engineer jobs in Sunnyvale, CA are:

What cities near Sunnyvale, CA are hiring for Reliability Engineer jobs?

Cities near Sunnyvale, CA with the most Reliability Engineer job openings:

Infographic showing various Reliability Engineer job openings in Sunnyvale, CA as of August 2026, with employment types broken down into 91% Full Time, and 9% Contract. Highlights an 89% In-person, and 11% Remote job distribution, with an average salary of $138,459 per year, or $66.6 per hour.

Site Reliability Engineer

Forward

Santa Clara, CA โ€ข On-site

$230K - $250K/yr

Full-time

Re-posted 20 days ago


Job description

Forward was founded in 2013 by four Stanford Ph.D.s, building the industry's first network digital twin: a mathematically accurate model of the production network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change before it touches production. That founding instinct still defines how we work. We're accurate and evidence-driven, relentless about clarity, and we'd rather be certain than comfortable, building a groundbreaking platform that transforms how teams run and secure networks across every major cloud and vendor environment.

Global leaders like Goldman Sachs, PayPal, S&P Global, IBM, and Dell trust Forward, alongside fast-growing enterprises and government agencies, realizing an average of $14.2 million in annual benefits, according to IDC. Backed by top-tier investors, including A. Capital, Andreessen Horowitz, Goldman Sachs, MSD Partners, Omega Venture Partners, Section 32, and Threshold Ventures, and headquartered in Santa Clara, we're most proud of our team: curious people who'd rather build what doesn't exist than accept how things have always been done.
Forward is looking for a Site Reliability Engineer

About the Role This is not a "keep the lights on" SRE role. As our first or early SRE hire you will be building the reliability engineering function at Forward - defining how we think about availability, observability, incident response, and operational excellence across a complex, distributed SaaS platform. You will work closely with engineering, infrastructure, and product to ensure our platform meets the reliability bar our enterprise customers demand.

If you thrive in environments where you're handed a problem rather than a playbook this role is for you.

What You'll Own

  • Define and drive SRE practices from the ground up - SLOs, SLIs, error budgets, and the frameworks the engineering org will actually use
  • Drive the reliability and operational excellence of the Forward SaaS platform
  • Build and maintain observability infrastructure - logging, metrics, tracing, and alerting - so the team always knows what's happening before customers do
  • Lead incident response: on-call rotations, runbooks, post-mortems, and the follow-through to make sure the same incident doesn't happen twice
  • Partner with engineering teams to embed reliability thinking into the SDLC - capacity planning, load testing, chaos engineering, and production readiness reviews
  • Help define and build the SRE team as the company scales - this is a foundational hire with a path to leadership

What We're Looking For

  • 6+ years of experience in site reliability engineering, DevOps, or infrastructure engineering in a SaaS or cloud environment
  • Proven experience building or significantly maturing an SRE function - not just operating within one someone else built
  • Strong fundamentals in networking - TCP/IP, DNS, routing, switching, firewalls, and load balancing. Experience with network management or observability platforms is a significant plus
  • Hands-on experience with Kubernetes and container orchestration in production environments
  • Deep proficiency with observability tooling - Prometheus, Grafana, Datadog, Splunk, or similar
  • Strong scripting and automation skills in Python, Bash, or similar
  • Experience with cloud platforms - AWS, GCP, or Azure - including infrastructure as code (Terraform, Ansible, or equivalent)
  • Track record of owning and improving incident response processes including blameless post-mortems and SLO-driven reliability improvements
  • Ability to communicate clearly with both engineering teams and non-technical stakeholders - you can explain an outage to a customer-facing team without jargon and explain an SLO to an executive without losing them

Nice to Have

  • Experience supporting enterprise or federal government customers with high availability requirements
  • Experience in a foundational or early SRE hire capacity at a growth stage company

What This Role Is Not

  • A pure ops or NOC role - you are building and engineering, not just monitoring
  • A siloed function - you will be deeply embedded with product and engineering teams
  • A ticket-taker - you will be proactively identifying and solving reliability problems before they become incidents

Why Forward

  • You'll be building something from scratch at a company with real enterprise traction and world-class investors behind it
  • Our customers include some of the most complex network environments on the planet - the reliability bar is high and the work is genuinely interesting
  • People-centric culture built by Stanford Ph.D.s who care deeply about doing things the right way
  • Competitive compensation, equity, and the opportunity to grow into a leadership role as the SRE function scales

The base pay range for this role is between $230,000 and $250,000. Base pay will depend on your skills, qualifications, experience, and location