1

Reliability Engineer Jobs in Alameda, CA (NOW HIRING)

Site Reliability Engineer

San Francisco, CA ยท On-site

$150 - $180/hr

As a SRE, you'll be responsible for the reliability, observability, performance, and security of our core platform. You'll work closely with our engineering team to develop and maintain the systems ...

Meet the Team The SRE Fleet team is responsible for maintaining the stability, scalability, and efficiency of the infrastructure that powers our global cloud platform. As a team of six engineers ...

The Global SRE team is responsible for owning and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Site Reliability Engineer who is ...

Senior SRE

San Francisco, CA ยท On-site

$135K - $159K/yr

The Global SRE team is responsible for owning and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Site Reliability Engineer who is ...

Senior SRE

San Francisco, CA ยท On-site

$167K - $196K/yr

The Global SRE team is responsible for owning and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Site Reliability Engineer who is ...

Senior SRE

San Francisco, CA ยท On-site

$135K - $159K/yr

The Global SRE team is responsible for owning and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Site Reliability Engineer who is ...

Senior SRE

San Francisco, CA ยท On-site

$135K - $159K/yr

The Global SRE team is responsible for owning and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Site Reliability Engineer who is ...

Site Reliability Engineer

Sunnyvale, CA ยท On-site

$145K - $175K/yr

Bolt Graphics is seeking a highly experienced Site Reliability Engineer (SRE) to design, build, and operate highly reliable developer and production systems. This role is mission-critical to ...

Meet the Team The SRE Fleet team is responsible for maintaining the stability, scalability, and efficiency of the infrastructure that powers our global cloud platform. As a team of six engineers ...

Senior Site Reliability Engineer

San Francisco, CA ยท On-site

$67.25 - $89.25/hr

The main goals of SRE are to create scalable and highly reliable systems. Our SREs ensure our production systems' reliability, performance, and scalability while enabling rapid development and ...

We're looking for a Systems Reliability Engineer to own the reliability of our system across cloud, edge, and real-world environments . Our platform runs across distributed infrastructure-connecting ...

Senior SRE

San Francisco, CA ยท On-site

$127K - $191K/yr

The Global SRE team is responsible for owning and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Site Reliability Engineer who is ...

Hardware Reliability Engineer

Mountain View, CA ยท On-site

$120K - $152K/yr

You'll define how we qualify 2.5D/3D and heterogeneously-integrated packages, model their physics of failure, drive root-cause when things fail, and build the reliability engineering that lets us ...

We are seeking a Site Reliability Engineer (SRE) to architect and manage the critical ground infrastructure for our satellite constellation. This role is responsible for the "last mile" of mission ...

SRE ARCHITECT

Fremont, CA ยท On-site

$62.50 - $83.25/hr

Info Way Solutions is seeking a highly experienced Site Reliability Engineering (SRE) Architect to lead the design, implementation, and governance of highly reliable, scalable, and resilient ...

Staff Reliability Engineer

Palo Alto, CA ยท On-site

$220K - $255K/yr

ALSO is looking for a Staff Reliability Engineer to play a key role in developing and leading the reliability of multiple electric mobility vehicles over the full development cycle. What You'll Do

Showing results 41-60

Reliability Engineer information

See Alameda, CA salary details

$69.1K

$133.7K

$159.8K

How much do reliability engineer jobs pay per year?

As of Aug 16, 2026, the average yearly pay for reliability engineer in Alameda, CA is $133,706.00, according to ZipRecruiter salary data. Most workers in this role earn between $116,200.00 and $146,200.00 per year, depending on experience, location, and employer.

What are some typical challenges reliability engineers face when implementing preventive maintenance strategies?

Reliability Engineers often encounter challenges such as balancing preventive maintenance schedules with production demands, ensuring buy-in from operations teams, and accurately predicting equipment failures. They must analyze large sets of historical data to identify trends and root causes, which can be complex in facilities with diverse machinery. Collaboration with maintenance, operations, and engineering teams is essential to develop effective strategies that minimize downtime while optimizing resources.

How much do reliability engineers get paid?

Reliability engineers typically earn a median annual salary ranging from $70,000 to $110,000, depending on experience, location, and industry. Senior or specialized reliability engineers with certifications and advanced skills can earn higher salaries, often exceeding $120,000 annually.

What are the key skills and qualifications needed to thrive as a reliability engineer, and why are they important?

To thrive as a Reliability Engineer, you need a solid background in engineering principles, failure analysis, and reliability modeling, typically with a degree in engineering or a related field. Familiarity with tools such as FMEA, Root Cause Analysis (RCA), reliability-centered maintenance (RCM) software, and certifications like Certified Reliability Engineer (CRE) are highly valued. Strong problem-solving abilities, attention to detail, and effective communication are crucial soft skills in this role. These skills ensure systems are dependable, downtime is minimized, and organizational performance and safety are optimized.

What is the difference between Reliability Engineer vs Maintenance Engineer?

AspectReliability EngineerMaintenance Engineer
CredentialsTypically requires engineering degree, certifications in reliability or asset managementOften requires engineering or technical diploma, certifications in maintenance or equipment repair
Work EnvironmentFocuses on analysis, design, and improvement of systems for reliabilityHands-on maintenance, repair, and troubleshooting of equipment
Industry UsageCommon in manufacturing, energy, aerospace, and industrial sectorsPrevalent in manufacturing, facilities, and industrial plants

Reliability Engineers focus on designing and improving systems to prevent failures, using data analysis and modeling. Maintenance Engineers perform hands-on repairs and upkeep of equipment to ensure operational continuity. While both roles aim to optimize equipment performance, Reliability Engineers work proactively on system reliability, whereas Maintenance Engineers handle reactive and scheduled maintenance tasks.

What does a reliability engineer do?

As a reliability engineer, your duties are to test and evaluate the manufacturing of products and components and ensure that the procedures are efficient and do not lead to abnormally high maintenance or operational costs. Your other responsibilities are to find solutions to product reliability risks. You may manage risk in a supply chain, develop loss prevention strategies, and track the entire lifecycle of product development, from building prototypes to moving a product into full-scale production. You analyze information from department heads and recommend strategies to reduce risk and ensure that the product works reliably.

What is a reliability engineer?

Reliability Engineers are professionals responsible for ensuring that systems, equipment, or processes function consistently and efficiently over time. They analyze data, identify potential points of failure, and develop maintenance strategies to improve system reliability and minimize downtime. Their work spans various industries, including manufacturing, energy, and technology, and often involves collaborating with design, operations, and maintenance teams. By implementing reliability-centered maintenance and predictive analysis, they help organizations save costs and increase safety.

What are popular job titles related to Reliability Engineer jobs in Alameda, CA?

For Reliability Engineer jobs in Alameda, CA, the most frequently searched job titles are:

What job categories do people searching Reliability Engineer jobs in Alameda, CA look for?

The top searched job categories for Reliability Engineer jobs in Alameda, CA are:

What cities near Alameda, CA are hiring for Reliability Engineer jobs?

Cities near Alameda, CA with the most Reliability Engineer job openings:

Infographic showing various Reliability Engineer job openings in Alameda, CA as of August 2026, with employment types broken down into 84% Full Time, 11% Part Time, and 5% Contract. Highlights an 84% Physical, 5% Hybrid, and 11% Remote job distribution, with an average salary of $133,706 per year, or $64.3 per hour.

Site Reliability Engineer

Runloop

San Francisco, CA โ€ข On-site

$150 - $180/hr

Other

Medical, Dental, Vision

Re-posted 2 days ago


Job description

About Runloop

Runloop.ai is pioneering the next generation of infrastructure and orchestration to power the Agentic Web/age of AI Agents. Our platform empowers developers to deploy agents that write code, browse the web, and use computers the way a human would. We're a small team of former Google and Stripe engineers, including the founding team of Google Wallet, dedicated to solving the complex challenges of productionizing AI for software engineering at scale.

The Role

We're looking for a skilled and passionate Site Reliability Engineer to join our team. As a SRE, you'll be responsible for the reliability, observability, performance, and security of our core platform. You'll work closely with our engineering team to develop and maintain the systems that power our code sandboxes, ensuring a seamless and stable experience for our customers. This is a critical role that blends a deep understanding of distributed systems with a software engineering mindset.

Responsibilities
  • Design and maintain our production infrastructure on cloud platforms like AWS, GCP, Azure, and emergent Neo-Clouds
  • Monitor and respond to system alerts and incidents using Grafana and Prometheus, ensuring high availability and a secure environment for our users
  • Collaborate with developers to ensure new features and services are designed with scalability and reliability in mind
  • Troubleshoot and resolve complex issues related to our infrastructure, networking, and the sandbox environment
  • Participate in an on-call rotation to support our production systems
  • Define and track SLIs/SLOs, manage error budgets, and proactively monitor distributed systems with logging and tracing
  • Automate deployments, scaling, provisioning, and recovery tasks to reduce toil and build self-healing systems
  • Lead incident response, conduct rootโ€‘cause analysis, and facilitate blameless postโ€‘mortems to drive continual improvement
  • Collaborate crossโ€‘functionally with product, engineering, and developer relations to ensure reliable releases and an outstanding developer experience
  • Plan for capacity growth, forecast system usage, and contribute to safe release and change management processes
Qualifications
  • Strong computer science fundamentals, backed by a degree from a topโ€‘tier CS/EE program, or equivalent experience
  • 5+ years of experience in software engineering, with at least 3 years focused explicitly on site reliability, DevOps, or infrastructure operations
  • Strong programming skills in languages like Python or Go
  • Deep expertise in containerization technologies such as Docker and Kubernetes
  • Experience with cloud infrastructure and tools like Terraform and/or Pulumi
  • Familiarity with monitoring and alerting tools like Prometheus, Grafana, or Datadog
  • A solid understanding of networking, security, and Linux systems administration
  • Experience designing, scaling, and maintaining distributed systems (backend platforms, APIs, or frontโ€‘end infrastructure)
  • Proficiency in implementing observability frameworks (metrics, logging, tracing) and aligning reliability goals with developer velocity
  • Handsโ€‘on experience managing incidents, running onโ€‘call operations, and producing actionable postโ€‘mortems
  • Ability to mentor engineers and influence reliability practices across teams, especially for frontโ€‘end infrastructure and performance
Bonus Points
  • Experience with chaos engineering techniques, frontโ€‘end observability tools (e.g., Sentry, RUM, synthetic monitoring), or building CI/CD pipelines for frontโ€‘end delivery
Benefits
  • Competitive salary and equity
  • Comprehensive health, dental, and vision insurance for employee and dependents
  • Opportunity to work on cuttingโ€‘edge technology and make a real impact on the future of software engineering
  • Daily catered lunch for all employees and a fridge full of your favorite snacks and drinks
Location
  • Onsite 4 days a week in San Francisco; Optional 1 day a week remote

Join Us! If you're excited about shaping the future of AIโ€‘driven software engineering and empowering developers to build the next generation of AI powered coding tools, we want to hear from you. Join the Runloop team and be at the forefront of the AI revolution in software development.

Runloop AI is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, disability status, protected veteran status, sexual orientation, gender identity, or any other characteristic protected by law.

#J-18808-Ljbffr