2

Remote Mechanical Reliability Engineer Jobs in California

Site Reliability Engineer

San Francisco, CA · Remote

$67.25 - $89.25/hr

Remote (US) Department: Cloud Platform Engineering / SRE/Reliability Position summary The Site Reliability Engineer (SRE) owns reliability, observability, and incident response for the GPU One ...

Senior Site Reliability Engineer

San Diego, CA · Remote

$60.50 - $80.50/hr

This is a remote, contract opportunity for a project Arctiq is delivering for a client. Candidates ... The Senior Site Reliability Engineer is a technical leader responsible for architecting the ...

$98K - $138K/yr

The Site Reliability Engineer II will be responsible for supporting, enhancing, and maintaining ... Wellness initiatives #BI-Remote Internal Employees - R365 is committed to growing talent from ...

$98K - $138K/yr

The Site Reliability Engineer II will be responsible for supporting, enhancing, and maintaining ... Wellness initiatives #BI-Remote Internal Employees - R365 is committed to growing talent from ...

Sr. Site Reliability Engineer

San Francisco, CA · On-site +1

$67.25 - $89.25/hr

Open to remote or San Francisco Bay Area, Nashville Metro Area, or Raleigh, NC Area What you\'ll do ... SRE, platform, or staff infrastructure role * Deep Kubernetes expertise across managed (EKS, GKE ...

Data Reliability Engineer

San Diego, CA · On-site +1

$200K - $240K/yr

None Potential for Remote Work: ORA_ON_SITE Description We are seeking a Data Reliability Engineer to design, build, and maintain real-time data ingestion pipelines. In this role, you will be ...

Meet the Team We are seeking a highly experienced and driven ASIC Mechanical Engineer to join the ... Own and drive solder joint reliability assessments, including thermal cycling, mechanical shock ...

Meet the Team We are seeking a highly experienced and driven ASIC Mechanical Engineer to join the ... Own and drive solder joint reliability assessments, including thermal cycling, mechanical shock ...

next page

Showing results 1-20

Remote Mechanical Reliability Engineer information

What is the difference between Remote Mechanical Reliability Engineer vs Mechanical Maintenance Engineer?

AspectRemote Mechanical Reliability EngineerMechanical Maintenance Engineer
CredentialsBachelor's in Mechanical Engineering, certifications like RCM or FMEABachelor's in Mechanical Engineering or related field, certifications like HVAC or CMMS experience
Work EnvironmentPrimarily office-based, remote collaboration, data analysisOn-site plant or facility, hands-on equipment maintenance
Industry UsageManufacturing, energy, oil & gas, where reliability analysis is keyManufacturing, facilities management, industrial plants

The Remote Mechanical Reliability Engineer focuses on analyzing and improving equipment reliability remotely, often through data analysis and predictive maintenance. In contrast, the Mechanical Maintenance Engineer performs hands-on repairs and maintenance on-site. Both roles require mechanical engineering knowledge, but their work environments and daily tasks differ significantly.

What are the key skills and qualifications needed to thrive as a Remote Mechanical Reliability Engineer, and why are they important?

To thrive as a Remote Mechanical Reliability Engineer, you need a solid background in mechanical engineering, experience with reliability analysis, and typically a bachelor's degree in engineering. Familiarity with reliability-centered maintenance (RCM) tools, predictive analytics software, and Computerized Maintenance Management Systems (CMMS) is often required. Strong problem-solving abilities, effective communication, and self-motivation are crucial soft skills for collaborating remotely and analyzing data independently. These competencies ensure accurate reliability assessments and effective maintenance strategies, driving equipment uptime and operational efficiency from a remote setting.

How does a Remote Mechanical Reliability Engineer typically collaborate with onsite teams to address equipment issues?

Remote Mechanical Reliability Engineers frequently work alongside onsite engineers, maintenance staff, and operations teams to troubleshoot and resolve equipment reliability concerns. Collaboration is facilitated through virtual meetings, real-time data sharing platforms, and remote monitoring tools. Clear communication and effective documentation are key, as remote engineers must interpret equipment data and provide actionable recommendations while relying on local teams for physical inspections and repairs. Building strong relationships with onsite personnel is essential for timely problem-solving and continuous improvement.

What does a Remote Mechanical Reliability Engineer do?

A Remote Mechanical Reliability Engineer is responsible for ensuring the reliability and optimal performance of mechanical systems and equipment, often working from a remote location. They analyze data, identify potential failure points, develop maintenance strategies, and recommend design improvements to increase equipment uptime and lifespan. These engineers collaborate with onsite teams, use specialized software for monitoring, and may also support troubleshooting and root cause analysis remotely. Their work helps organizations minimize unplanned downtime and reduce maintenance costs.
What are the most commonly searched types of Mechanical Reliability Engineer jobs in California? The most popular types of Mechanical Reliability Engineer jobs in California are:
What cities in California are hiring for Remote Mechanical Reliability Engineer jobs? Cities in California with the most Remote Mechanical Reliability Engineer job openings:

Site Reliability Engineer

STN Inc

San Francisco, CA • Remote

$67.25 - $89.25/hr

Full-time

Re-posted 13 days ago


Job description

Site Reliability Engineer

Platform and software · shared across customers

Reports to: Director, Site Reliability

Location: Remote (US)

Department: Cloud Platform Engineering / SRE/Reliability

Position summary

The Site Reliability Engineer (SRE) owns reliability, observability, and incident response for the GPU One (GPUaaS) platform. The SRE defines and enforces SLOs aligned with contractual SLAs, builds the observability stack, and leads major incidents to resolution.

Key responsibilities
  • Define and operate Service Level Objectives (SLOs) aligned with customer SLAs

  • Build and maintain the observability stack including metrics, logs, traces, and alerting

  • Lead incident response and chair post-incident reviews

  • Drive automation to reduce toil and improve mean-time-to-recover (MTTR)

  • Author and maintain operational runbooks alongside the NOC

  • Manage on-call rotation, escalation paths, and incident-management tooling

  • Coordinate cross-functionally with NOC, Platform Engineering, and Network Engineering

  • Drive chaos engineering, game days, and reliability testing programs

  • Produce SLA performance reports in coordination with the SLA Manager

  • Mentor junior engineers and contribute to engineering culture

Required qualifications
  • 5+ years in SRE, DevOps, or production engineering roles

  • Strong programming skills in Go, Python, or both

  • Hands-on experience operating Kubernetes-based platforms at scale

  • Deep familiarity with observability tooling (Prometheus, Grafana, Datadog, OpenTelemetry)

  • Strong incident management experience including major-incident command

Preferred qualifications
  • GPU or HPC platform operational experience

  • Familiarity with SLA-driven customer environments and credit calculations

  • Experience with chaos engineering tools (Gremlin, Litmus, or similar)

  • Published SRE content or contributions