2

Remote Reliability Manager Jobs in Dripping Springs, TX

Senior Software Engineer (Remote)

Austin, TX · Remote

$121K - $160K/yr

We partner closely with Design, Product Management, and other teams to create solutions for our ... Review and test code to identify problems and improve scalability, speed, and reliability.

As a premier hard money lender, we are known for speed, reliability, and deep market expertise. We ... Assign each new hire a mentor, and make sure remote reps get the same structure and visibility as ...

Senior Frontend Engineer

Austin, TX · Remote

$121K - $167K/yr

... Remote Control system that powers teleoperation and fleet management for autonomous cars and ... We care about performance, reliability, and thoughtful UX design - all in service of safe and ...

Showing results 41-60

Remote Reliability Manager information

See Dripping Springs, TX salary details

$65.8K

$124.7K

$178.8K

How much do remote reliability manager jobs pay per year?

As of Aug 22, 2026, the average yearly pay for remote reliability manager in Dripping Springs, TX is $124,692.00, according to ZipRecruiter salary data. Most workers in this role earn between $100,300.00 and $148,600.00 per year, depending on experience, location, and employer.

What is the difference between Remote Reliability Manager vs Remote Maintenance Engineer?

AspectRemote Reliability ManagerRemote Maintenance Engineer
CredentialsEngineering degree, certifications in reliability or asset managementEngineering degree, certifications in maintenance or technical skills
Work EnvironmentOversees reliability strategies remotely, collaborates with teamsPerforms maintenance tasks remotely or on-site, technical troubleshooting
Industry UsageUsed across manufacturing, energy, and industrial sectorsCommon in manufacturing, utilities, and industrial facilities
Search IntentComparing reliability management roles with maintenance rolesLooking for maintenance-focused remote engineering jobs

The Remote Reliability Manager focuses on developing and implementing strategies to improve asset reliability remotely, often overseeing teams and analyzing data. In contrast, the Remote Maintenance Engineer handles technical maintenance tasks, troubleshooting, and repairs remotely or on-site. Both roles require engineering credentials and are prevalent in industrial sectors, but their core responsibilities differ—one emphasizes strategic reliability management, the other technical maintenance execution.

What job categories do people searching Remote Reliability Manager jobs in Dripping Springs, TX look for?

The top searched job categories for Remote Reliability Manager jobs in Dripping Springs, TX are:

What cities near Dripping Springs, TX are hiring for Remote Reliability Manager jobs?

Cities near Dripping Springs, TX with the most Remote Reliability Manager job openings:

Senior Software Engineer [REMOTE]

Upbound - Job Posting

Austin, TX • Remote

$121K - $160K/yr

Full-time

Re-posted 23 days ago


Job description

Upbound is hiring a Senior Software Engineer to help us build and operate Upbound Spaces, the multiple control plane management software at the heart of the Upbound Platform. As part of the Spaces team, you will help us scale Upbound to reliably support thousands of control planes, while also extending enterprise control plane management and operations both in the cloud and on premises. Our team is expanding, and this is the perfect opportunity for you to make a significant engineering impact in both development and production operations.

What You'll Do
  • Actively build and operate Upbound Spaces in production, troubleshooting and resolving issues across multi-tenant SaaS environments, as well as contributing to Upbound's open-source projects, including Crossplane.
  • Take ownership of building features in high demand by Upbound's customers and deliver new functionality that will delight and amaze our users.
  • Investigate and debug complex issues in customer environments, including multi-control plane scenarios, resource reconciliation problems, and performance bottlenecks.
  • Communicate through thoughtful and thorough design documents for new initiatives and detailed post-incident reviews that drive system improvements.
  • Support the full project lifecycle for highly scalable and reliable services running in a cloud environment - discovery, analysis, architecture, design, review, documentation, building, migration, automation, deployment, production-readiness, and ongoing operational support.
  • Write and maintain Go code that interfaces with the Kubernetes API, such as operators, controllers, add-ons, etc., with a focus on observability, debuggability, and operational excellence.
  • Deploy, manage, and troubleshoot our Kubernetes services in production, using metrics, logs, and traces to identify and resolve issues quickly.
  • Build and maintain operational tooling for debugging customer environments, analyzing control plane health, and automating incident response.
  • Author documentation, user guides, runbooks, and blog posts to support and promote new features that you release.
  • Support the software release cycle for Spaces self-hosted distributions, including diagnosing issues in customer-managed deployments.
  • Participate in on-call rotation to support Upbound Cloud, responding to incidents and driving them to resolution.
What You'll Bring
  • Have experience operating production cloud services at scale: monitoring, alerting, incident response, post-mortems, and continuous improvement of service reliability.
  • Have strong debugging skills across distributed systems, including experience with observability tools (Prometheus, Grafana, OpenTelemetry, distributed tracing) and techniques for diagnosing issues in production environments.
  • Have experience building and operating controllers that interact with the Kubernetes API server, including troubleshooting reconciliation loops, managing API rate limits, and optimizing controller performance.
  • Are comfortable working directly with customers to understand, reproduce, and resolve complex technical issues in their environments.
  • Take responsibility and ownership for solving problems even if they are outside your lane, especially during incidents affecting customer workloads.
  • Demonstrate excellence in your work, constantly trying to improve your skills and the operational posture of the systems you build.
  • Have empathy for customers and keep them in mind as you build solutions, understanding that reliability and debuggability are features.
  • Realize the importance of clear communication and effective collaboration to work as a team, deliver great results, and support customers through technical challenges.
  • Help create a safe environment where everyone can contribute, learn from failures, share on-call knowledge, and help each other grow as operators and engineers.

 #LI-REMOTE