2

Remote Reliability Manager Jobs in Leander, TX (NOW HIRING)

Sr. SRE Platform Architect

Austin, TX · On-site +1

$56.50 - $75/hr

Coordinate with Security - joint ownership of vulnerability management, exposure management, joint ... Qualifications * 10+years of production SRE / platform-engineering / infra-architecture, including ...

This includes remote caching, remote execution, target identification and selection, etc. * Operate ... on pipeline reliability * Reduce developer cycle time through pre-commit tooling, caching ...

Software Engineer, CI/CD

Austin, TX · Remote

$123K - $216K/yr

This includes remote caching, remote execution, target identification and selection, etc. * Operate ... on pipeline reliability * Reduce developer cycle time through pre-commit tooling, caching ...

Showing results 21-40

Remote Reliability Manager information

See Leander, TX salary details

$59.2K

$112.3K

$161K

How much do remote reliability manager jobs pay per year?

As of Aug 15, 2026, the average yearly pay for remote reliability manager in Leander, TX is $112,260.00, according to ZipRecruiter salary data. Most workers in this role earn between $90,300.00 and $133,800.00 per year, depending on experience, location, and employer.

What does a Remote Reliability Manager do?

A Remote Reliability Manager oversees the maintenance and reliability of equipment and systems in a remote or distributed environment. They develop maintenance strategies, analyze performance data, and coordinate teams to ensure operational efficiency and minimize downtime, often using tools like CMMS software and data analysis skills.

What is the difference between Remote Reliability Manager vs Remote Maintenance Engineer?

AspectRemote Reliability ManagerRemote Maintenance Engineer
CredentialsEngineering degree, certifications in reliability or asset managementEngineering degree, certifications in maintenance or technical skills
Work EnvironmentOversees reliability strategies remotely, collaborates with teamsPerforms maintenance tasks remotely or on-site, technical troubleshooting
Industry UsageUsed across manufacturing, energy, and industrial sectorsCommon in manufacturing, utilities, and industrial facilities
Search IntentComparing reliability management roles with maintenance rolesLooking for maintenance-focused remote engineering jobs

The Remote Reliability Manager focuses on developing and implementing strategies to improve asset reliability remotely, often overseeing teams and analyzing data. In contrast, the Remote Maintenance Engineer handles technical maintenance tasks, troubleshooting, and repairs remotely or on-site. Both roles require engineering credentials and are prevalent in industrial sectors, but their core responsibilities differ—one emphasizes strategic reliability management, the other technical maintenance execution.

What are popular job titles related to Remote Reliability Manager jobs in Leander, TX?

For Remote Reliability Manager jobs in Leander, TX, the most frequently searched job titles are:

What job categories do people searching Remote Reliability Manager jobs in Leander, TX look for?

The top searched job categories for Remote Reliability Manager jobs in Leander, TX are:

What cities near Leander, TX are hiring for Remote Reliability Manager jobs?

Cities near Leander, TX with the most Remote Reliability Manager job openings:

Infographic showing various Remote Reliability Manager job openings in Leander, TX as of August 2026, with employment types broken down into 88% Full Time, 11% Part Time, and 1% Contract. Highlights an 84% Physical, 3% Hybrid, and 13% Remote job distribution, with an average salary of $112,260 per year, or $54 per hour.

Sr. SRE Platform Architect

BitDeer

Austin, TX • On-site, Remote

$56.50 - $75/hr

Full-time

Re-posted 23 days ago


Job description

About Bitdeer Technologies Group
Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure.
Bitdeer is committed to providing comprehensive Bitcoin mining solutions for its customers and building AI computational infrastructure to support the AI revolution. Bitdeer handles complex processes involved in computing such as equipment procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer also offers advanced cloud capabilities to customers with high demand for artificial intelligence.
Headquartered in Singapore, Bitdeer has deployed data centers across multiple countries, including the United States, Norway, Bhutan, and Ethiopia.
Position Overview
Bitdeer is seeking a visionary and hands-on Cloud SRE Architect to lead the design, development, and evolution of our next-generation public cloud platform. This role will oversee the end-to-end architecture across CPU, GPU, RDS, storage, networking, serverless, and AI services, ensuring global scalability, reliability, and performance. The ideal candidate is a strategic thinker with deep technical expertise in cloud infrastructure, platform engineering and AI systems, capable of bridging architecture vision with real-world engineering execution. You will collaborate closely with cross-functional teams and global partners to define our cloud technology roadmap, optimize multi-region deployments, and deliver world-class infrastructure and platform solutions that power large-scale AI and enterprise workloads.
Key Responsibilities
Own the end-to-end architecture of the NeoCloud SRE platform - the substrate that observes, protects, and operates a multi-region GPU rental fleet across self-built and OEM-rented data centers. You are the single point of architectural accountability across the platform's ~57 bounded contexts, ~12 frameworks, and three operational tiers (Edge DC → Regional Controller → Global Hub).
This role is for someone who writes the design, defends it under review, and shepherds it through the engineering squads that build it.
What You'll Do
  1. Write and maintain the platform architecture document - keep the design coherent across all sections, frameworks, and tiers. The current document is your starting point.
  2. Review every framework-level change - new bounded context, new plugin kind, tier-deployment shift, schema change, naming change, cross-context contract change. Architecture changes ride GitOps PRs like any other artifact.
  3. Set design invariants - residency rules (raw data stays in Region), Tier 2 self-sufficiency budget (≥ 24 h), survival-uplink contracts, naming conventions, SLO catalogues, redaction-at-boundary rules.
  4. Run the plugin framework - every extension uses one uniform contract (Common + Domain manifest, lifecycle, observability). You author and evolve this contract.
  5. Decide tier placement - what runs at Edge DC vs Regional Controller vs Global Hub, with data-residency / compliance / availability tradeoffs explicit.
  6. Coordinate with cloud-service teams and tenants - they author plugins, SDKs, dashboards, agent recipes that ride the platform. You set the contracts they consume.
  7. Coordinate with Security - joint ownership of vulnerability management, exposure management, joint operations. Security owns policy and risk acceptance; you own the operational mechanisms they ride.
  8. Pre-flight roadmap items - for any new capability, produce a one-page design that fits the existing layered model (L1-L6), tier topology, naming conventions, and extension contracts before implementation starts.
  9. Defend the design under review - say no to scope creep, special-case workarounds, and one-off integrations that don't fit the framework model. Say yes when a new plugin kind is genuinely needed.

Qualifications
  • 10+years of production SRE / platform-engineering / infra-architecture, including ≥ 3 years at architect level.
  • Hands-on with GPU / AI-compute infrastructure - NVIDIA GPU ops (DCGM, MIG, vGPU, NVLink/NVSwitch, XID semantics, NCCL), InfiniBand or RoCE fabrics (subnet manager, fabric partitioning, optical health), HPC storage (Lustre, NetApp/Pure/DDN/VAST, NVMe-oF).
  • Multi-region observability at scale - metrics / logs / traces / profiles / analytics-lake substrate; recording rules, MWMBR burn-rate alerting, SLI/SLO discipline.
  • Cluster platforms - first-hand experience with Kubernetes (control plane + GPU Operator + topology-aware scheduling) AND at least one of Slurm / Volcano / Kueue / Ray / KubeRay.
  • Data-center operations - ZTP, BMC/IPMI/Redfish, BIOS/firmware lifecycle, RMA, multi-vendor OEM management (self-built + leased DC mix).
  • Strong DDD instincts - bounded contexts, public contracts, no shared databases, one-context-one-repo discipline.
  • Plugin framework design - you have built (or substantively contributed to) a real extension framework with a uniform manifest + lifecycle.
  • Writing fluency - you can author and maintain a multi-thousand-line architecture document under review without it drifting; you can also write a one-pager an executive will read.
  • Cross-team operating tempo - design reviews, runbook authorship, on-call shadowing, post-mortem facilitation
  • Hyperscale or NeoCloud experience
  • BS/MS in Computer Science or similar

Bitdeer is committed to providing equal employment opportunities in accordance with country, state, and local laws. Bitdeer does not discriminate against employees or applicants based on conditions such as race, color, gender identity and/or expression, sexual orientation, marital and/or parental status, religion, political opinion, nationality, ethnic background or social origin, social status, disability, age, indigenous status, and union.