2

Remote Reliability Engineer Jobs in Miami, FL (NOW HIRING)

Software Engineer, Site Reliability

Miami, FL · On-site +1

$54.50 - $72.50/hr

Role As a Software Engineer working on Site Reliability at OpenEvidence, you will build and harden the mission-critical infrastructure powering our medical AI platform used by healthcare providers ...

Devops Automation Engineer

Miami, FL · Remote

$54 - $74/hr

This is a remote position, however it is US only. We do not offer relocation or VISA. Our contract ... Who You Are: * A strong platform, DevOps, SRE, or infrastructure automation engineer with ...

Devops Automation Engineer

Miami, FL · On-site +1

$50.50 - $69/hr

This is a remote position, however it is US only. We do not offer relocation or VISA. Our contract ... Who You Are: * A strong platform, DevOps, SRE, or infrastructure automation engineer with ...

Software Engineer

Miami, FL · On-site +1

$128K - $184K/yr

This role is remote and can be performed from anywhere in the U.S. Meet the Team The Cryptographic ... Participate in upgrades, capacity planning, hardening, and reliability work for a production ...

DevOps Engineer

Miami, FL · Remote

$54 - $74/hr

Participate in evaluating and integrating new technologies to enhance the scalability, reliability ... We offer a hybrid work schedule to perfectly combine the benefits of remote work and the essential ...

Senior DevOps Engineer

Miami, FL · Remote

$140K - $170K/yr

Define and own our reliability practices: SLOs, incident response, post-mortems, and the production ... Fully remote * Competitive salary and equity package * Health, dental, and vision insurance ...

Senior Software Engineer (Remote)

Miami, FL · On-site +1

$117K - $154K/yr

Practical understanding of reliability patterns: retries, timeouts, idempotency, fallback ... the Engineering Manager. No direct people management responsibilities. Location Remote. USA ...

Senior Software Engineer (Remote)

Miami, FL · Remote

$125K - $165K/yr

Practical understanding of reliability patterns: retries, timeouts, idempotency, fallback ... the Engineering Manager. No direct people management responsibilities. Location Remote. USA ...

We are currently looking for a Backend Engineer for a 100% remote position on a large federal ... Optimize backend service performance, reliability, monitoring, logging, and observability across ...

next page

Showing results 1-20

Remote Reliability Engineer information

See Miami, FL salary details

$58.3K

$112.8K

$134.9K

How much do remote reliability engineer jobs pay per year?

As of Aug 5, 2026, the average yearly pay for remote reliability engineer in Miami, FL is $112,834.00, according to ZipRecruiter salary data. Most workers in this role earn between $98,000.00 and $123,400.00 per year, depending on experience, location, and employer.

What is the difference between Remote Reliability Engineer vs Remote Site Reliability Engineer?

AspectRemote Reliability EngineerRemote Site Reliability Engineer
CredentialsTypically requires certifications like AWS Certified Solutions Architect, Linux Foundation certificationsSimilar credentials, often with additional focus on site-specific tools and monitoring
Work EnvironmentPrimarily remote, focusing on cloud infrastructure and system reliabilityRemote with some on-site responsibilities, focusing on infrastructure and operational stability
Industry UsageUsed across tech, cloud providers, SaaS companiesCommon in data centers, cloud providers, and large enterprise IT
Search & Comparison IntentOften compared due to overlapping roles in system reliability and cloud infrastructureCompared for on-site vs remote operational responsibilities

The main difference is that Remote Reliability Engineers focus on cloud and system reliability remotely, while Remote Site Reliability Engineers may have some on-site duties related to infrastructure. Both roles require similar skills and certifications but differ in their work environment and specific responsibilities.

What are the key skills and qualifications needed to thrive as a remote reliability engineer?

To thrive as a Remote Reliability Engineer, you need a strong background in systems engineering, software development, and infrastructure management, often supported by a degree in computer science or a related field. Proficiency with cloud platforms (such as AWS, Azure, or GCP), monitoring tools (like Prometheus, Grafana), and relevant certifications (e.g., AWS Certified DevOps Engineer) is highly valuable. Excellent problem-solving, communication, and collaboration skills are crucial for working effectively across distributed teams and responding to incidents. These abilities ensure system reliability, quick incident resolution, and seamless remote teamwork, which are vital for maintaining high service uptime and user satisfaction.

How do remote reliability engineers typically collaborate with on-site teams to address urgent technical issues?

Remote Reliability Engineers often utilize a combination of video conferencing, instant messaging, and collaborative monitoring tools to stay closely connected with on-site teams. When urgent technical issues arise, they participate in real-time troubleshooting sessions, analyze system logs remotely, and may guide on-site staff through step-by-step resolution procedures. Building strong communication channels and regular check-ins are essential to ensure swift and effective collaboration, even across different time zones. This structure allows Remote Reliability Engineers to contribute significantly to system uptime while working from a distance.

What is a remote reliability engineer?

A Remote Reliability Engineer is a professional who works from a remote location to ensure that systems, applications, or infrastructure are reliable, available, and performing well. Their responsibilities typically include monitoring system health, diagnosing issues, implementing preventative measures, and collaborating with teams to improve system reliability. They often use tools for automation, incident response, and performance monitoring, all while working offsite. This role is critical in minimizing downtime and ensuring a smooth user experience, especially for companies with complex technical environments. Remote Reliability Engineers must have strong problem-solving skills and be proficient in cloud technologies, automation, and incident management.
What are the most commonly searched types of Reliability Engineer jobs in Miami, FL? The most popular types of Reliability Engineer jobs in Miami, FL are:
What are popular job titles related to Remote Reliability Engineer jobs in Miami, FL? For Remote Reliability Engineer jobs in Miami, FL, the most frequently searched job titles are:
What job categories do people searching Remote Reliability Engineer jobs in Miami, FL look for? The top searched job categories for Remote Reliability Engineer jobs in Miami, FL are:
What cities near Miami, FL are hiring for Remote Reliability Engineer jobs? Cities near Miami, FL with the most Remote Reliability Engineer job openings:
Infographic showing various Remote Reliability Engineer job openings in Miami, FL as of July 2026, with employment types broken down into 93% Full Time, 4% Part Time, and 3% Contract. Highlights an 89% Physical, 4% Hybrid, and 7% Remote job distribution, with an average salary of $112,834 per year, or $54.2 per hour.

Site Reliability Engineering Manager (Remote)

NationsBenefits, LLC

Plantation, FL • Remote

$58.25 - $77.50/hr

Full-time

Medical, PTO

Posted 7 days ago


NationsBenefits rating

6.6

Company rating: 6.6 out of 10

Based on 16 frontline employees who took The Breakroom Quiz

309th of 481 rated business services


Job description

NationsBenefits is recognized as one of the fastest-growing companies in America and a Healthcare Fintech provider of supplemental benefits, flex cards, and member engagement solutions. We partner with managed care organizations to provide innovative healthcare solutions that drive growth, improve outcomes, reduce costs, and bring value to their members.Through our comprehensive suite of innovative supplemental benefits, fintech payment platforms, and member engagement solutions, we help health plans deliver high-quality benefits to their members that address the social determinants of health and improve member health outcomes and satisfaction.Our compliance-focused infrastructure, proprietary technology systems, and premier service delivery model allow our health plan partners to deliver high-quality, value-based care to millions of members.We offer a fulfilling work environment that attracts top talent and encourages all associates to contribute to delivering premier service to internal and external customers alike. Our goal is to transform the healthcare industry for the better! We provide career advancement opportunities from within the organization across multiple locations in the US, South America, and India.
Location: Remote (US-based candidates only)
Manager, Site Reliability Engineering (SRE)Position Overview

We are seeking a Manager, Site Reliability Engineering (SRE) to lead our US-based SRE team and drive operational excellence across our production platforms.

This is a player-coach leadership role that combines people management with hands-on technical leadership. You will mentor and grow a team of Site Reliability Engineers while actively participating in major incident response, reliability initiatives, and operational reviews. The role is a key part of our global follow-the-sun support model and requires close collaboration with SRE leadership in India.

Key ResponsibilitiesTeam Leadership & Development
  • Lead, mentor, and develop a US-based team of Site Reliability Engineers.
  • Conduct regular 1:1s, performance reviews, and career development discussions.
  • Own hiring, onboarding, and retention efforts as the team scales.
  • Foster a culture of ownership, blameless postmortems, and continuous improvement.
Operational Excellence & Incident Management
  • Lead day-to-day production operations and ensure timely incident triage, resolution, and escalation.
  • Serve as an escalation point and incident commander for major production incidents.
  • Drive problem management and root cause analysis processes.
  • Carry PagerDuty on-call escalation responsibilities for critical issues.
  • Track and report operational KPIs, SLAs, and SLOs, including availability, MTTR, and incident trends.
Reliability & Automation
  • Improve system reliability, observability, and resilience using Datadog and related tooling.
  • Drive automation, self-healing capabilities, and runbook maturity.
  • Partner with Development, DevOps, DevSecOps, and Engineering teams to embed reliability into the SDLC.
  • Contribute hands-on to tooling, automation, and technical reviews as needed.
Collaboration & Global Alignment
  • Coordinate closely with SRE leadership in India to ensure seamless follow-the-sun coverage.
  • Represent the US SRE organization in cross-functional planning and operational reviews.
  • Communicate effectively with both technical and non-technical stakeholders.
Documentation & Compliance
  • Maintain high-quality documentation for incidents, postmortems, runbooks, and operational procedures.
  • Ensure adherence to healthcare and fintech compliance standards, including HIPAA, PCI DSS, SOC 2, ISO 27001, and HITRUST.
Required Qualifications
  • 58 years of experience in Site Reliability Engineering, DevOps, Production Support, or Platform Engineering.
  • 12+ years of experience leading, mentoring, or managing engineers.
  • Demonstrated success operating in a player-coach leadership model.
  • Strong hands-on experience with production incident management and escalation processes.
  • Proficiency with Datadog or similar observability platforms.
  • Hands-on experience with Kubernetes and Docker in production environments.
  • Strong scripting or programming skills in PowerShell, Bash, Python, Java, or C#.
  • Experience with Helm, CI/CD pipelines, and deployment automation.
  • Working knowledge of ITIL processes and Agile methodologies.
  • Experience working with SQL, MySQL, or NoSQL databases.
  • Excellent communication and stakeholder management skills.
  • Willingness to participate in PagerDuty on-call escalation and work within a global follow-the-sun operating model.
Preferred Qualifications
  • Experience with cloud platforms such as AWS, Azure, or GCP.
  • Experience building or scaling SRE teams and on-call programs.
  • Experience defining and managing SLOs, SLIs, and error budgets.
  • Prior experience in the healthcare or fintech industry.
  • Knowledge of security and compliance frameworks relevant to regulated environments.
Why Join NationsBenefits?
  • Competitive compensation and comprehensive benefits.
  • Unlimited PTO.
  • Fully remote work environment (US-based).
  • Opportunity to lead and grow a high-impact SRE organization.
  • Exposure to modern cloud-native technologies and large-scale reliability challenges.
  • Collaborative culture focused on innovation, learning, and continuous improvement.
  • Meaningful work that directly impacts healthcare technology and millions of members.
Ideal Candidate

We are looking for a technically strong SRE leader who enjoys building teams, improving operational maturity, and remaining hands-on during critical production events. The ideal candidate combines leadership, systems thinking, and automation expertise to help scale reliability practices across a fast-growing Healthcare FinTech organization.
NationsBenefits is an Equal Opportunity Employer.


What NationsBenefits employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom