2

Overnight Site Reliability Engineer Remote Jobs in Florida

Site Reliability Engineer (SRE)

Jupiter, FL · Remote

$55.75 - $74/hr

We are looking for a Site Reliability Engineer to help maintain the health, stability, and ... This is a remote position. * Flexible Vacation . Never get denied a vacation request ever again.

Site Reliability Engineer (SRE)

Jupiter, FL · On-site +1

$55.75 - $74/hr

We are looking for a Site Reliability Engineer to help maintain the health, stability, and ... This is a remote position. * Flexible Vacation . Never get denied a vacation request ever again.

If not, this role is fully remote. We do not restrict applicants based on job site or posting location. Job Title: Senior Site Reliability Engineer (SRE) Location: Open (U.S.-based). No job site ...

If not, this role is fully remote. We do not restrict applicants based on job site or posting location. Job Title: Senior Site Reliability Engineer (SRE) Location: Open (U.S.-based). No job site ...

Software Engineer, Site Reliability

Miami, FL · On-site +1

$54.50 - $72.50/hr

Role As a Software Engineer working on Site Reliability at OpenEvidence, you will build and harden the mission-critical infrastructure powering our medical AI platform used by healthcare providers ...

Devops Automation Engineer

Miami, FL · On-site +1

$50.50 - $69/hr

This is a remote position, however it is US only. We do not offer relocation or VISA. Our contract ... Who You Are: * A strong platform, DevOps, SRE, or infrastructure automation engineer with ...

next page

Showing results 1-20

Overnight Site Reliability Engineer Remote information

What is an overnight site reliability engineer?

An Overnight Site Reliability Engineer (SRE) is a professional responsible for ensuring the reliability, performance, and uptime of software systems during overnight or off-peak hours, typically working remotely. Their main tasks include monitoring system health, responding to incidents, troubleshooting outages, and implementing fixes to maintain service availability. SREs also work to automate processes, improve system resilience, and collaborate with other engineering teams to prevent future issues. Working overnight ensures that critical systems remain operational and issues are addressed promptly, even outside of standard business hours.

What skills and qualifications are needed to be an overnight site reliability engineer?

To thrive as an Overnight Site Reliability Engineer (Remote), you need strong expertise in systems administration, incident response, automation, and a solid background in computer science or related fields. Proficiency with monitoring tools (like Prometheus or Datadog), cloud platforms (such as AWS or GCP), scripting languages (Python, Bash), and certifications like AWS Certified SysOps Administrator are highly beneficial. Exceptional problem-solving skills, attention to detail, and effective remote communication help you excel in high-pressure overnight scenarios. These skills ensure system reliability, minimize downtime, and maintain seamless operations during critical off-hours.

What are the unique challenges and expectations for an overnight site reliability engineer working remotely?

As an Overnight Site Reliability Engineer working remotely, you'll often handle critical incidents that arise outside of standard business hours, so strong problem-solving skills and the ability to work independently are crucial. Communication is key, as you'll need to coordinate with team members in different time zones and document incidents clearly for seamless handoffs. You may also be tasked with proactive monitoring and maintenance activities during quieter periods, making self-motivation and attention to detail especially important. The role offers valuable exposure to high-impact issues and can accelerate your expertise in incident management and system reliability.

What is the difference between Overnight Site Reliability Engineer Remote vs Overnight DevOps Engineer Remote?

AspectOvernight Site Reliability Engineer RemoteOvernight DevOps Engineer Remote
Primary FocusEnsuring system reliability, uptime, and incident responseAutomating deployment, integration, and infrastructure management
Required SkillsMonitoring, incident management, scripting, system troubleshootingCI/CD pipelines, automation, cloud platforms, scripting
Work EnvironmentRemote, on-call shifts, collaboration with SRE teamsRemote, development and deployment focus, collaboration with development teams
CertificationsLinux, AWS, Google Cloud, or Azure certifications often preferredCloud certifications, Docker, Kubernetes, CI/CD tools

While both roles involve working remotely and require cloud and scripting skills, the Overnight Site Reliability Engineer Remote primarily focuses on maintaining system reliability and incident response, whereas the Overnight DevOps Engineer Remote emphasizes automation, deployment, and infrastructure management. Understanding these differences helps candidates align their skills with the right role.

What are the most commonly searched types of Site Reliability Engineer Remote jobs in Florida?

The most popular types of Site Reliability Engineer Remote jobs in Florida are:

What job categories do people searching Overnight Site Reliability Engineer Remote jobs in Florida look for?

The top searched job categories for Overnight Site Reliability Engineer Remote jobs in Florida are:

What cities in Florida are hiring for Overnight Site Reliability Engineer Remote jobs?

Cities in Florida with the most Overnight Site Reliability Engineer Remote job openings:

Site Reliability Engineer (SRE)

Rocket.net

Jupiter, FL • Remote

$55.75 - $74/hr

Full-time

PTO

Posted 22 days ago


Job description

Description

At Rocket.net, reliability, performance, and customer experience are at the center of everything we build. We are looking for a Site Reliability Engineer to help maintain the health, stability, and performance of our hosting platform while providing advanced technical support to our customers.

The Platform Operations team acts as a critical escalation layer between WordPress Support and Engineering. This role combines infrastructure operations, platform monitoring, troubleshooting, and advanced customer support.

As a Site Reliability Engineer, you will help ensure Rocket.net's servers, services, and customer environments are operating at the highest standards. You will assist WordPress Support Engineers with complex issues, support VIP customers with advanced technical requests, investigate platform-level problems, and work with internal teams to deliver fast and effective solutions.

Responsibilities

Platform Monitoring & Reliability

  • Monitor the health, availability, and performance of Rocket.net servers, services, and customer environments.
  • Proactively identify infrastructure issues, performance degradation, and potential service disruptions.
  • Investigate alerts and operational events to maintain platform stability.
  • Perform regular platform health checks and ensure critical systems are operating correctly.
  • Participate in incident response and coordinate troubleshooting during customer-impacting events.
  • Communicate platform issues, updates, and resolutions to relevant internal teams.

Advanced Technical Support & Escalations

  • Provide advanced technical support for VIP customers and customers with complex hosting-related issues.
  • Act as a senior escalation point for WordPress Support Engineers when issues require deeper technical investigation.
  • Troubleshoot complex issues involving servers, websites, networking, DNS, performance, caching, and hosting infrastructure.
  • Assist customers with advanced technical problems beyond standard WordPress troubleshooting.
  • Investigate and resolve issues involving server resources, application performance, connectivity, and platform behavior.
  • Work directly with customers when required to provide expert-level technical assistance.
  • Ensure escalated customer issues are handled with urgency, ownership, and clear communication.

Infrastructure Operations

  • Troubleshoot and maintain Linux-based production environments.
  • Investigate issues related to NGINX, Apache, PHP-FPM, MySQL/MariaDB, Redis, and other platform services.
  • Assist with server maintenance, configuration changes, and operational improvements.
  • Support security updates, system hardening, and infrastructure best practices.
  • Monitor resource usage and identify capacity or performance concerns.
  • Help improve monitoring, automation, and operational workflows.

Team Collaboration

  • Work closely with WordPress Support Engineers, Shift Leads, Site Reliability Engineers, and Engineering teams.
  • Provide technical guidance and knowledge sharing to Support teams.
  • Help create internal documentation, troubleshooting guides, and knowledge base articles.
  • Identify recurring issues and recommend improvements to reduce future incidents.
  • Participate in incident reviews and root cause analysis.

Requirements

  • 3+ years of experience in SRE, DevOps, Platform Engineering, or similar roles.
  • Strong experience troubleshooting Linux production environments.
  • Experience supporting customer-facing technical environments.
  • Strong understanding of web hosting technologies including NGINX, Apache, PHP-FPM, MySQL/MariaDB, and Redis.
  • Advanced troubleshooting skills across WordPress, servers, DNS, networking, and performance issues.
  • Experience with Linux command line (SSH).
  • Strong understanding of DNS, HTTP/HTTPS, SSL/TLS, CDN, and caching technologies.
  • Experience with Cloudflare, WAF, and web performance optimization.
  • Experience with monitoring tools and incident response processes.
  • Ability to troubleshoot complex issues independently and communicate technical solutions clearly.
  • Excellent written and verbal communication skills (English).
  • Ability to work under pressure during customer-impacting incidents.

Nice to Have

  • Experience supporting managed WordPress hosting platforms.
  • Experience handling VIP customers or enterprise-level support.
  • Experience with high-traffic websites and performance optimization.
  • Knowledge of Bash scripting, yum/dnf , or automation tools.
  • Familiarity with observability tools such as Nedata or Datadog or similar.
  • Experience with incident management and postmortems.

Benefits

So what are the advantages of working with a pretty amazing tech company?

  • Ability to work from anywhere in the world. You can travel and work in a new city every month. Work and travel without ever using your vacation time. This is a remote position.
  • Flexible Vacation. Never get denied a vacation request ever again.
  • Paid Education. We care about your career. We will help you gain new skills, and guide you on where you want to go.

Come work with a fun, driven, and amazing team at Rocket.net!

Rocket.net is an equal opportunity employer committed to diversity and inclusion. As a multicultural organization, we encourage individual achievement and recognize the strength of our diverse team.

Rocket.net is committed to providing accommodations for people with disabilities. If you require accommodation, we will work with you to meet your needs. Accommodation may be provided in all parts of the hiring process.

We would like to thank each applicant; however, only qualified candidates will be contacted for an interview.