2

Remote Reliability Engineer Jobs in New York (NOW HIRING)

AI Platform / SRE Lead

New York, NY · Remote

$62.25 - $82.75/hr

AI Platform / SRE Lead Location: Remote - USA Company Overview Glint Tech Solutions is a women-owned, global IT staffing and recruiting firm serving enterprise clients across the USA and Canada.

New

Senior SRE Engineer

New York, NY · On-site +1

$62.25 - $82.75/hr

New York, NY (Hybrid) / Remote Department: Engineering The Role Flowcode is seeking a Senior Site Reliability Engineer (SRE) to work on reliability and infrastructure efforts across our platforms.

Senior SRE Engineer

New York, NY · On-site +1

$62.25 - $82.75/hr

New York, NY (Hybrid) / Remote Department: Engineering The Role Flowcode is seeking a Senior Site Reliability Engineer (SRE) to work on reliability and infrastructure efforts across our platforms.

SRE Engineer

Newark, NJ · Remote

$58.25 - $77.50/hr

Join us as an Expert SRE to ensure operational excellence during the integration of merging systems. You'll design resilient infrastructure, automate incident response, and drive system performance ...

Site Reliability Engineer

New York, NY · On-site +1

$165K - $330K/yr

THE ROLE As a Site Reliability Engineer at Baseten, you'll define and codify the gold standards of day 2 operations for our ML infrastructure platform. You'll envision and build robust systems ...

Senior Site Reliability Engineer

New York, NY · Remote

$62.25 - $82.75/hr

... remote work and occasional travel to HQ. What you will do: * Run the Machine: Own the end-to-end ... scale in an SRE, DevOps, or platform engineering role * Deep expertise with Kubernetes and ...

You will lead and scale the SRE team to ensure our infrastructure stays ahead of demand, operates efficiently, and meets the needs of our growing healthcare customers. What You'll Do * Ensure 99.99 ...

Site Reliability Engineer

New York, NY · On-site +1

$120K - $160K/yr

As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and monitor production-critical infrastructure and ...

next page

Showing results 1-20

Remote Reliability Engineer information

What is a remote reliability engineer?

A Remote Reliability Engineer is a professional who works from a remote location to ensure that systems, applications, or infrastructure are reliable, available, and performing well. Their responsibilities typically include monitoring system health, diagnosing issues, implementing preventative measures, and collaborating with teams to improve system reliability. They often use tools for automation, incident response, and performance monitoring, all while working offsite. This role is critical in minimizing downtime and ensuring a smooth user experience, especially for companies with complex technical environments. Remote Reliability Engineers must have strong problem-solving skills and be proficient in cloud technologies, automation, and incident management.

What are the key skills and qualifications needed to thrive as a remote reliability engineer?

To thrive as a Remote Reliability Engineer, you need a strong background in systems engineering, software development, and infrastructure management, often supported by a degree in computer science or a related field. Proficiency with cloud platforms (such as AWS, Azure, or GCP), monitoring tools (like Prometheus, Grafana), and relevant certifications (e.g., AWS Certified DevOps Engineer) is highly valuable. Excellent problem-solving, communication, and collaboration skills are crucial for working effectively across distributed teams and responding to incidents. These abilities ensure system reliability, quick incident resolution, and seamless remote teamwork, which are vital for maintaining high service uptime and user satisfaction.

How do remote reliability engineers typically collaborate with on-site teams to address urgent technical issues?

Remote Reliability Engineers often utilize a combination of video conferencing, instant messaging, and collaborative monitoring tools to stay closely connected with on-site teams. When urgent technical issues arise, they participate in real-time troubleshooting sessions, analyze system logs remotely, and may guide on-site staff through step-by-step resolution procedures. Building strong communication channels and regular check-ins are essential to ensure swift and effective collaboration, even across different time zones. This structure allows Remote Reliability Engineers to contribute significantly to system uptime while working from a distance.

What is the difference between Remote Reliability Engineer vs Remote Site Reliability Engineer?

AspectRemote Reliability EngineerRemote Site Reliability Engineer
CredentialsTypically requires certifications like AWS Certified Solutions Architect, Linux Foundation certificationsSimilar credentials, often with additional focus on site-specific tools and monitoring
Work EnvironmentPrimarily remote, focusing on cloud infrastructure and system reliabilityRemote with some on-site responsibilities, focusing on infrastructure and operational stability
Industry UsageUsed across tech, cloud providers, SaaS companiesCommon in data centers, cloud providers, and large enterprise IT
Search & Comparison IntentOften compared due to overlapping roles in system reliability and cloud infrastructureCompared for on-site vs remote operational responsibilities

The main difference is that Remote Reliability Engineers focus on cloud and system reliability remotely, while Remote Site Reliability Engineers may have some on-site duties related to infrastructure. Both roles require similar skills and certifications but differ in their work environment and specific responsibilities.

What are the most commonly searched types of Reliability Engineer jobs in New York?

The most popular types of Reliability Engineer jobs in New York are:

What are popular job titles related to Remote Reliability Engineer jobs in New York?

For Remote Reliability Engineer jobs in New York, the most frequently searched job titles are:

What job categories do people searching Remote Reliability Engineer jobs in New York look for?

The top searched job categories for Remote Reliability Engineer jobs in New York are:

What cities in New York are hiring for Remote Reliability Engineer jobs?

Cities in New York with the most Remote Reliability Engineer job openings:

Infographic showing various Remote Reliability Engineer job openings in New York as of September 2026, with employment types broken down into 91% Full Time, and 9% Contract. Highlights an 100% Remote job distribution.

Senior Site Reliability Engineer (SRE

New York, NY • On-site, Remote

$62.25 - $82.75/hr

Full-time

Posted 18 days ago


Job description

Title: Senior Site Reliability Engineer (SRE)
Location: Remote

About
January

At
January, we’re transforming the lives of borrowers by bringing humanity to
consumer finance. Our data-driven products empower financial institutions to
streamline collections and help borrowers regain financial stability and
control over their lives. We’re not just expanding access to credit — we’re
restoring dignity and paving the way for millions to achieve financial freedom.

About
the Role

As a Senior
Site Reliability Engineer (SRE)
, you will establish SRE practices from the
ground up — ensuring reliability, scalability, and performance as January
scales from thousands to millions of borrowers. You’ll architect resilient
infrastructure, design modern observability solutions, and build sustainable
on-call processes that evolve with our rapid growth.

Your work
will directly address scaling challenges including database optimization, async
workflow infrastructure, and data pipeline reliability — enabling the
engineering team to ship confidently and efficiently.

Key
Responsibilities

  • Lead incident response and develop
    sustainable on-call practices, including runbooks, blameless postmortems,
    and continuous improvement to reduce MTTR.
  • Build and maintain self-service observability tools
    (Datadog, Prometheus, ELK) for proactive monitoring and troubleshooting.
  • Create and maintain Infrastructure as Code (IaC) using Terraform or CloudFormation for consistent, secure
    AWS environments.
  • Partner with development teams to architect
    resilient, scalable infrastructure
     for critical components like
    databases, networking, async workflows, and data pipelines.
  • Design and implement robust CI/CD pipelines (GitHub
    Actions) with advanced deployment strategies (blue/green, canary).
  • Drive best practices in reliability and
    performance early in the design phase to future-proof January’s systems.

Required
Skills & Experience

  • Proven experience leading incident response and
    postmortem processes for high-availability production systems.
  • Deep expertise in designing highly available
    architectures
     (EC2, Fargate, auto-scaling, health checks,
    graceful degradation).
  • Strong experience with AWS cloud infrastructure and IaC tools (Terraform, CloudFormation).
  • Hands-on experience with CI/CD automation using GitHub Actions or equivalent tools.
  • Proficiency in observability and monitoring stacks (Datadog,
    Prometheus, ELK
    ).
  • Solid scripting/programming skills in Python (for
    automation, tooling, and debugging).
  • Excellent communication and documentation skills, with
    the ability to collaborate across engineering and platform teams.




Requirements

Tools
& Technologies

  • Cloud: AWS
  • IaC: Terraform,
    CloudFormation
  • CI/CD: GitHub
    Actions
  • Monitoring: Datadog,
    Prometheus, ELK
  • Languages: Python
  • Infrastructure: EC2,
    Fargate

Additional
Details

  • Remote role (NYC-based preferred for hybrid
    collaboration).
  • Opportunity to build and own the entire SRE practice for
    a growing FinTech startup.
  • Fast-paced, innovative environment working on
    AI-forward consumer finance products.