2

Remote Reliability Centered Maintenance Jobs in New York

Senior SRE Engineer

New York, NY · On-site +1

$62.25 - $82.75/hr

New York, NY (Hybrid) / Remote Department: Engineering The Role Flowcode is seeking a Senior Site ... Experience maintaining and scaling CI/CD automation using GitHub Actions within collaborative ...

Senior SRE Engineer

New York, NY · On-site +1

$62.25 - $82.75/hr

New York, NY (Hybrid) / Remote Department: Engineering The Role Flowcode is seeking a Senior Site ... Experience maintaining and scaling CI/CD automation using GitHub Actions within collaborative ...

Staff Site Reliability Engineer

New York, NY · Remote

$62.25 - $82.75/hr

Because our systems directly influence health outcomes for millions of patients, maintaining the ... remote work and occasional travel to HQ. What you will do: * Own the Reliability Strategy:

Senior Site Reliability Engineer

New York, NY · Remote

$62.25 - $82.75/hr

... remote work and occasional travel to HQ. What you will do: * Run the Machine: Own the end-to-end ... Build and maintain the monitoring, alerting, and observability systems that let us detect and ...

New

Remote Job Summary: Join our team as a Senior Database Reliability Engineer, where you'll play a ... maintenance tasks and promote operational excellence across all database environments. • ...

next page

Showing results 1-20

Remote Reliability Centered Maintenance information

What is the difference between Remote Reliability Centered Maintenance vs Remote Maintenance Technician?

AspectRemote Reliability Centered MaintenanceRemote Maintenance Technician
CertificationsReliability, RCM, or maintenance planning certificationsHVAC, electrical, or mechanical certifications
Work EnvironmentIndustrial plants, manufacturing facilities, remote monitoringField sites, industrial equipment, remote or on-site
Industry UsageAsset management, reliability engineering, maintenance planningEquipment repair, troubleshooting, routine maintenance

Remote Reliability Centered Maintenance focuses on optimizing asset reliability through analysis and planning, often involving remote monitoring and strategic decision-making. In contrast, Remote Maintenance Technicians perform hands-on repairs and routine maintenance tasks, often on-site or remotely troubleshooting equipment. Both roles are essential in industrial settings but differ mainly in scope, responsibilities, and required skills.

What are the most commonly searched types of Reliability Centered Maintenance jobs in New York?

The most popular types of Reliability Centered Maintenance jobs in New York are:

What are popular job titles related to Remote Reliability Centered Maintenance jobs in New York?

For Remote Reliability Centered Maintenance jobs in New York, the most frequently searched job titles are:

What job categories do people searching Remote Reliability Centered Maintenance jobs in New York look for?

The top searched job categories for Remote Reliability Centered Maintenance jobs in New York are:

What cities in New York are hiring for Remote Reliability Centered Maintenance jobs?

Cities in New York with the most Remote Reliability Centered Maintenance job openings:

Senior Site Reliability Engineer (SRE

Gov Services Hub

New York, NY • On-site, Remote

$62.25 - $82.75/hr

Full-time

Posted 9 days ago


Job description

Title: Senior Site Reliability Engineer (SRE)
Location: Remote

About
January

At
January, we’re transforming the lives of borrowers by bringing humanity to
consumer finance. Our data-driven products empower financial institutions to
streamline collections and help borrowers regain financial stability and
control over their lives. We’re not just expanding access to credit — we’re
restoring dignity and paving the way for millions to achieve financial freedom.

About
the Role

As a Senior
Site Reliability Engineer (SRE)
, you will establish SRE practices from the
ground up — ensuring reliability, scalability, and performance as January
scales from thousands to millions of borrowers. You’ll architect resilient
infrastructure, design modern observability solutions, and build sustainable
on-call processes that evolve with our rapid growth.

Your work
will directly address scaling challenges including database optimization, async
workflow infrastructure, and data pipeline reliability — enabling the
engineering team to ship confidently and efficiently.

Key
Responsibilities

  • Lead incident response and develop
    sustainable on-call practices, including runbooks, blameless postmortems,
    and continuous improvement to reduce MTTR.
  • Build and maintain self-service observability tools
    (Datadog, Prometheus, ELK) for proactive monitoring and troubleshooting.
  • Create and maintain Infrastructure as Code (IaC) using Terraform or CloudFormation for consistent, secure
    AWS environments.
  • Partner with development teams to architect
    resilient, scalable infrastructure
     for critical components like
    databases, networking, async workflows, and data pipelines.
  • Design and implement robust CI/CD pipelines (GitHub
    Actions) with advanced deployment strategies (blue/green, canary).
  • Drive best practices in reliability and
    performance early in the design phase to future-proof January’s systems.

Required
Skills & Experience

  • Proven experience leading incident response and
    postmortem processes for high-availability production systems.
  • Deep expertise in designing highly available
    architectures
     (EC2, Fargate, auto-scaling, health checks,
    graceful degradation).
  • Strong experience with AWS cloud infrastructure and IaC tools (Terraform, CloudFormation).
  • Hands-on experience with CI/CD automation using GitHub Actions or equivalent tools.
  • Proficiency in observability and monitoring stacks (Datadog,
    Prometheus, ELK
    ).
  • Solid scripting/programming skills in Python (for
    automation, tooling, and debugging).
  • Excellent communication and documentation skills, with
    the ability to collaborate across engineering and platform teams.




Requirements

Tools
& Technologies

  • Cloud: AWS
  • IaC: Terraform,
    CloudFormation
  • CI/CD: GitHub
    Actions
  • Monitoring: Datadog,
    Prometheus, ELK
  • Languages: Python
  • Infrastructure: EC2,
    Fargate

Additional
Details

  • Remote role (NYC-based preferred for hybrid
    collaboration).
  • Opportunity to build and own the entire SRE practice for
    a growing FinTech startup.
  • Fast-paced, innovative environment working on
    AI-forward consumer finance products.