2

Remote Devops Sre Engineer Jobs in New York (NOW HIRING)

DevOps/SRE

Manhattan, NY ยท On-site +1

$62.50 - $83/hr

The Role We are seeking an experienced New York-based DevOps / Site Reliability Engineer to join our DevOps team and own the reliability, stability, and operational support of our production systems.

SRE Engineer

Newark, NJ ยท Remote

$58.25 - $77.50/hr

Join us as an Expert SRE to ensure operational excellence during the integration of merging systems. You'll design resilient infrastructure, automate incident response, and drive system performance ...

Senior DevOps Engineer

New York, NY ยท On-site +1

$142K - $182K/yr

Senior DevOps Engineer NYC (Hybrid) or US Remote | Competitive + Equity About Avantos Avantos is ... The Role We're seeking a Senior DevOps Engineer / Site Reliability Engineer to own and evolve our ...

New

SRE - Remote

New York, NY ยท Remote

$62.25 - $82.75/hr

Glassbox is looking for an SRE to join our global Cloud team. We are Glassbox, and our mission is ... Work closely with the Cloud DevOps team to transition products from development to the production ...

DevOps Engineer

New York, NY ยท On-site +1

$160K - $200K/yr

Requirements REQUIRED QUALIFICATIONS: * 3-5+ years of experience in DevOps, Site Reliability Engineering (SRE), Platform Engineering, or similar roles. * Strong knowledge of cloud platforms (AWS, GCP ...

Site Reliability Engineer

New York, NY ยท On-site +1

$165K - $330K/yr

By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable ... 2 operations for our ML infrastructure platform. You'll envision and build robust systems ...

Location - We are flexible on remote working from home, if you are located in the USA and reside in ... Other duties as needed About You * 10+ years' experience in DevOps and/or Site Reliability ...

SRE Leader

New York, NY ยท Remote

$58.25 - $77.50/hr

Yet inside hospitals, where every second and every decision can affect a patient's life, operations ... You will lead and scale the SRE team to ensure our infrastructure stays ahead of demand, operates ...

Site Reliability Engineer

New York, NY ยท On-site +1

$120K - $160K/yr

As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and monitor production-critical infrastructure and ...

... operational overhead, and cost overruns due to over provisioning of workloads. Traditional ... Enable every developer, SRE, and platform engineer to operate Kubernetes like an expert, without ...

... operational overhead, and cost overruns due to over provisioning of workloads. Traditional ... Enable every developer, SRE, and platform engineer to operate Kubernetes like an expert, without ...

Solutions Engineer (Pre-Sales)

New York, NY ยท Remote

$220K - $250K/yr

... remote position, but candidates must be located within the US Eastern Time Zone (EST/EDT ... DevOps, SRE, Platform Engineering, or Security Engineering. * Strong hands-on technical ability ...

next page

Showing results 1-20

Remote Devops Sre Engineer information

What are the key skills and qualifications needed to thrive in the Remote Devops Sre Engineer position, and why are they important?

To thrive as a Remote DevOps SRE Engineer, you need deep expertise in systems administration, cloud platforms (such as AWS, Azure, or Google Cloud), automation, and infrastructure-as-code, typically backed by experience in scripting languages and a relevant technical degree. Proficiency with tools like Docker, Kubernetes, Terraform, Jenkins, and monitoring solutions, as well as certifications such as AWS Certified DevOps Engineer or Google Professional SRE, is highly valued. Strong problem-solving skills, effective communication, and the ability to collaborate virtually are key soft skills in a remote setting. These capabilities ensure reliability, scalability, and seamless operations of critical systems in distributed environments.

What are the primary responsibilities and challenges faced by a Remote DevOps SRE Engineer day-to-day?

Remote DevOps SRE Engineers are responsible for maintaining the availability, scalability, and performance of production systems, often using automation to streamline deployments and incident response. A typical day may involve refining CI/CD pipelines, responding to system alerts, conducting root cause analysis of outages, and collaborating with software engineers to improve system reliability. One of the main challenges is proactively identifying and mitigating potential infrastructure issues before they impact customers, all while communicating effectively across remote teams. The role often requires balancing the implementation of new technologies with maintaining operational stability, making adaptability and strong time management essential for success.

What is a Remote DevOps SRE Engineer job?

A Remote DevOps SRE (Site Reliability Engineer) job involves managing and automating infrastructure, ensuring system reliability, and optimizing deployment processes from a remote location. These professionals work with cloud platforms, CI/CD pipelines, monitoring tools, and configuration management to enhance system performance and availability. They collaborate with development and operations teams to prevent service disruptions and improve scalability. The role requires expertise in coding, automation, troubleshooting, and cloud technologies like AWS, Azure, or GCP.

What are popular job titles related to Remote Devops Sre Engineer jobs in New York? For Remote Devops Sre Engineer jobs in New York, the most frequently searched job titles are:
What job categories do people searching Remote Devops Sre Engineer jobs in New York look for? The top searched job categories for Remote Devops Sre Engineer jobs in New York are:
Infographic showing various Remote Devops Sre Engineer job openings in New York as of July 2026, with employment types broken down into 100% Full Time. Highlights an 100% Remote job distribution.

DevOps/SRE

Solidus Labs

Manhattan, NY โ€ข On-site, Remote

$62.50 - $83/hr

Full-time

Re-posted 14 days ago


Job description

Description
About Solidus Labs
At Solidus, we are shaping the financial markets of tomorrow by providing cutting-edge trade surveillance technology that protects investors, enhances transparency, and ensures regulatory compliance across traditional assets, prediction markets, and crypto.
With over 20 years of experience in developing Wall Street-grade FinTech, our team delivers innovative solutions that financial institutions and regulators worldwide rely on to detect, investigate, and report market manipulation, financial crime, and fraud. Headquartered in NYC, with offices in Singapore, Tel Aviv, and London, we safeguard millions of retail and institutional entities globally, monitoring over a trillion events each day.
The Role
We are seeking an experienced New York-based DevOps / Site Reliability Engineer to join our DevOps team and own the reliability, stability, and operational support of our production systems.
This role focuses on production ownership, monitoring, incident response, and on-call support, providing critical coverage. You will work with a modern cloud-native stack and play a key role in keeping systems highly available, secure, and performant.
Day-to-Day Responsibilities
  • Own the reliability, availability, and performance of our production environments.
  • Operate production Kubernetes (EKS), including cluster upgrades and Helm deployments.
  • Manage scaling and capacity using KEDA, Karpenter, and HPA for resource optimization.
  • Manage AWS Cloud environments including EC2, Lambda, AWS Batch, Elasticache, RDS, and more.
  • Evolve infrastructure as code using Terraform and Helm with security best practices.
  • Support GitLab CI/CD pipelines, resolving deployment issues and improving stability.
  • Design observability systems using Prometheus, Grafana, and EFK to reduce alert fatigue.
  • Solve networking issues involving TLS, Load Balancing, VPCs, NAT, and VPN.
  • Support compliance initiatives and respond to security-related incidents.
  • Leverage AI-powered tools as a standard part of your workflow for automation and productivity.
  • Lead incident response end-to-end, including troubleshooting, mitigation, and resolution.
  • Perform deep-dive RCA to drive long-term corrective and preventive actions.
  • Participate in on-call rotations to provide consistent operational coverage.

Requirements
Minimum Qualifications
  • 3+ years of hands-on DevOps / SRE experience
  • Strong production experience with Docker and Kubernetes
  • Solid knowledge of AWS (EKS, EC2, Organizations, RDS, S3, CloudWatch, Lambda, DynamoDB)
  • Experience with monitoring, logging, and alerting systems
  • Proficiency with Terraform, Helm, and GitLab CI (or similar)
  • Strong troubleshooting skills across infrastructure, CI/CD, and networking
  • Scripting experience with Bash and Python
  • Willingness to participate in on-call rotations
  • Familiarity with pub/sub systems (SQS, Kafka, or similar)

Nice to Have
  • Experience with Redis, Airflow, Databricks, Spark/EMR
  • GitOps workflows and advanced Git usage
  • Experience supporting databases such as Postgres, Snowflake, or ClickHouse

Why Join Us?
Join a team where you'll own and improve the reliability of critical production systems end to end, with real autonomy and impact, directly supporting premier clients globally. You'll work on a modern, cloud-native stack operating at scale, tackling meaningful performance and resilience challenges. And you'll do it alongside a highly collaborative, global DevOps and R&D team-sharing standards, tooling, and operational expertise across regions.