1

Senior Reliability Engineer Jobs in New York (NOW HIRING)

Email Reliability Engineer

New York, NY ยท On-site

$62.25 - $82.75/hr

Senior Systems Engineer - Email Reliability (Hybrid SRE) New York, NY (Hybrid, 3 days in office) Top-tier compensation package Join an elite technology group where systems engineering is a ...

Senior SRE Engineer

Parsippany, NJ ยท On-site

$57.25 - $76.25/hr

JOB SUMMARY We are seeking a highly technical, hands-on Senior Site Reliability Engineer (SRE) / DevOps Engineer to join our engineering organization. This individual will play a critical role in ...

What a Senior Site Reliability Engineer does at Clover As a Senior Site Reliability Engineer, you are responsible for managing, deploying, and architecting infrastructure solutions that scale, are ...

New

About The Opportunity As a Senior Engineer on the Runtime Automation team, you will design ... This is a high-impact position driving continuous reliability, deep system optimization, and ...

Showing results 21-40

Senior Reliability Engineer information

See New York salary details

$23

$70

$100

How much do senior reliability engineer jobs pay per hour?

As of Aug 8, 2026, the average hourly pay for senior reliability engineer in New York is $70.47, according to ZipRecruiter salary data. Most workers in this role earn between $58.12 and $84.42 per hour, depending on experience, location, and employer.

How much do senior reliability engineers get paid?

Senior reliability engineers typically earn between $90,000 and $130,000 annually, depending on experience, industry, and location. They often have expertise in systems analysis, failure modes, and reliability tools like FMEA and RCM, which can influence compensation levels.

What are the key skills and qualifications needed to thrive as a senior reliability engineer?

To thrive as a Senior Reliability Engineer, you need expertise in reliability engineering principles, root cause analysis, and a relevant engineering degree such as mechanical, electrical, or industrial engineering. Familiarity with tools like FMEA, RCA software, CMMS, and certifications such as Certified Reliability Engineer (CRE) are often required. Strong analytical thinking, communication skills, and the ability to lead cross-functional teams set top performers apart. These skills are essential for minimizing downtime, improving system reliability, and ensuring safe, efficient operations.

What are some common challenges faced by senior reliability engineers, and how are they typically addressed within the team?

Senior Reliability Engineers often encounter challenges such as diagnosing complex system failures, balancing proactive maintenance with urgent reactive fixes, and ensuring consistent communication across multidisciplinary teams. These challenges are typically addressed through root cause analysis, prioritization frameworks, and fostering a culture of knowledge sharing. Regular collaboration with operations, maintenance, and engineering teams helps in developing effective solutions and continuous improvement strategies.

What does a senior reliability engineer do?

A Senior Reliability Engineer is responsible for ensuring that systems, products, or processes operate reliably and efficiently over time. They analyze failure data, design reliability tests, develop maintenance strategies, and work with cross-functional teams to improve system performance and reduce downtime. Their expertise helps organizations minimize risk, optimize lifecycle costs, and maintain high standards of quality and safety. Senior Reliability Engineers often mentor junior team members and play a key role in developing reliability standards and best practices.

What is the difference between Senior Reliability Engineer vs Reliability Engineer?

AspectSenior Reliability EngineerReliability Engineer
CredentialsTypically requires 5+ years experience, certifications like CRE or Six SigmaEntry to mid-level, often with 2-4 years experience, similar certifications
Work EnvironmentDesigns and oversees reliability programs, leads projectsPerforms analysis, supports reliability improvements
Industry UsageUsed across manufacturing, energy, aerospaceCommon in same industries, often as a stepping stone to senior roles

The main difference between a Senior Reliability Engineer and a Reliability Engineer lies in experience, leadership responsibilities, and scope of work. Senior Reliability Engineers typically lead projects and develop strategies, while Reliability Engineers focus on analysis and supporting reliability initiatives. Both roles are vital in ensuring equipment and system dependability across industries.

What are the most commonly searched types of Reliability Engineer jobs in New York? The most popular types of Reliability Engineer jobs in New York are:
What cities in New York are hiring for Senior Reliability Engineer jobs? Cities in New York with the most Senior Reliability Engineer job openings:
Infographic showing various Senior Reliability Engineer job openings in New York as of July 2026, with employment types broken down into 93% Full Time, 4% Part Time, and 3% Contract. Highlights an 89% Physical, 4% Hybrid, and 7% Remote job distribution, with an average salary of $146,579 per year, or $70.5 per hour.

Senior Site Reliability Engineer

Castleton Commodities International, LLC

Stamford, CT โ€ข On-site

$60.75 - $80.75/hr

Full-time

Medical, Dental, Life, Retirement, PTO

Re-posted 5 days ago


Job description

The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering, Security, and Infrastructure teams to design resilient cloud-native architectures, implement Infrastructure as Code (IaC) and CI/CD standards, and drive measurable reliability outcomes. The Senior Site Reliability Engineer will also lead efforts to define and validate recovery objectives (RTO/RPO), design and implement Business Continuity / Disaster Recovery (BCP/DR) plans, and coordinate structured testing to ensure readiness.
Responsibilities:
Reliability Engineering & Operations
  • Own and improve service reliability through SLO/SLI definition, error budgets, and operational best practices.
  • Design, implement, and maintain observability (monitoring, logging, tracing, alerting) to reduce MTTR and improve proactive detection.
  • Lead incident response practices including on-call improvements, runbooks, post-incident reviews (RCA), and preventative actions.
  • Partner with application teams to improve performance, capacity planning, and resiliency under failure scenarios.

Infrastructure & Cloud Architecture
  • Design and operate highly available, fault-tolerant Cloud architectures (multi-AZ and, where required, multi-region).
  • Implement resilient patterns across compute, storage, networking, and managed services (e.g., autoscaling, load balancing, backups, replication).
  • Drive cloud governance best practices (tagging, account/landing zone patterns, least privilege, guardrails) in partnership with security and platform teams.

Infrastructure as Code (IaC) & DevOps Enablement
  • Build and maintain IaC modules and standards (e.g., Terraform, CloudFormation, CDK) for repeatable, auditable infrastructure delivery.
  • Develop, standardize, and optimize CI/CD pipelines to enable safe, automated deployments (e.g., GitHub Actions, GitLab CI, Jenkins, AWS CodePipeline).
  • Promote DevOps practices: version-controlled infrastructure, automated testing, immutable deployments, and progressive delivery patterns.
  • Establish environment consistency across dev/test/stage/prod and ensure infrastructure drift detection and remediation.

BCP/DR, RTO/RPO Definition & Testing
  • Collaborate with stakeholders to evaluate and define service-level RTO and RPO targets based on business and technical requirements.
  • Design and implement BCP/DR architectures and procedures (backups, restore workflows, replication, failover/failback, data integrity validation).
  • Coordinate and execute structured DR tests (tabletop, simulation, partial failover, full failover) and document outcomes.
  • Maintain DR runbooks, dependency maps, and recovery checklists; drive remediation of gaps identified during testing.
  • Produce metrics and reporting on DR readiness, test results, and continuous improvement actions.

Qualifications:
  • 7+ years of experience in SRE, DevOps, Platform Engineering, or Systems Engineering roles supporting production environments.
  • Strong proficiency with observability platforms (e.g., Datadog, Prometheus/Grafana, ELK/OpenSearch, Nagios, Nimsoft, etc).
  • Strong hands-on AWS experience building and operating production systems.
  • Proven expertise with Infrastructure as Code (Terraform and/or CloudFormation/CDK).
  • Strong CI/CD and automation background (pipeline design, deployment strategies, testing automation).
  • Experience defining and validating RTO/RPO, and implementing BCP/DR plans with structured testing.
  • Experience with Kubernetes and auto-scaling container platforms (EKS, ECS, or Kubernetes on-prem).
  • Strong Linux fundamentals, networking concepts (DNS, TCP/IP, load balancing), and troubleshooting skills.
  • Proficiency in at least one scripting/programming language (Python, Go, Bash, or similar).
  • Ability to write clear operational documentation, runbooks, and post-incident reports.
  • Ability to work effectively in a fast-paced, dynamic and high-intensity environment including open-floor plan if applicable to the position, with timely responsiveness and the ability to work beyond normal business hours when required.

Preferred Qualifications:
  • Familiarity with Azure and/or Oracle Cloud (OCI).
  • Familiarity with Service Mesh, API Gateways, and distributed tracing tooling.
  • Familiarity with OpenTelemetry, client instrumentations and collector configurations.
  • Security and compliance familiarity in cloud environments (IAM design, secrets management, audit logging).
  • Experience implementing progressive delivery (blue/green, canary), feature flags, and automated rollback.
  • Relevant certifications (AWS Solutions Architect/DevOps Engineer, Kubernetes CKA/CKAD).
  • Experience with ArgoCD & Karpenter.

Employee Programs & Benefits:
CCI offers competitive benefits and programs to support our employees, their families and local communities. These include:
  • Competitive comprehensive medical, dental, retirement and life insurance benefits
  • Employee assistance & wellness programs
  • Parental and family leave policies
  • CCI in the Community: Each office has a Charity Committee and as a part of this program employees are allocated 2 days annually to volunteer at the selected charities.
  • Charitable contribution match program
  • Tuition assistance & reimbursement
  • Quarterly Innovation & Collaboration Awards
  • Employee discount program, including access to fitness facilities
  • Competitive paid time off
  • Continued learning opportunities

Visit https://www.cci.com/careers/life-at-cci/# to learn more!
#LI-CD1