2

Remote Reliability Engineer Jobs in Philadelphia, PA

OpenSearch SRE

Newtown Square, PA ยท Remote

$70 - $80/hr

Site Reliability Engineer - OpenSearch Location-Type: 100% Remote Start Date Is: September 21st Duration: 12 months (Contract-to-Hire Potential) Compensation Range: $70/hr - $80/hr W2 Benefits:

Site Reliability Engineer

Camden, NJ ยท Remote

$58.25 - $77.50/hr

A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who treats the exchange like the production trading system it is: sessions open and close on schedule ...

DevOps Engineer - Remote

Philadelphia, PA ยท Remote

$50 - $150/hr

... Engineer Role Type: Contractor Location: Remote Job Overview We are seeking experienced DevOps ... Maintain high standards for correctness, reliability, and reproducibility. * Provide structured ...

New

DevOps Engineer

Cherry Hill, NJ ยท On-site +1

$52.25 - $71.75/hr

... #SRE #DevOps #programming #kubernetes #security #compliance (ISO 27001, SOC). What are you looking ... Hybrid & remote work. So, what's up? If you liked what you read, don't hesitate.. hit on apply, we ...

Senior Software Engineer (Remote)

Philadelphia, PA ยท Remote

$123K - $163K/yr

This is a remote role anywhere in the USA. Meet the Team Our software engineering team develops ... Review and test code to identify problems and improve scalability, speed, and reliability.

Remote (must be able to work Eastern Standard Time hours) Start Date: ASAP Duration: Contract-to ... and reliability Build and support Node.js and Next.js layers in a Backend for Frontend (BFF ...

Senior DevOps Automation Engineer

Philadelphia, PA ยท On-site +1

$104K - $137K/yr

We are currently looking for a Senior DevOps Automation Engineer for a 100% remote position ... and reliability. * Take on additional tasks and responsibilities as needed to support team ...

Senior DevOps Automation Engineer

Philadelphia, PA ยท On-site +1

$104K - $137K/yr

We are currently looking for a Senior DevOps Automation Engineer for a 100% remote position ... and reliability. * Take on additional tasks and responsibilities as needed to support team ...

AI Engineer

Broomall, PA ยท On-site +1

... speed, reliability, and accuracy - while also closing more loans more quickly with greater ... Xactus is proud to provide a friendly work environment that is primarily remote. Our workforce ...

Data Engineer - Databricks, AWS, Python

Philadelphia, PA ยท Remote

$109K - $131K/yr

Data Engineer - Databricks, AWS, Python 100% Remote / MUST interview on-site: Phila., PA 19103 $125 ... Perform testing and quality assurance to validate data accuracy, pipeline reliability, and ...

next page

Showing results 1-20

Remote Reliability Engineer information

See Philadelphia, PA salary details

$61.6K

$119K

$142.3K

How much do remote reliability engineer jobs pay per year?

As of Aug 28, 2026, the average yearly pay for remote reliability engineer in Philadelphia, PA is $119,045.00, according to ZipRecruiter salary data. Most workers in this role earn between $103,400.00 and $130,200.00 per year, depending on experience, location, and employer.

What is a remote reliability engineer?

A Remote Reliability Engineer is a professional who works from a remote location to ensure that systems, applications, or infrastructure are reliable, available, and performing well. Their responsibilities typically include monitoring system health, diagnosing issues, implementing preventative measures, and collaborating with teams to improve system reliability. They often use tools for automation, incident response, and performance monitoring, all while working offsite. This role is critical in minimizing downtime and ensuring a smooth user experience, especially for companies with complex technical environments. Remote Reliability Engineers must have strong problem-solving skills and be proficient in cloud technologies, automation, and incident management.

What are the key skills and qualifications needed to thrive as a remote reliability engineer?

To thrive as a Remote Reliability Engineer, you need a strong background in systems engineering, software development, and infrastructure management, often supported by a degree in computer science or a related field. Proficiency with cloud platforms (such as AWS, Azure, or GCP), monitoring tools (like Prometheus, Grafana), and relevant certifications (e.g., AWS Certified DevOps Engineer) is highly valuable. Excellent problem-solving, communication, and collaboration skills are crucial for working effectively across distributed teams and responding to incidents. These abilities ensure system reliability, quick incident resolution, and seamless remote teamwork, which are vital for maintaining high service uptime and user satisfaction.

How do remote reliability engineers typically collaborate with on-site teams to address urgent technical issues?

Remote Reliability Engineers often utilize a combination of video conferencing, instant messaging, and collaborative monitoring tools to stay closely connected with on-site teams. When urgent technical issues arise, they participate in real-time troubleshooting sessions, analyze system logs remotely, and may guide on-site staff through step-by-step resolution procedures. Building strong communication channels and regular check-ins are essential to ensure swift and effective collaboration, even across different time zones. This structure allows Remote Reliability Engineers to contribute significantly to system uptime while working from a distance.

What is the difference between Remote Reliability Engineer vs Remote Site Reliability Engineer?

AspectRemote Reliability EngineerRemote Site Reliability Engineer
CredentialsTypically requires certifications like AWS Certified Solutions Architect, Linux Foundation certificationsSimilar credentials, often with additional focus on site-specific tools and monitoring
Work EnvironmentPrimarily remote, focusing on cloud infrastructure and system reliabilityRemote with some on-site responsibilities, focusing on infrastructure and operational stability
Industry UsageUsed across tech, cloud providers, SaaS companiesCommon in data centers, cloud providers, and large enterprise IT
Search & Comparison IntentOften compared due to overlapping roles in system reliability and cloud infrastructureCompared for on-site vs remote operational responsibilities

The main difference is that Remote Reliability Engineers focus on cloud and system reliability remotely, while Remote Site Reliability Engineers may have some on-site duties related to infrastructure. Both roles require similar skills and certifications but differ in their work environment and specific responsibilities.

What are the most commonly searched types of Reliability Engineer jobs in Philadelphia, PA?

The most popular types of Reliability Engineer jobs in Philadelphia, PA are:

What job categories do people searching Remote Reliability Engineer jobs in Philadelphia, PA look for?

The top searched job categories for Remote Reliability Engineer jobs in Philadelphia, PA are:

What cities near Philadelphia, PA are hiring for Remote Reliability Engineer jobs?

Cities near Philadelphia, PA with the most Remote Reliability Engineer job openings:

Infographic showing various Remote Reliability Engineer job openings in Philadelphia, PA as of August 2026, with employment types broken down into 91% Full Time, 5% Part Time, and 4% Contract. Highlights an 87% Physical, 5% Hybrid, and 8% Remote job distribution, with an average salary of $119,045 per year, or $57.2 per hour.

OpenSearch SRE

Newtown Square, PA โ€ข Remote

$70 - $80/hr

Contractor

Medical, Dental, Vision

Posted 7 days ago


Job description

Job Title: Site Reliability Engineer – OpenSearchLocation-Type: 100% RemoteStart Date Is: September 21st Duration: 12 months (Contract-to-Hire Potential)Compensation Range: $70/hr - $80/hr W2Benefits: Eligible for Health, Dental, VisionNot eligible for Visa sponsorship

Job Description:Build, deploy, maintain, and optimize high-performance OpenSearch/Elasticsearch environments on Kubernetes supporting mission-critical cloud services.

Day-to-Day Responsibilities:

  • Architect, build, deploy, and maintain OpenSearch/Elasticsearch clusters on Kubernetes
  • Monitor cluster health, node performance, indexing throughput, search latency, shard allocation, replication, and storage
  • Troubleshoot production issues across infrastructure, platform, and application layers
  • Lead incident response, root cause analysis, and post-incident remediation
  • Handle platform installations, upgrades, patching, backup/restore, and disaster recovery
  • Automate testing, deployment, scaling, recovery, and operational workflows
  • Build and maintain CI/CD pipelines
  • Support capacity planning across compute, memory, storage, and network resources
  • Manage log ingestion, index management, retention policies, and search performance
  • Develop monitoring, alerting, and operational runbooks
  • Participate in an on-call rotation and occasional weekend/after-hours releases

Minimum Requirements:

  • 8 years of SRE, DevOps, Platform Engineering, or related experience
  • Expert-level Kubernetes experience in complex production environments
  • Hands-on experience building, deploying, and maintaining  OpenSearch or Elasticsearch clusters on Kubernetes in production
  • Strong OpenSearch/Elasticsearch cluster administration, architecture, performance tuning, scaling, upgrades, and troubleshooting
  • Experience with index design, shard/replica strategy, cluster sizing, snapshot/restore, and disaster recovery
  • Strong Linux experience
  • Experience with Concourse CI/CD pipelines
  • Kafka and ZooKeeper experience
  • Strong Git and automation/scripting experience
  • Experience supporting distributed, highly available production systems
  • Hands-on production incident response, RCA, and on-call experience
  • Experience working within highly secure, complex enterprise environments

Preferred Qualifications:

  • AWS experience (EC2, S3, Route 53, CloudWatch, IAM, VPC, RDS)
  • Cloud Foundry / PCF experience
  • Terraform, Jenkins, and/or Chef
  • Prometheus and Grafana
  • Log ingestion and index lifecycle/retention management
  • Capacity forecasting, performance benchmarking, and resilience testing
  • SaaS / multi-tenant security experience