1

Site Reliability Engineer Jobs in Philadelphia, PA

Site Reliability Engineer

Camden, NJ · On-site

$130K - $150K/yr

Site Reliability Engineer (SRE) Engineer Reliability into the Systems That Move the Nation's Food Supply Who We Are US Cold owns and operates one of the most complex temperature-controlled logistics ...

Site Reliability Engineer

Pennington, NJ · On-site

$57.50 - $76.50/hr

Site Reliability Engineer We are seeking a hands-on Site Reliability Engineer (SRE) to support the reliability, observability, automation, and operational health of a large-scale production platform.

New

DeVops SRE

Wilmington, DE · On-site

$55.25 - $73.50/hr

Responsibilities : • Implement and maintain SRE DevOps practices to enhance system reliability and performance. • Utilize Splunk for monitoring and observability to ensure system health and ...

Staff Site Reliability Engineer

Crum Lynne, PA

$54.50 - $72.25/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work on the platform. As a Staff Site ...

Staff Site Reliability Engineer

Crum Lynne, PA

$54.50 - $72.25/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work on the platform. As a Staff Site ...

Site Reliability Engineer

Camden, NJ · Remote

$58.25 - $77.50/hr

A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who treats the exchange like the production trading system it is: sessions open and close on schedule ...

Expert SRE UI Engineer

Malvern, PA · On-site

$56 - $74.25/hr

The role involves leading SRE initiatives, architecting resiliency solutions, and integrating AI capabilities into user interfaces. Responsibilities : • Join Personal Investor Technologies Site ...

Site Reliability Engineer

Camden, NJ · On-site

$150K - $170K/yr

A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who treats the exchange like the production trading system it is: sessions open and close on schedule ...

Site Reliability Engineer

Camden, NJ · On-site

$57.50 - $76.50/hr

A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who treats the exchange like the production trading system it is: sessions open and close on schedule ...

next page

Showing results 1-20

Site Reliability Engineer information

See Philadelphia, PA salary details

$10

$64

$92

How much do site reliability engineer jobs pay per hour?

As of Aug 27, 2026, the average hourly pay for site reliability engineer in Philadelphia, PA is $64.32, according to ZipRecruiter salary data. Most workers in this role earn between $55.29 and $73.51 per hour, depending on experience, location, and employer.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often requires strong problem-solving skills, familiarity with monitoring tools, and the ability to work in high-pressure situations, but it also offers opportunities for skill development and process improvements.

What are the most commonly searched types of Site Reliability Engineer jobs in Philadelphia, PA?

The most popular types of Site Reliability Engineer jobs in Philadelphia, PA are:

What are popular job titles related to Site Reliability Engineer jobs in Philadelphia, PA?

For Site Reliability Engineer jobs in Philadelphia, PA, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer jobs in Philadelphia, PA look for?

The top searched job categories for Site Reliability Engineer jobs in Philadelphia, PA are:

What cities near Philadelphia, PA are hiring for Site Reliability Engineer jobs?

Cities near Philadelphia, PA with the most Site Reliability Engineer job openings:

Infographic showing various Site Reliability Engineer job openings in Philadelphia, PA as of August 2026, with employment types broken down into 75% Full Time, and 25% Contract. Highlights an 100% In-person job distribution, with an average salary of $133,788 per year, or $64.3 per hour.

Site Reliability Engineer

Camden, NJ • On-site


United States Cold Storage
Warehousing and Storage • 1 - 5K employees

7.8

Company rating: 7.8 out of 10

Based on 49 frontline employees who took The Breakroom Quiz

90th of 366 rated logistics

Good employer

Paid breaks

Respectful managers


$130K - $150K/hr

Full-time

Re-posted 11 days ago


Job description

Site Reliability Engineer (SRE)
Engineer Reliability into the Systems That Move the Nation’s Food SupplyWho We AreUS Cold owns and operates one of the most complex temperature-controlled logistics networks in North America. Every day, our systems coordinate the storage and movement of food at national scale across a network of state-of-the-art distribution centers, including multiple highly automated warehouse facilities.We continue to advance our core warehouse and logistics platforms. Our current focus is on modular, event-driven, API-first and cloud architectures. We continue to enhance reliability and accelerate engineering productivity by strengthening our SRE and AI practices. This is a large investment in innovation to continue to drive operational excellence at our facilities.If you want to build durable systems that operate in the physical world at scale, this is that opportunity. The RoleThe Site Reliability Engineer is a founding member of US Cold’s SRE practice.This role exists to move the organization from reactive operations to engineered reliability. You will study how our most critical systems fail — particularly our Phenix WMS and facility automation interfaces — and design controls, automation, and observability that reduce incidents over time.Success in this role means fewer false alerts, faster recovery, less manual intervention, and systems that heal themselves when possible.You will work closely with application, infrastructure, and operations teams and participate directly in oncall and incident response.What You Will Own
  • Reliability of the Phenix WMS and its integration with facility automation systems (robotics, conveyors, and control interfaces)
  • Definition and implementation of SLIs and SLOs that measure meaningful system health, not just availability
  • Observability across the full stack, correlating cloud services, APIs, and onpremise facility operations
  • Automation to eliminate operational toil, including patching, data corrections, restarts, and recovery tasks
  • Development of selfhealing behaviors for common failure modes
  • Participation in oncall rotations and leadership of blameless postincident reviews
  • Design and execution of disaster recovery tests across SaaS, cloud, and onpremise environments
This is handson reliability engineering. The systems you improve will directly impact daily warehouse operations.Technical Environment
  • Hybrid environments spanning cloud and onpremise infrastructure
  • Azure cloud services
  • Warehouse Management Systems (Phenix WMS) and facility automation interfaces
  • Java Development
  • Observability tooling across logs, metrics, and alerting
  • Automation using Python, PowerShell, Bash, or Ansible
  • CI/CD tools and modern deployment practices
  • Exposure to containerized and distributed systems environments
What We’re Looking For
  • 3+ years of experience in SRE, DevOps, Systems Engineering, or related roles
  • Strong Linux and Windows systems administration and troubleshooting skills
  • Handson experience with automation and scripting
  • Experience designing and operating monitoring, alerting, and observability solutions
  • Practical experience working in Azure environments
  • Strong analytical skills and a bias toward eliminating root causes, not symptoms
  • Ability to collaborate across application, infrastructure, and operations teams
  • Experience supporting warehouse management systems or industrial automation platforms
  • Exposure to Kubernetes, microservices, or container orchestration
  • Hands on experience with infrastructureascode tools such as Terraform or Ansible
  • Understanding of distributed systems and highavailability design
  • Experience with SRE practices such as SLObased operations, runbook automation, or chaos testing
Why This Role Is DifferentThis is not an inherited SRE function.
 There is no mature framework to maintain.You will:
  • Help define what reliability means at US Cold
  • Work on systems that operate in the physical world
  • Engineer solutions that reduce toil and operational load
  • See the direct impact of your work on warehouse uptime and performance
  • Build practices that scale as the platform modernizes
This is an opportunity to grow as an SRE while helping establish the reliability foundation of a missioncritical platform.Compensation & Structure 
  • Location: Hybrid – Camden NJ 
  • Reports to: IT – Site Reliability Engineering Manager
  • Salary Range: $130,000- $150,000
Operational Context
  • Systems operate continuously across warehouse facilities
  • Reliability failures have physical and operational consequences
  • Oncall participation is part of the role
  • Work occurs across cloud, SaaS, and onpremise environments


What United States Cold Storage employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom