1

Site Reliability Engineer Manager Jobs in Philadelphia, PA

Site Reliability Engineer

Camden, NJ · On-site

$130K - $150K/yr

Site Reliability Engineer (SRE) Engineer Reliability into the Systems That Move the Nation's Food ... Warehouse Management Systems (Phenix WMS) and facility automation interfaces * Java Development

Site Reliability Engineer

Pennington, NJ · On-site

$57.50 - $76.50/hr

Site Reliability Engineer We are seeking a hands-on Site Reliability Engineer (SRE) to support the reliability, observability, automation, and operational health of a large-scale production platform.

New

DeVops SRE

Wilmington, DE · On-site

$55.25 - $73.50/hr

Responsibilities : • Implement and maintain SRE DevOps practices to enhance system reliability and performance. • Utilize Splunk for monitoring and observability to ensure system health and ...

Site Reliability Engineer

Camden, NJ · Remote

$58.25 - $77.50/hr

A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who ... Hands-on experience operating Kubernetes in production and managing infrastructure as code ...

Site Reliability Engineer

Camden, NJ · On-site

$150K - $170K/yr

A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who ... Hands-on experience operating Kubernetes in production and managing infrastructure as code ...

Staff Site Reliability Engineer

Crum Lynne, PA

$54.50 - $72.25/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work on the platform. As a Staff Site ...

Site Reliability Engineer

Camden, NJ · On-site

$57.50 - $76.50/hr

A successful Site Reliability Engineer at Sporttrade will be an operations-minded self-starter who ... Hands-on experience operating Kubernetes in production and managing infrastructure as code ...

Staff Site Reliability Engineer

Crum Lynne, PA

$54.50 - $72.25/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work on the platform. As a Staff Site ...

Expert SRE UI Engineer

Malvern, PA · On-site

$56 - $74.25/hr

The role involves leading SRE initiatives, architecting resiliency solutions, and integrating AI capabilities into user interfaces. Responsibilities : • Join Personal Investor Technologies Site ...

next page

Showing results 1-20

Site Reliability Engineer Manager information

See Philadelphia, PA salary details

$10

$64

$92

How much do site reliability engineer manager jobs pay per hour?

As of Aug 27, 2026, the average hourly pay for site reliability engineer manager in Philadelphia, PA is $64.32, according to ZipRecruiter salary data. Most workers in this role earn between $55.29 and $73.51 per hour, depending on experience, location, and employer.

What is a site reliability engineer manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.

How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How much do site reliability engineer managers get paid?

Site Reliability Engineer Managers typically earn between $120,000 and $180,000 annually, depending on experience, location, and company size. They often oversee teams responsible for system reliability, incident response, and infrastructure automation, requiring strong leadership and technical skills.

Is a Site Reliability Engineer Manager a stressful job?

A Site Reliability Engineer Manager role can be stressful due to the responsibility of maintaining system uptime, managing incident responses, and ensuring reliability across complex infrastructure. The job often involves working under pressure, handling outages, and coordinating teams, but it also offers opportunities for problem-solving and leadership. Stress levels vary depending on company size, team structure, and workload management skills.

What are the most commonly searched types of Site Reliability Engineer jobs in Philadelphia, PA?

The most popular types of Site Reliability Engineer jobs in Philadelphia, PA are:

What cities near Philadelphia, PA are hiring for Site Reliability Engineer Manager jobs?

Cities near Philadelphia, PA with the most Site Reliability Engineer Manager job openings:

Infographic showing various Site Reliability Engineer Manager job openings in Philadelphia, PA as of August 2026, with employment types broken down into 1% As Needed, 80% Full Time, 16% Part Time, 1% Temporary, and 2% Contract. Highlights an 93% Physical, 2% Hybrid, and 5% Remote job distribution, with an average salary of $133,788 per year, or $64.3 per hour.

Site Reliability Engineer

Camden, NJ • On-site


United States Cold Storage
Warehousing and Storage • 1 - 5K employees

7.8

Company rating: 7.8 out of 10

Based on 49 frontline employees who took The Breakroom Quiz

90th of 366 rated logistics

Good employer

Paid breaks

Respectful managers


$130K - $150K/hr

Full-time

Re-posted 11 days ago


Job description

Site Reliability Engineer (SRE)
Engineer Reliability into the Systems That Move the Nation’s Food SupplyWho We AreUS Cold owns and operates one of the most complex temperature-controlled logistics networks in North America. Every day, our systems coordinate the storage and movement of food at national scale across a network of state-of-the-art distribution centers, including multiple highly automated warehouse facilities.We continue to advance our core warehouse and logistics platforms. Our current focus is on modular, event-driven, API-first and cloud architectures. We continue to enhance reliability and accelerate engineering productivity by strengthening our SRE and AI practices. This is a large investment in innovation to continue to drive operational excellence at our facilities.If you want to build durable systems that operate in the physical world at scale, this is that opportunity. The RoleThe Site Reliability Engineer is a founding member of US Cold’s SRE practice.This role exists to move the organization from reactive operations to engineered reliability. You will study how our most critical systems fail — particularly our Phenix WMS and facility automation interfaces — and design controls, automation, and observability that reduce incidents over time.Success in this role means fewer false alerts, faster recovery, less manual intervention, and systems that heal themselves when possible.You will work closely with application, infrastructure, and operations teams and participate directly in oncall and incident response.What You Will Own
  • Reliability of the Phenix WMS and its integration with facility automation systems (robotics, conveyors, and control interfaces)
  • Definition and implementation of SLIs and SLOs that measure meaningful system health, not just availability
  • Observability across the full stack, correlating cloud services, APIs, and onpremise facility operations
  • Automation to eliminate operational toil, including patching, data corrections, restarts, and recovery tasks
  • Development of selfhealing behaviors for common failure modes
  • Participation in oncall rotations and leadership of blameless postincident reviews
  • Design and execution of disaster recovery tests across SaaS, cloud, and onpremise environments
This is handson reliability engineering. The systems you improve will directly impact daily warehouse operations.Technical Environment
  • Hybrid environments spanning cloud and onpremise infrastructure
  • Azure cloud services
  • Warehouse Management Systems (Phenix WMS) and facility automation interfaces
  • Java Development
  • Observability tooling across logs, metrics, and alerting
  • Automation using Python, PowerShell, Bash, or Ansible
  • CI/CD tools and modern deployment practices
  • Exposure to containerized and distributed systems environments
What We’re Looking For
  • 3+ years of experience in SRE, DevOps, Systems Engineering, or related roles
  • Strong Linux and Windows systems administration and troubleshooting skills
  • Handson experience with automation and scripting
  • Experience designing and operating monitoring, alerting, and observability solutions
  • Practical experience working in Azure environments
  • Strong analytical skills and a bias toward eliminating root causes, not symptoms
  • Ability to collaborate across application, infrastructure, and operations teams
  • Experience supporting warehouse management systems or industrial automation platforms
  • Exposure to Kubernetes, microservices, or container orchestration
  • Hands on experience with infrastructureascode tools such as Terraform or Ansible
  • Understanding of distributed systems and highavailability design
  • Experience with SRE practices such as SLObased operations, runbook automation, or chaos testing
Why This Role Is DifferentThis is not an inherited SRE function.
 There is no mature framework to maintain.You will:
  • Help define what reliability means at US Cold
  • Work on systems that operate in the physical world
  • Engineer solutions that reduce toil and operational load
  • See the direct impact of your work on warehouse uptime and performance
  • Build practices that scale as the platform modernizes
This is an opportunity to grow as an SRE while helping establish the reliability foundation of a missioncritical platform.Compensation & Structure 
  • Location: Hybrid – Camden NJ 
  • Reports to: IT – Site Reliability Engineering Manager
  • Salary Range: $130,000- $150,000
Operational Context
  • Systems operate continuously across warehouse facilities
  • Reliability failures have physical and operational consequences
  • Oncall participation is part of the role
  • Work occurs across cloud, SaaS, and onpremise environments


What United States Cold Storage employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom