1

Site Reliability Engineer Jobs (NOW HIRING)

SITE RELIABILITY ENGINEER

Camden, NJ ยท On-site

$130K - $150K/yr

Site Reliability Engineer (SRE) Engineer Reliability into the Systems That Move the Nation's Food Supply Who We Are US Cold owns and operates one of the most complex temperature-controlled logistics ...

Site Reliability Engineer

Chicago, IL ยท On-site

$58.75 - $78/hr

W (flexible on other 2 days) Site Reliability Engineer - Northern Trust, Goals Driven Wealth Management We are searching for a candidate who has extensive experience in Site Reliability Engineering ...

Site Reliability Engineer

Alpharetta, GA

$55.75 - $74/hr

Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it-from-zero role at startup speed. You're our first ...

Site Reliability Engineer

Atlanta, GA

$54.75 - $72.75/hr

Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it-from-zero role at startup speed. You're our first ...

Site Reliability Engineer

Beaverton, OR ยท Hybrid

$59.25 - $78.75/hr

Overview As our Site Reliability Engineer, you'll help drive Concora Credit's Mission to enable customers to Do More with Credit - every single day. The impact you'll have at Concora Credit: We are ...

Site Reliability Engineer

Alpharetta, GA ยท On-site

$55.75 - $74/hr

Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it-from-zero role at startup speed. You're our first ...

Site Reliability Engineer

Beaverton, OR ยท Hybrid

$59.25 - $78.75/hr

Overview As our Site Reliability Engineer, you'll help drive Concora Credit's Mission to enable customers to Do More with Credit - every single day. The impact you'll have at Concora Credit: We are ...

Site Reliability Engineer (SRE)

Vienna, VA ยท On-site

$57.25 - $76/hr

The AWS Site Reliability Engineer (SRE) is responsible for the operational health, availability, and performance of the AWS and Databricks environments built by the Platform Engineering team. You ...

Site Reliability Engineer

Beaverton, OR ยท On-site

$59.25 - $78.75/hr

Overview As our Site Reliability Engineer, you'll help drive Concora Credit's Mission to enable customers to Do More with Credit - every single day. The impact you'll have at Concora Credit: We are ...

Site Reliability Engineer (SRE)

San Diego, CA ยท On-site

$60.50 - $80.50/hr

We are looking for the right Site Reliability Engineer to help us take our efforts to the next level. In this role, you will help lead our cloud based infrastructure team for Apple's Video Computer ...

Site Reliability Engineer

Jersey City, NJ ยท On-site

$59.50 - $79/hr

Site Reliability Engineer Location: U.S. (Hybrid) This role requires U.S. citizenship and eligibility for a U.S. security clearance. Role Summary: Exiger is transforming how governments and global ...

Site Reliability Engineer

Santa Clara, CA ยท On-site

$230K - $250K/yr

As our first or early SRE hire you will be building the reliability engineering function at Forward - defining how we think about availability, observability, incident response, and operational ...

Site Reliability Engineer

Frederick, MD ยท Hybrid

$56.75 - $75.25/hr

You'll combine DevOps and SRE practices to support mission-driven scientific and clinical programs, emphasizing automation, reliability, compliance, and proactive monitoring while enabling innovation ...

Site Reliability Engineer

Beaverton, OR ยท Hybrid

$59.25 - $78.75/hr

As our Site Reliability Engineer, you'll help drive Concora Credit's Mission to enable customers to Do More with Credit - every single day. The impact you'll have at Concora Credit: We are in search ...

Site Reliability Engineer

Palo Alto, CA ยท On-site

$67 - $89/hr

About the DevOps / SRE Team The DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable, scalable, and high-performing production systems. We work closely with ...

Showing results 41-60

Site Reliability Engineer information

See salary details

$10

$63

$91

How much do site reliability engineer jobs pay per hour?

As of Aug 6, 2026, the average hourly pay for site reliability engineer in the United States is $63.74, according to ZipRecruiter salary data. Most workers in this role earn between $54.81 and $72.84 per hour, depending on experience, location, and employer.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often involves on-call duties, troubleshooting complex issues, and working with automation tools, which can contribute to work-related stress but also offers opportunities for skill development and problem-solving.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.
What cities are hiring for Site Reliability Engineer jobs? Cities with the most Site Reliability Engineer job openings:
What are the most commonly searched types of Site Reliability Engineer jobs? The most popular types of Site Reliability Engineer jobs are:
Who are the top companies hiring for Site Reliability Engineer jobs? The top employers for Site Reliability Engineer jobs are:
What states have the most Site Reliability Engineer jobs? States with the most job openings for Site Reliability Engineer jobs include:
What job categories do people searching Site Reliability Engineer jobs look for? The top searched job categories for Site Reliability Engineer jobs are:
Infographic showing various Site Reliability Engineer job openings in the United States as of July 2026, with employment types broken down into 96% Full Time, 1% Part Time, and 3% Contract. Highlights an 87% Physical, 5% Hybrid, and 8% Remote job distribution, with an average salary of $132,583 per year, or $63.7 per hour.

SITE RELIABILITY ENGINEER

USCS

Camden, NJ โ€ข On-site

$130K - $150K/yr

Full-time

Re-posted 19 days ago


Job description

Site Reliability Engineer (SRE)
Engineer Reliability into the Systems That Move the Nation's Food Supply
Who We Are
US Cold owns and operates one of the most complex temperature-controlled logistics networks in North America. Every day, our systems coordinate the storage and movement of food at national scale across a network of state-of-the-art distribution centers, including multiple highly automated warehouse facilities.
We continue to advance our core warehouse and logistics platforms. Our current focus is on modular, event-driven, API-first and cloud architectures. We continue to enhance reliability and accelerate engineering productivity by strengthening our SRE and AI practices. This is a large investment in innovation to continue to drive operational excellence at our facilities.
If you want to build durable systems that operate in the physical world at scale, this is that opportunity.
The Role
The Site Reliability Engineer is a founding member of US Cold's SRE practice.
This role exists to move the organization from reactive operations to engineered reliability. You will study how our most critical systems fail - particularly our Phenix WMS and facility automation interfaces - and design controls, automation, and observability that reduce incidents over time.
Success in this role means fewer false alerts, faster recovery, less manual intervention, and systems that heal themselves when possible.
You will work closely with application, infrastructure, and operations teams and participate directly in on-call and incident response.
What You Will Own
  • Reliability of the Phenix WMS and its integration with facility automation systems (robotics, conveyors, and control interfaces)
  • Definition and implementation of SLIs and SLOs that measure meaningful system health, not just availability
  • Observability across the full stack, correlating cloud services, APIs, and on-premise facility operations
  • Automation to eliminate operational toil, including patching, data corrections, restarts, and recovery tasks
  • Development of self-healing behaviors for common failure modes
  • Participation in on-call rotations and leadership of blameless post-incident reviews
  • Design and execution of disaster recovery tests across SaaS, cloud, and on-premise environments

This is hands-on reliability engineering. The systems you improve will directly impact daily warehouse operations.
Technical Environment
  • Hybrid environments spanning cloud and on-premise infrastructure
  • Azure cloud services
  • Warehouse Management Systems (Phenix WMS) and facility automation interfaces
  • Java Development
  • Observability tooling across logs, metrics, and alerting
  • Automation using Python, PowerShell, Bash, or Ansible
  • CI/CD tools and modern deployment practices
  • Exposure to containerized and distributed systems environments

What We're Looking For
  • 3+ years of experience in SRE, DevOps, Systems Engineering, or related roles
  • Strong Linux and Windows systems administration and troubleshooting skills
  • Hands-on experience with automation and scripting
  • Experience designing and operating monitoring, alerting, and observability solutions
  • Practical experience working in Azure environments
  • Strong analytical skills and a bias toward eliminating root causes, not symptoms
  • Ability to collaborate across application, infrastructure, and operations teams
  • Experience supporting warehouse management systems or industrial automation platforms
  • Exposure to Kubernetes, microservices, or container orchestration
  • Hands on experience with infrastructure-as-code tools such as Terraform or Ansible
  • Understanding of distributed systems and high-availability design
  • Experience with SRE practices such as SLO-based operations, runbook automation, or chaos testing

Why This Role Is Different
This is not an inherited SRE function.
There is no mature framework to maintain.

You will:
  • Help define what reliability means at US Cold
  • Work on systems that operate in the physical world
  • Engineer solutions that reduce toil and operational load
  • See the direct impact of your work on warehouse uptime and performance
  • Build practices that scale as the platform modernizes

This is an opportunity to grow as an SRE while helping establish the reliability foundation of a mission-critical platform.
Compensation & Structure
  • Location: Hybrid - Camden NJ

  • Reports to: IT - Site Reliability Engineering Manager

  • Salary Range: $130,000- $150,000

Operational Context
  • Systems operate continuously across warehouse facilities
  • Reliability failures have physical and operational consequences
  • On-call participation is part of the role
  • Work occurs across cloud, SaaS, and on-premise environments

Equal Opportunity Employer
This employer is required to notify all applicants of their rights pursuant to federal employment laws.
For further information, please review the Know Your Rights notice from the Department of Labor.