1

Site Reliability Engineer Jobs in Raleigh, NC (NOW HIRING)

Site Reliability Engineer Intern 2027

Durham, NC · Hybrid

$55 - $73.25/hr

Your role and responsibilities As a Site Reliability Engineer, you will work in an agile, collaborative environment to build, deploy, configure, and maintain systems for the IBM client business. In ...

Site Reliability Engineer Intern 2027

Durham, NC · On-site

$55 - $73.25/hr

Your role and responsibilities As a Site Reliability Engineer, you will work in an agile, collaborative environment to build, deploy, configure, and maintain systems for the IBM client business. In ...

Observability Engineer

Raleigh, NC · On-site

$70 - $85/hr

This position is ideal for someone who is passionate about monitoring, telemetry, distributed systems, and reliability engineering and enjoys partnering with development, SRE, and platform teams to ...

Site Reliability Engineer Spring Co-op 2027

Durham, NC · On-site

$55 - $73.25/hr

As a Site Reliability Engineer, you will work in an agile, collaborative environment to build, deploy, configure, and maintain systems for the IBM client business. In this role, you will lead the ...

As a Site Reliability Engineer, you will work in an agile, collaborative environment to build, deploy, configure, and maintain systems for the IBM client business. In this role, you will lead the ...

Lead DevOps Engineer

Raleigh, NC · On-site

$51.25 - $70.25/hr

... SRE practices across the organization. Key Responsibilities Technical Leadership & Ownership Act as a hands‑on technical resource, setting the bar for DevOps engineering excellence Mentor and guide ...

Lead DevOps Engineer

Raleigh, NC · On-site

$51.25 - $70.25/hr

... SRE practices across the organization. Key Responsibilities Technical Leadership & Ownership Act as a hands‑on technical resource, setting the bar for DevOps engineering excellence Mentor and guide ...

Showing results 41-60

Site Reliability Engineer information

See Raleigh, NC salary details

$10

$61

$89

How much do site reliability engineer jobs pay per hour?

As of Sep 14, 2026, the average hourly pay for site reliability engineer in Raleigh, NC is $61.96, according to ZipRecruiter salary data. Most workers in this role earn between $53.27 and $70.82 per hour, depending on experience, location, and employer.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often requires strong problem-solving skills, familiarity with monitoring tools, and the ability to work in high-pressure situations, but it also offers opportunities for skill development and process improvements.

What are the most commonly searched types of Site Reliability Engineer jobs in Raleigh, NC?

The most popular types of Site Reliability Engineer jobs in Raleigh, NC are:

What are popular job titles related to Site Reliability Engineer jobs in Raleigh, NC?

For Site Reliability Engineer jobs in Raleigh, NC, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer jobs in Raleigh, NC look for?

The top searched job categories for Site Reliability Engineer jobs in Raleigh, NC are:

What cities near Raleigh, NC are hiring for Site Reliability Engineer jobs?

Cities near Raleigh, NC with the most Site Reliability Engineer job openings:

Infographic showing various Site Reliability Engineer job openings in Raleigh, NC as of September 2026, with employment types broken down into 1% Internship, 2% As Needed, 77% Full Time, 14% Part Time, and 6% Contract. Highlights an 88% Physical, 2% Hybrid, and 10% Remote job distribution, with an average salary of $128,882 per year, or $62 per hour.

CaaS Private Site Reliability Engineer - Assistant Vice President

Cary, NC • On-site

Deutsche Bank
Banking and Credit Intermediation • 10K+ employees

$53.25 - $70.75/hr

Full-time

Medical, Retirement, PTO

Re-posted 5 days ago


Deutsche Bank rating

7.7

Company rating: 7.7 out of 10

Based on 14 frontline employees who took The Breakroom Quiz


Job description

Job Description:
Job Title CaaS Private Site Reliability Engineer
Corporate Title Assistant Vice President
Location Cary, NC
Who we are:
In short - an essential part of Deutsche Bank's technology solution, developing applications for key business areas.
Our Technologists drive Cloud, Cyber and business technology strategy while transforming it within a robust, hands-on engineering culture. Learning is a key element of our people strategy, and we have a variety of options for you to develop professionally. Our approach to the future of work champions flexibility and is rooted in the understanding that there have been dramatic shifts in the ways we work.
Having first established a presence in the Americas in the 19th century, Deutsche Bank opened its US technology center in Cary, North Carolina in 2009. Learn more about us here
Overview
As a Site Reliability Engineer on the CaaS Private platform team, you will help operate and improve an on-prem, multi-tenant Kubernetes platform running on bare metal. You will strengthen the reliability, observability, scalability, and operational excellence of a platform that supports critical, low-latency, and regulated workloads. You will partner closely with platform, network, security, and application teams to define service level objectives, improve resilience, reduce operational toil, and build automation that allows the platform to run safely at scale. Join us here, and you will turn operational challenges into measurable engineering improvements that application teams can rely on every day.
What We Offer You
  • A diverse and inclusive environment that embraces change, innovation, and collaboration
  • A hybrid working model, allowing for in-office / work from home flexibility, generous vacation, personal and volunteer days
  • Employee Resource Groups support an inclusive workplace for everyone and promote community engagement
  • Competitive compensation packages including health and wellbeing benefits, retirement savings plans, parental leave, and family building benefits
  • Educational resources, matching gift and volunteer programs

What You'll Do
  • Define, implement, and continuously improve SLI/SLOs, alerting standards, and error budgets for the CaaS Private platform and critical services
  • Build and maintain observability across metrics, logs, alerts, and dashboards to provide clear insight into platform health, saturation, latency, and failure modes
  • Lead or coordinate incident response for platform-impacting events, ensuring timely mitigation, clear communication, blameless postmortems, and durable follow-up actions
  • Automate repetitive operational tasks and remediation workflows to reduce toil, improve platform consistency, and accelerate recovery time
  • Improve reliability, upgrade safety, and operational readiness for Kubernetes clusters, ingress paths, service mesh components, node services, and critical platform dependencies
  • Partner with platform, network, security, and application teams on capacity planning, release readiness, troubleshooting, operational documentation, and adoption of best practices

Skills You'll Need
  • Hands-on Kubernetes expertise, including operating clusters on bare metal or private cloud environments and supporting platform services at scale
  • Proven experience in Site Reliability Engineering, Production Engineering, DevOps, or a closely related infrastructure role
  • Strong Linux system administration capability and infrastructure-level scripting experience using Python, Ansible, and Bash
  • Practical knowledge of observability stacks and telemetry pipelines, including Prometheus, Grafana, Splunk, metrics, logging, alerting, dashboards, and Open Telemetry-style concepts
  • Strong understanding of incident management, root cause analysis, operational readiness, computer networking, virtualization, containerization, and distributed systems behavior under failure
  • Proven ability to leverage AI tools to enhance productivity, optimize workflows to solve business problems, while applying critical judgment to ensure responsible and ethical use of data and AI outputs

Skills That Will Help You Excel
  • Experience with Istio / Envoy, service mesh observability, traffic management, OPA Gatekeeper, admission controls, or policy-driven operational guardrails
  • Familiarity supporting stateful services such as PostgreSQL, Kafka, MongoDB, or comparable platform dependencies
  • Practical knowledge of capacity planning, load testing, chaos testing, failure-injection techniques, alert tuning, and self-healing automation
  • Exposure to low-latency or regulated environments with strict uptime, change control, compliance, time synchronization, deterministic performance, or SR-IOV workload constraints
  • Ability to read and understand Golang code when troubleshooting platform components, with strong written and verbal communication skills and a continuous learning mindset

Expectations
It is the Bank's expectation that employees hired into this role will work in the Cary, NC office in accordance with the Bank's hybrid working model.
Deutsche Bank provides reasonable accommodations to candidates and employees with a substantiated need based on disability and/or religion.
The salary range for this position in Cary is $100,000 to $153,000. Actual salaries may be based on a number of factors including, but not limited to, a candidate's skill set, experience, education, work location and other qualifications. Posted salary ranges do not include incentive compensation or any other type of remuneration.
Deutsche Bank Benefits
At Deutsche Bank, we recognize that our benefit programs have a profound impact on our colleagues. That's why we are focused on providing benefits and perks that enable our colleagues to live authentically and be their whole selves, at every stage of life. We provide access to physical, emotional, and financial wellness benefits that allow our colleagues to stay financially secure and strike balance between work and home. Click here to learn more!
Learn more about your life at Deutsche Bank through the eyes of our current employees: https://careers.db.com/life
The California Consumer Privacy Act outlines how companies can use personal information. If you are interested in receiving a copy of Deutsche Bank's California Privacy Notice please email HR.Direct@DB.com.
#LI-HYBRID
We strive for a culture in which we are empowered to excel together every day. This includes acting responsibly, thinking commercially, taking initiative and working collaboratively.
Together we share and celebrate the successes of our people. Together we are Deutsche Bank Group.
We welcome applications from all people and promote a positive, fair and inclusive work environment.
Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, protected veteran status or other characteristics protected by law. Click these links to view Deutsche Bank's Equal Opportunity Policy Statement and the following notices: EEOC Know Your Rights; Employee Rights and Responsibilities under the Family and Medical Leave Act; and Employee Polygraph Protection Act.

What Deutsche Bank employees say

Pay

Hours and flexibility

Workplace

Get the full story on Breakroom


Deutsche Bank logo

About Deutsche Bank

Sourced by ZipRecruiter

Deutsche Bank is the leading German bank with strong European roots and a global network. We're driving growth through our strong client franchise. Against a backdrop of increasing globalization in the world economy, Deutsche Bank is very well-positioned, with significant regional diversification and substantial revenue streams from all the major regions of the world. We serve our clients' real economic needs in commercial banking, investment banking, private banking and asset management. We are investing heavily in digital technologies, prioritizing long term success over short-term gains, and serving society with ambition and integrity.

Industry

Banking and credit intermediation

Company size

10,000+ Employees

Headquarters location

New York, NY, US

Social media