RBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience ...
RBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience ...
Site Reliability Engineer
Toronto, ON · On-site
About the Role We are looking for a Site Reliability Engineer to help design, build, and operate ... Our hiring process is managed in-house and the best way for candidates to express interest is by ...
Site Reliability Engineer
Toronto, ON · On-site
About the Role We are looking for a Site Reliability Engineer to help design, build, and operate ... Our hiring process is managed in-house and the best way for candidates to express interest is by ...
... management practices Strong communication, problem-solving skills and willingness to learn This ... site reliability engineering. Apply today!
... management practices Strong communication, problem-solving skills and willingness to learn This ... site reliability engineering. Apply today!
Manager, Site Reliability Engineering and DevOps
Toronto, ON · Hybrid
CA$142K - CA$186K/yr
Manager, Site Reliability Engineering (SRE) Your Moneris Career - The Opportunity The Manager, Site ... You will oversee SRE delivery across assigned domains while embedding reliability principles ...
Manager, Site Reliability Engineering and DevOps
Toronto, ON · Hybrid
CA$142K - CA$186K/yr
Manager, Site Reliability Engineering (SRE) Your Moneris Career - The Opportunity The Manager, Site ... You will oversee SRE delivery across assigned domains while embedding reliability principles ...
We are looking for a Site Reliability Engineer to help design and deploy, and operate the platforms ... Our hiring process is managed in-house, and the best way for candidates to express interest is by ...
We are looking for a Site Reliability Engineer to help design and deploy, and operate the platforms ... Our hiring process is managed in-house, and the best way for candidates to express interest is by ...
Manage cross-functional teams and stakeholders to execute upgrade and operational change management ... Decent knowledge of the following SRE practices and technologies: Python, YAML, Shell scripting ...
Manage cross-functional teams and stakeholders to execute upgrade and operational change management ... Decent knowledge of the following SRE practices and technologies: Python, YAML, Shell scripting ...
Site Reliability Engineer
Toronto, ON · Hybrid
Site Reliability Engineer The Site Reliability Engineer is responsible for ensuring the ... management. What You Bring: Technical Proficiency: * CI/CD tools * Cloud platforms (AWS, Azure)
Site Reliability Engineer
Toronto, ON · Hybrid
Site Reliability Engineer The Site Reliability Engineer is responsible for ensuring the ... management. What You Bring: Technical Proficiency: * CI/CD tools * Cloud platforms (AWS, Azure)
Site Reliability Engineer
London, ON · On-site
We are looking to hire a Site Reliability Engineer who will help in building and maintaining the ... Experience managing cloud infrastructure with Infrastructure as Code. Terraform preferred.
Site Reliability Engineer
London, ON · On-site
We are looking to hire a Site Reliability Engineer who will help in building and maintaining the ... Experience managing cloud infrastructure with Infrastructure as Code. Terraform preferred.
Sectigo's automated, cloud-native CLM platform issues and manages digital certificates across ... The Site Reliability Engineer designs and implements solutions to reduce toil and ensure ...
Quick apply
Sectigo's automated, cloud-native CLM platform issues and manages digital certificates across ... The Site Reliability Engineer designs and implements solutions to reduce toil and ensure ...
As a SRE, you will implement, measure and gather insights from Operational Level Indicators ... You excel at managing the communication of production releases, and their service availability ...
As a SRE, you will implement, measure and gather insights from Operational Level Indicators ... You excel at managing the communication of production releases, and their service availability ...
SRE is part of a global organization that leverages the latest technology to communicate with our ... You'll be a key voice in observability, change management, and service scalability, providing ...
Quick apply
SRE is part of a global organization that leverages the latest technology to communicate with our ... You'll be a key voice in observability, change management, and service scalability, providing ...
Senior Site Reliability Engineer
Ottawa, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Quick apply
Senior Site Reliability Engineer
Ottawa, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Senior Site Reliability Engineer
Kitchener, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Quick apply
Senior Site Reliability Engineer
Kitchener, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Senior Site Reliability Engineer
Ottawa, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Quick apply
Senior Site Reliability Engineer
Ottawa, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Senior Site Reliability Engineer
Kitchener, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Quick apply
Senior Site Reliability Engineer
Kitchener, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Site Reliability Engineer
Toronto, ON · On-site +1
CA$125K - CA$250K/yr
Improve provisioning, configuration management, testing, and deployment automation * Help plan ... site reliability engineering, infrastructure engineering, systems engineering, or a related ...
Site Reliability Engineer
Toronto, ON · On-site +1
CA$125K - CA$250K/yr
Improve provisioning, configuration management, testing, and deployment automation * Help plan ... site reliability engineering, infrastructure engineering, systems engineering, or a related ...
Site Reliability Engineer (.Net)
Toronto, ON · Hybrid
CA$100K - CA$125K/yr
As a Site Reliability Engineer, you will play a crucial role in enhancing the reliability ... supplier management, tax compliance, and treasury. Tipalti partners with leading financial ...
Site Reliability Engineer (.Net)
Toronto, ON · Hybrid
CA$100K - CA$125K/yr
As a Site Reliability Engineer, you will play a crucial role in enhancing the reliability ... supplier management, tax compliance, and treasury. Tipalti partners with leading financial ...
The Lead is also expected to implement and maintain the infrastructure and tools that are necessary to manage the software development process. Reporting to the SRE, Developer Platform Engineering ...
The Lead is also expected to implement and maintain the infrastructure and tools that are necessary to manage the software development process. Reporting to the SRE, Developer Platform Engineering ...
Senior Site Reliability Engineer
Toronto, ON · On-site
CA$90K - CA$132K/yr
We enable the Managed Investment Data (MID), Reference Entity Data (RED), Fixed Income, and Third ... Collaborate with software engineers and data engineers to embed SRE best practices into the ...
Senior Site Reliability Engineer
Toronto, ON · On-site
CA$90K - CA$132K/yr
We enable the Managed Investment Data (MID), Reference Entity Data (RED), Fixed Income, and Third ... Collaborate with software engineers and data engineers to embed SRE best practices into the ...
Senior Site Reliability Engineer
Toronto, ON · Hybrid
CA$99K/yr
... Manage infrastructure deployment pipelines and troubleshoot onboarding and operational issues ... in Site Reliability, DevOps, or Cloud Engineering roles Expertise with Microsoft Azure; AWS ...
Senior Site Reliability Engineer
Toronto, ON · Hybrid
CA$99K/yr
... Manage infrastructure deployment pipelines and troubleshoot onboarding and operational issues ... in Site Reliability, DevOps, or Cloud Engineering roles Expertise with Microsoft Azure; AWS ...
Site Reliability Engineer Manager information
See Ontario salary details
$130.5K - $136.5K
6% of jobs
$136.5K - $142.5K
6% of jobs
$142.5K - $148.5K
6% of jobs
$152.5K is the 25th percentile. Wages below this are outliers.
$148.5K - $154.5K
9% of jobs
The median wage is $160.5K / yr.
$154.5K - $160.5K
22% of jobs
$160.5K - $166.5K
18% of jobs
$170.7K is the 75th percentile. Wages above this are outliers.
$166.5K - $172.5K
10% of jobs
$172.5K - $178.5K
10% of jobs
$178.5K - $184.5K
6% of jobs
$184.5K - $190.5K
4% of jobs
$190.5K - $196.5K
1% of jobs
$130.5K
$163.3K
$196.5K
How much do site reliability engineer manager jobs pay per year?
What is a site reliability engineer manager?
What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?
| Aspect | Site Reliability Engineer (SRE) | Site Reliability Engineer Manager |
|---|---|---|
| Responsibilities | Focuses on designing, implementing, and maintaining reliable systems and automation | Oversees SRE teams, manages projects, and aligns reliability goals with business objectives |
| Required Skills | Strong coding, system design, and troubleshooting skills | Leadership, team management, strategic planning |
| Certifications | Google Cloud, AWS certifications, Linux, scripting | Same as SRE, plus management certifications (e.g., PMP) often preferred |
| Work Environment | Technical, hands-on with systems and automation | Managerial, coordinating teams and projects |
The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.
How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?
What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

Full-time
Posted 25 days ago
Job description
Job Description
What is the opportunity?
RBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and platforms that support the wealth management business. Working at the intersection of software engineering, cloud-native operations, observability, and automation, the team plays a central role in delivering reliable digital services for both internal users and clients.
As a Senior Site Reliability Engineer, you will bring an engineering-first mindset, strong operational judgment, and a passion for automation to improve system reliability at scale. You will work closely with development, infrastructure, platform, and support teams to build modern observability practices, improve incident response, strengthen reliability engineering standards, and drive the evolution toward intelligent, self-healing operations.
This role is ideal for a hands-on engineer who is equally comfortable improving production resilience, building automation, defining service-level objectives, and shaping the future of AI-enhanced operations. You will help design and implement scalable SRE solutions across the technology estate using tools and platforms such as Elasticsearch, Ansible, GitHub Actions, Dynatrace, PagerDuty, Moogsoft, Kubernetes, OpenShift, Kafka, and emerging AIOps capabilities.
What will you do?
- Build and enhance the SRE product base, including intelligent monitoring, alerting, reliability testing, anomaly detection, and automated remediation.
- Implement modern observability practices across supported applications, including metrics, logs, traces, dashboards, and actionable alerting.
- Design and pilot machine learning-based anomaly detection capabilities to improve signal quality and move from reactive to predictive operations.
- Architect and implement self-healing solutions that automatically remediate recurring operational issues with appropriate controls and governance.
- Design human-in-the-loop workflows that balance automation speed with accountability, risk management, and operational oversight.
- Standardize telemetry and instrumentation across platforms to improve visibility, coverage, and correlation of operational signals.
- Contribute to the centralization and evolution of observability and monitoring backends to enable deeper analytics and faster incident triage.
- Partner with cross-functional teams to improve monitoring, logging, alerting, incident response, and production readiness practices.
- Automate operational workflows and platform tasks using Ansible, GitHub Actions, and scripting languages such as Bash, Python, and PowerShell.
- Develop and maintain custom tooling that improves operational efficiency, reliability, and scale.
- Work closely with development teams to understand application changes, production risks, and release readiness, ensuring services meet reliability standards before and after deployment.
- Define, track, and improve SLIs, SLOs, error budgets, and other critical service health indicators.
- Evolve runbooks into automation-first remediation patterns and intelligent operational workflows.
- Support production deployments by advocating for reliability, resilience, and performance improvements.
- Lead or contribute to incident management, problem management, and root cause analysis, ensuring corrective actions are implemented and sustained.
- Troubleshoot production issues across application, middleware, infrastructure, and platform layers.
- Participate in an on-call rotation and provide senior operational support for business-critical systems.
- Continuously identify opportunities to simplify, automate, and modernize operations using engineering and AI-driven approaches.
What do you need to succeed?
Must-have
- 5+ years of experience in Site Reliability Engineering, Production Engineering, DevOps, Platform Engineering, or Systems Engineering roles with strong operational depth.
- Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
- Strong experience with infrastructure automation and configuration management, particularly Ansible.
- Strong scripting and automation skills in Bash, Python, PowerShell, or similar languages.
- Hands-on experience with modern reliability and observability tooling such as Elasticsearch, Dynatrace, GitHub, Kubernetes, OpenShift, Kafka, PagerDuty, Moogsoft, or related platforms.
- Strong understanding of production operations, incident management, root cause analysis, and reliability engineering practices.
- Experience defining and operating SLIs, SLOs, alerting strategies, and service health metrics.
- Knowledge of cloud-native and distributed systems concepts, including resiliency, scalability, fault isolation, and performance tuning.
- Understanding of AIOps, AI/ML concepts, or intelligent automation as applied to observability and operations.
- Ability to work across teams, influence engineering practices, and communicate clearly with technical and non-technical stakeholders.
Nice-to-have
- Experience in financial services, wealth management, banking, insurance, or other highly regulated environments.
- Experience with OpenTelemetry and telemetry standardization across distributed systems.
- Hands-on experience with Prometheus, Grafana, Splunk, Catchpoint, Azure Automation, or similar SRE and observability platforms.
- Experience with CI/CD and developer platform tools such as Jenkins, Artifactory, and Vault.
- Familiarity with containerization and cloud platform patterns, including Docker and Kubernetes-based deployments.
- Experience building or operating anomaly detection, predictive alerting, or self-healing automation solutions.
- Familiarity with AI governance, model validation, and operational controls in regulated environments.
- Experience with reliability testing, chaos engineering, or resilience validation practices.
What's in it for you?
We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.
A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable
Leaders who support your development through coaching and managing opportunities
Ability to make a difference and lasting impact
Work in a dynamic, collaborative, progressive, and high-performing team
A world-class training program in financial services
Opportunities to do challenging work
#LI-POST
#TECHPJ
Job Skills
Agile Methodology, Group Problem Solving, IT Systems Integration, Organizational Leadership, Product Services, Software Development Life Cycle (SDLC), System Applications, System Integration Testing (SIT), Systems SoftwareAdditional Job Details
Address:
City:
Country:
Work hours/week:
Employment Type:
Platform:
Job Type:
Pay Type:
Posted Date:
Application Deadline:
Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above
Our Employment Opportunities
At RBC, we are guided by living shared values of Client First, Integrity, Collaboration, Respect and Excellence and winning together as One RBC. We believe an inclusive workplace that has diverse perspectives is core to our continued growth as one of the largest and most successful banks in the world. Maintaining a workplace where our employees feel supported to perform at their best, effectively collaborate, drive innovation, and grow professionally helps to bring our Purpose to life and create value for our clients and communities. RBC strives to deliver this through policies and programs intended to foster a workplace based on respect, belonging and opportunity for all.
Join our Talent Community
Stay in-the-know about great career opportunities at RBC. Sign up and get customized info on our latest jobs, career tips and Recruitment events that matter to you.
Expand your limits and create a new future together at RBC. Find out how we use our passion and drive to enhance the well-being of our clients and communities at jobs.rbc.com.
RBC is presently inviting candidates to apply for this existing vacancy. Applying to this posting allows you to express your interest in this current career opportunity at RBC. Qualified applicants may be contacted to review their resume in more detail.
Employment Type: FULL_TIMEAbout Royal Bank of Canada
Sourced by ZipRecruiter
Industry
Banking and credit intermediation
Company size
10,000+ Employees
Headquarters location
Toronto, Ontario, CA