Senior SRE/AIOps Engineer
Toronto, ON ยท On-site
Lead problem management and eliminate recurring issues * Develop and maintain runbooks, playbooks ... Must Have: * 3+ years of SRE or Systems Engineering experience with strong technical expertise.
Toronto, ON ยท On-site
Lead problem management and eliminate recurring issues * Develop and maintain runbooks, playbooks ... Must Have: * 3+ years of SRE or Systems Engineering experience with strong technical expertise.
Toronto, ON ยท On-site
Lead problem management and eliminate recurring issues * Develop and maintain runbooks, playbooks ... Must Have: * 3+ years of SRE or Systems Engineering experience with strong technical expertise.
Toronto, ON ยท Hybrid
CA$99K/yr
... Manage infrastructure deployment pipelines and troubleshoot onboarding and operational issues ... in Site Reliability, DevOps, or Cloud Engineering roles Expertise with Microsoft Azure; AWS ...
Toronto, ON ยท Hybrid
CA$99K/yr
... Manage infrastructure deployment pipelines and troubleshoot onboarding and operational issues ... in Site Reliability, DevOps, or Cloud Engineering roles Expertise with Microsoft Azure; AWS ...
The Lead is also expected to implement and maintain the infrastructure and tools that are necessary to manage the software development process. Reporting to the SRE, Developer Platform Engineering ...
The Lead is also expected to implement and maintain the infrastructure and tools that are necessary to manage the software development process. Reporting to the SRE, Developer Platform Engineering ...
Brampton, ON ยท On-site
CA$23.50 - CA$26.50/hr
... Management System (CMMS)with interpretation and trending of highly critical equipment with common repeating failure patterns at the plant * Supporting Reliability Engineer Data collection and ...
Brampton, ON ยท On-site
CA$23.50 - CA$26.50/hr
... Management System (CMMS)with interpretation and trending of highly critical equipment with common repeating failure patterns at the plant * Supporting Reliability Engineer Data collection and ...
... E practices and modern infrastructure. If you're ready to make a strategic impact in a ... Lead incident response and change management , coordinating cross-functional teams, conducting ...
... E practices and modern infrastructure. If you're ready to make a strategic impact in a ... Lead incident response and change management , coordinating cross-functional teams, conducting ...
Toronto, ON ยท Hybrid
CA$170K - CA$185K/yr
Interested in joining one of Canada's top-performing asset managers? We are seeking a Head of Platform Engineering, Reliability & Control to lead our horizontal engineering function and shared ...
Toronto, ON ยท Hybrid
CA$170K - CA$185K/yr
Interested in joining one of Canada's top-performing asset managers? We are seeking a Head of Platform Engineering, Reliability & Control to lead our horizontal engineering function and shared ...
Team Summary: The AI SRE team is a focused group of SRE engineers dedicated to making ... Strong programming skills for automation, operational tooling, and infrastructure management
Team Summary: The AI SRE team is a focused group of SRE engineers dedicated to making ... Strong programming skills for automation, operational tooling, and infrastructure management
Role Overview We're seeking a Senior Site Reliability Engineer to join our SaaS-Ops team within ... Manage and evolve Helm chart definitions and ArgoCD GitOps workflows for multi-region SaaS ...
Role Overview We're seeking a Senior Site Reliability Engineer to join our SaaS-Ops team within ... Manage and evolve Helm chart definitions and ArgoCD GitOps workflows for multi-region SaaS ...
CA$22 - CA$25/hr
... Management System (CMMS)with interpretation and trending of highly critical equipment with common repeating failure patterns at the plant * Supporting Reliability Engineer Data collection and ...
CA$22 - CA$25/hr
... Management System (CMMS)with interpretation and trending of highly critical equipment with common repeating failure patterns at the plant * Supporting Reliability Engineer Data collection and ...
The OPS Site Reliability Engineer will be a focal role owning and ensuring the fluent operations of Managed Services offerings in the KPMG production cloud environment. The role will be focusing on ...
The OPS Site Reliability Engineer will be a focal role owning and ensuring the fluent operations of Managed Services offerings in the KPMG production cloud environment. The role will be focusing on ...
The OPS Site Reliability Engineer will be a focal role owning and ensuring the fluent operations of Managed Services offerings in the KPMG production cloud environment. The role will be focusing on ...
The OPS Site Reliability Engineer will be a focal role owning and ensuring the fluent operations of Managed Services offerings in the KPMG production cloud environment. The role will be focusing on ...
Toronto, ON ยท Remote
CA$120/hr
Incident management / reliability / SRE Evaluator Type: Contract Compensation: $80-$120/hour Location: Remote Role Responsibilities * Evaluate AI-generated artifacts against domain-specific quality ...
Quick apply
Toronto, ON ยท Remote
CA$120/hr
Incident management / reliability / SRE Evaluator Type: Contract Compensation: $80-$120/hour Location: Remote Role Responsibilities * Evaluate AI-generated artifacts against domain-specific quality ...
Toronto, ON ยท On-site
We're searching for a Senior Site Reliability Engineer who not only brings deep technical expertise ... Create and refine tooling and processes that improve incident management, driving proactive ...
Toronto, ON ยท On-site
We're searching for a Senior Site Reliability Engineer who not only brings deep technical expertise ... Create and refine tooling and processes that improve incident management, driving proactive ...
Mississauga, ON ยท Hybrid
CA$70K - CA$80K/yr
The Site Reliability Engineer works to ensure services operate smoothly, quickly restore service ... Reporting to the Sr Manager, Applications Operations Engineering Key Responsibilities: Participate ...
Mississauga, ON ยท Hybrid
CA$70K - CA$80K/yr
The Site Reliability Engineer works to ensure services operate smoothly, quickly restore service ... Reporting to the Sr Manager, Applications Operations Engineering Key Responsibilities: Participate ...
Toronto, ON ยท On-site
... for the management and execution of Reliability projects and programs. * Drive change by ... Bachelor of Engineering Degree or equivalent related field. * Training in Reliability Engineering ...
Toronto, ON ยท On-site
... for the management and execution of Reliability projects and programs. * Drive change by ... Bachelor of Engineering Degree or equivalent related field. * Training in Reliability Engineering ...
Toronto, ON ยท Hybrid
CA$99K/yr
... Manage infrastructure deployment pipelines and troubleshoot onboarding and operational issues ... in Site Reliability, DevOps, or Cloud Engineering roles Expertise with Microsoft Azure; AWS ...
Toronto, ON ยท Hybrid
CA$99K/yr
... Manage infrastructure deployment pipelines and troubleshoot onboarding and operational issues ... in Site Reliability, DevOps, or Cloud Engineering roles Expertise with Microsoft Azure; AWS ...
Staff Software Reliability Engineer - Data Platform About the Team The Data Platform team is ... Participate in the on-call rotation, and incident management Required Knowledge, Skills, and ...
Staff Software Reliability Engineer - Data Platform About the Team The Data Platform team is ... Participate in the on-call rotation, and incident management Required Knowledge, Skills, and ...
CA$113K - CA$163K/yr
Senior Platform Reliability Engineer Join Manulife Global Wealth & Asset Management (GWAM) and help ... Manage and enhance Azure Kubernetes Services (AKS), Azure PaaS resources, and CI/CD pipelines.
CA$113K - CA$163K/yr
Senior Platform Reliability Engineer Join Manulife Global Wealth & Asset Management (GWAM) and help ... Manage and enhance Azure Kubernetes Services (AKS), Azure PaaS resources, and CI/CD pipelines.
At its heart, the Smile platform enables people and organizations to better manage healthcare data ... The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability, scalability ...
At its heart, the Smile platform enables people and organizations to better manage healthcare data ... The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability, scalability ...
Toronto, ON ยท On-site
CA$144K - CA$202K/yr
Proven experience managing SaaS or PaaS systems at enterprise scale (multi-region, multi-tenant ... reliability practices. #LI-KC1 About Kong: Kong Inc., a leading developer of API and AI ...
Quick apply
Toronto, ON ยท On-site
CA$144K - CA$202K/yr
Proven experience managing SaaS or PaaS systems at enterprise scale (multi-region, multi-tenant ... reliability practices. #LI-KC1 About Kong: Kong Inc., a leading developer of API and AI ...
| Aspect | Reliability Engineer | Reliability Engineer Manager |
|---|---|---|
| Required Credentials | Bachelor's in Engineering or related field; certifications like CRC, CRE | Same as Reliability Engineer, plus leadership experience |
| Work Environment | Design, analyze, and improve system reliability; often in teams | Oversees Reliability Engineers; manages projects and teams |
| Employer & Industry Usage | Manufacturing, aerospace, energy, automotive | Same industries, with added managerial responsibilities |
| Common Search & Comparison | Focuses on technical skills and hands-on reliability tasks | Focuses on leadership, team management, and strategic planning |
The main difference between a Reliability Engineer and a Reliability Engineer Manager lies in their responsibilities. The Reliability Engineer focuses on technical analysis and system improvements, while the Reliability Engineer Manager oversees teams, manages projects, and develops strategies to enhance reliability across the organization.

Job Description
WHAT IS THE OPPORTUNITY?
This role is responsible for designing, implementing, and maintaining SRE (Site Reliability Engineering) and AIOps (Artificial Intelligence for IT Operations) capabilities to ensure system reliability, proactive monitoring, and automation of self-healing operations. In addition to day-to-day support, the position provides end-to-end operational ownership across systems managed by multiple enterprise teams, including incident coordination, dependency management, and escalation. The role is also responsible for key security and compliance functions such as service ID and certificate management, SSO updates, vulnerability remediation, and lifecycle management of end-of-life components.
Our team supports a portfolio of multi-platform HR data pipelines that move and process data into Snowflake through multiple integrated components, requiring end-to-end monitoring, coordination, and support across systems.
In addition, we support SaaS-based applications that are primarily vendor-managed, while we retain responsibility for integration, access management, monitoring, and operational oversight.
WHAT WILL YOU DO?
Key Responsibilities:
SRE & Reliability Engineering
AIOps, Observability & Logging
Automation & Self-Healing Systems
Application & Data Platform Support
Infrastructure & Platform Expertise
Disaster Recovery & Resiliency
Governance, Compliance & Collaboration
WHAT DO YOU NEED TO SUCCEED?
Must Have:
Nice to Have:
WHAT'S IN IT FOR YOU?
We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.
A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable
Leaders who support your development through coaching and managing opportunities
Ability to make a difference and lasting impact
Work in a dynamic, collaborative, progressive, and high-performing team
A world-class training program in financial services
Opportunities to do challenging work
#LI-POST
#TECHPJ
Job Skills
Critical Thinking, Customer Support Systems, Group Problem Solving, Installation Support, IT Service Level Management, IT Service Management (ITSM), IT Standards, Technical TroubleshootingAdditional Job Details
Address:
City:
Country:
Work hours/week:
Employment Type:
Platform:
Job Type:
Pay Type:
Posted Date:
Application Deadline:
Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above
Our Employment Opportunities
At RBC, we are guided by living shared values of Client First, Integrity, Collaboration, Respect and Excellence and winning together as One RBC. We believe an inclusive workplace that has diverse perspectives is core to our continued growth as one of the largest and most successful banks in the world. Maintaining a workplace where our employees feel supported to perform at their best, effectively collaborate, drive innovation, and grow professionally helps to bring our Purpose to life and create value for our clients and communities. RBC strives to deliver this through policies and programs intended to foster a workplace based on respect, belonging and opportunity for all.
Join our Talent Community
Stay in-the-know about great career opportunities at RBC. Sign up and get customized info on our latest jobs, career tips and Recruitment events that matter to you.
Expand your limits and create a new future together at RBC. Find out how we use our passion and drive to enhance the well-being of our clients and communities at jobs.rbc.com.
RBC is presently inviting candidates to apply for this existing vacancy. Applying to this posting allows you to express your interest in this current career opportunity at RBC. Qualified applicants may be contacted to review their resume in more detail.
Employment Type: FULL_TIMESourced by ZipRecruiter
Banking and credit intermediation
10,000+ Employees
Toronto, Ontario, CA