1

Site Reliability Engineer Jobs in Toronto, ON (NOW HIRING)

Site Reliability Engineer

Toronto, ON ยท Hybrid

CA$100K - CA$125K/yr

As a Site Reliability Engineer, you will play a crucial role in enhancing the reliability, performance, and scalability of our systems and services. You will be a part of a global "commando" team of ...

Site Reliability Engineer Hybrid GTA 12+ months This is an opportunity for someone early in their career who has hands-on experience with Azure data technologies and wants to build deeper experience ...

We're looking for a Site Reliability Engineer to help shape the future of intelligent operations by leveraging AIOps, GenAI, automation, and self-healing technologies across enterprise platforms. In ...

We're looking for a Site Reliability Engineer to help shape the future of intelligent operations by leveraging AIOps, GenAI, automation, and self-healing technologies across enterprise platforms. In ...

Site Reliability Engineer

Toronto, ON ยท On-site +1

CA$125K - CA$250K/yr

We are looking for a Site Reliability Engineer to help build and operate the infrastructure behind that work. Based in Toronto or remote, you will work across the systems that enable large-scale AI ...

SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams -- SRE teams are anchored ...

Mentor and guide engineers on cloud-native technologies, site reliability engineering principles, and operational excellence practices, fostering a culture of continuous learning and accountability.

... (SRE) to manage critical cloud infrastructure and site reliability operations for the Autodesk Platform Services and Emerging Technologies organization. The team delivers high-value, exabyte-scale ...

Senior Site Reliability Engineer

Toronto, ON ยท On-site

CA$90K - CA$132K/yr

Collaborate with software engineers and data engineers to embed SRE best practices into the development lifecycle, including SLOs, error budgets, and capacity planning. * Write scripts and tooling in ...

WHY THIS ROLE IS IMPORTANT TO US As a Senior Site Reliability Engineer, you will be embedded within one of our Product Areas, taking ownership of specific responsibility domains where your experience ...

next page

Showing results 1-20

Site Reliability Engineer information

See Toronto, ON salary details

$59.6K

$124.2K

$169.9K

How much do site reliability engineer jobs pay per year?

As of Sep 6, 2026, the average yearly pay for site reliability engineer in Toronto, ON is $124,163.00, according to ZipRecruiter salary data. Most workers in this role earn between $104,500.00 and $142,673.00 per year, depending on experience, location, and employer.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often requires strong problem-solving skills, familiarity with monitoring tools, and the ability to work in high-pressure situations, but it also offers opportunities for skill development and process improvements.

What are the most commonly searched types of Site Reliability Engineer jobs in Toronto, ON?

The most popular types of Site Reliability Engineer jobs in Toronto, ON are:

What are popular job titles related to Site Reliability Engineer jobs in Toronto, ON?

For Site Reliability Engineer jobs in Toronto, ON, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer jobs in Toronto, ON look for?

The top searched job categories for Site Reliability Engineer jobs in Toronto, ON are:

What cities near Toronto, ON are hiring for Site Reliability Engineer jobs?

Cities near Toronto, ON with the most Site Reliability Engineer job openings:

Infographic showing various Site Reliability Engineer job openings in Toronto, ON as of August 2026, with employment types broken down into 1% As Needed, 83% Full Time, 12% Part Time, 1% Temporary, 2% Contract, and 1% Nights. Highlights an 92% Physical, 4% Hybrid, and 4% Remote job distribution, with an average salary of $124,163 per year, or $59.7 per hour.

Site Reliability Engineer - SRE

Royal Bank of Canada

Toronto, ON โ€ข On-site

Full-time

Posted 9 days ago


Key responsibilities

  • Assist in development and implementation of SRE solutions such as monitoring, alerting, anomaly detection, self-healing, and reliability testing.

  • Perform production support activities including incident management, problem management, and ensuring application availability and uptime.

  • Support automation efforts, maintain system compliance, and evaluate opportunities for automation and improvement.


Job description

Job Description

What is job opportunity?
We are seeking a Site Resiliency Engineer to be a member of Production Support team and focus on automation, development, implementation and support of Site Reliability Engineering (SRE) solutions. The incumbent will need introductory knowledge of Azure Databricks, OpenShift, Splunk. Perform Production Support role and partner with SRE Delivery team in incident management and problem management.


What will you do?

Engineering:

  • Assist in development of SRE solutions (monitoring and alerting, machine learning anomaly detection, self-healing and reliability testing)

  • Implement monitoring and alerting, anomaly detection, self-healing and reliability testing for applications in scope

  • Supports enterprise goals to adopt automation solutions for applications in scope

  • Creating AI Agents
    Production Support:

  • Perform production support role on a shift-based work schedule, including off-hours support and rotational on-call supportto be compensated accordingly with overtime pay, lieu time, and on-call allowance

  • Assist in incident management and problem management for applications in scope

  • Evaluate continuously - what went well, what went wrong, what can be done to improve and prevent in future

  • Assist in maintaining technology currency (perform server patching, certificate renewal, etc.) with keen eye on automating opportunities

  • Assist in ensuring availability and uptime of applications in scope, as per service level objectives

  • Assist in ensuring compliance of all systems and applications in scope, including maintaining segregation of duties

Innovation and Learning:

  • Stay abreast of technology change and learn constantly, through official training assignments and self-assigned learning

What you need to succeed?

Must Have:

  • Azure Databricks

  • Be able to read application logs

  • Basic AI knowledge

  • PowerBI

  • Be able to work off business hours (after 5 PM, weekends)


Nice to Have:

  • Experience with IAM platforms and security-focused systems (Microsoft Entra, Okta, HashiCorp Vault, Service Now etc.)

  • Expertise with scripting (Python, PowerShell, Bash) to reduce manual toil and streamline deployments

  • Familiarity with identity, authentication, authorization, and privileged access management concepts and protocols

  • Knowledge of enterprise security architecture and compliance frameworks as they relate to IAM and service resilience

  • Exposure to AIOps or ML-based anomaly detection for proactive reliability management

What's in it for you?

We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.

  • A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable

  • Leaders who support your development through coaching and managing opportunities

  • Ability to make a difference and lasting impact

  • Work in a dynamic, collaborative, progressive, and high-performing team

  • A world-class training program in financial services

  • Opportunities to do challenging work

#TechPJ

#LI-ASPOST

Job Skills

Agile Methodology, Application Infrastructure, Group Problem Solving, IT Automation, IT Monitoring, Operations Support, Production Support, Software Development Life Cycle (SDLC), Software Engineering, Software Product Technical Knowledge, System Applications, Systems Software

Additional Job Details

Address:

RBC CENTRE, 155 WELLINGTON ST W:TORONTO

City:

Toronto

Country:

Canada

Work hours/week:

37.5

Employment Type:

Full time

Platform:

Job Type:

Regular

Pay Type:

Salaried

Posted Date:

2026-08-27

Application Deadline:

2026-09-17

Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above

Our Employment Opportunities

At RBC, we are guided by living shared values of Client First, Integrity, Collaboration, Respect and Excellence and winning together as One RBC. We believe an inclusive workplace that has diverse perspectives is core to our continued growth as one of the largest and most successful banks in the world. Maintaining a workplace where our employees feel supported to perform at their best, effectively collaborate, drive innovation, and grow professionally helps to bring our Purpose to life and create value for our clients and communities. RBC strives to deliver this through policies and programs intended to foster a workplace based on respect, belonging and opportunity for all.

Join our Talent Community
Stay in-the-know about great career opportunities at RBC. Sign up and get customized info on our latest jobs, career tips and Recruitment events that matter to you.
Expand your limits and create a new future together at RBC. Find out how we use our passion and drive to enhance the well-being of our clients and communities at jobs.rbc.com.

RBC is presently inviting candidates to apply for this existing vacancy. Applying to this posting allows you to express your interest in this current career opportunity at RBC. Qualified applicants may be contacted to review their resume in more detail.

Employment Type: FULL_TIME