1

Site Reliability Engineer Jobs in Toronto, ON (NOW HIRING)

Site Reliability Engineer Hybrid GTA 12+ months This is an opportunity for someone early in their career who has hands-on experience with Azure data technologies and wants to build deeper experience ...

Site Reliability Engineer

Toronto, ON · Hybrid

CA$100K - CA$125K/yr

As a Site Reliability Engineer, you will play a crucial role in enhancing the reliability, performance, and scalability of our systems and services. You will be a part of a global "commando" team of ...

About the Role We are looking for a Site Reliability Engineer to help design, build, and operate the platforms that power AI CoWorkers. This is a handson role for an engineer who enjoys owning ...

As a SRE, you will implement, measure and gather insights from Operational Level Indicators identifying areas for service improvements covering availability, performance, resilience, incidents and ...

SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams -- SRE teams are anchored ...

Site Reliability Engineer

Toronto, ON · On-site +1

CA$125K - CA$250K/yr

We are looking for a Site Reliability Engineer to help build and operate the infrastructure behind that work. Based in Toronto or remote, you will work across the systems that enable large-scale AI ...

Mentor and guide engineers on cloud-native technologies, site reliability engineering principles, and operational excellence practices, fostering a culture of continuous learning and accountability.

Senior Site Reliability Developer

Toronto, ON · On-site

CA$107K - CA$157K/yr

... (SRE) to manage critical cloud infrastructure and site reliability operations for the Autodesk Platform Services and Emerging Technologies organization. The team delivers high-value, exabyte-scale ...

WHY THIS ROLE IS IMPORTANT TO US As a Senior Site Reliability Engineer, you will be embedded within one of our Product Areas, taking ownership of specific responsibility domains where your experience ...

We are seeking a Site Reliability Engineer to ensure the availability, performance, and reliability of our containerized business applications. You will bridge the gap between development and ...

The OPS Site Reliability Engineer will be a focal role owning and ensuring the fluent operations of Managed Services offerings in the KPMG production cloud environment. The role will be focusing on ...

next page

Showing results 1-20

Site Reliability Engineer information

See Toronto, ON salary details

$59.6K

$124.2K

$169.9K

How much do site reliability engineer jobs pay per year?

As of Aug 30, 2026, the average yearly pay for site reliability engineer in Toronto, ON is $124,163.00, according to ZipRecruiter salary data. Most workers in this role earn between $104,500.00 and $142,673.00 per year, depending on experience, location, and employer.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often requires strong problem-solving skills, familiarity with monitoring tools, and the ability to work in high-pressure situations, but it also offers opportunities for skill development and process improvements.

What are the most commonly searched types of Site Reliability Engineer jobs in Toronto, ON?

The most popular types of Site Reliability Engineer jobs in Toronto, ON are:

What are popular job titles related to Site Reliability Engineer jobs in Toronto, ON?

For Site Reliability Engineer jobs in Toronto, ON, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer jobs in Toronto, ON look for?

The top searched job categories for Site Reliability Engineer jobs in Toronto, ON are:

What cities near Toronto, ON are hiring for Site Reliability Engineer jobs?

Cities near Toronto, ON with the most Site Reliability Engineer job openings:

Infographic showing various Site Reliability Engineer job openings in Toronto, ON as of August 2026, with employment types broken down into 1% As Needed, 85% Full Time, 12% Part Time, and 2% Contract. Highlights an 91% Physical, 4% Hybrid, and 5% Remote job distribution, with an average salary of $124,163 per year, or $59.7 per hour.

Site Reliability Engineer

Toronto, ON

Full-time

Re-posted 7 days ago


Job description

Job Description Site Reliability Engineer Hybrid GTA 12+ months This is an opportunity for someone early in their career who has hands-on experience with Azure data technologies and wants to build deeper experience across data platforms, production support, monitoring and Microsoft Fabric. What you will do: Monitor and support live data pipelines and data platform components Troubleshoot failed jobs, incidents and data-related issues Support ETL/ELT processes using Azure Data Factory, Synapse Analytics, Databricks and Microsoft Fabric Use SQL and Python or PySpark for data investigation, transformation and pipeline support Improve monitoring, alerting, logging and operational reliability Automate recurring support tasks and manual processes Work with data engineers, analysts, data scientists and business stakeholders Complete root cause analysis and help prevent recurring incidents What we are looking for: Approximately one to two years of relevant experience in data engineering, data platform support, cloud operations or a related area Hands-on exposure to the Microsoft Azure data environment, particularly Azure Synapse, Azure Data Lake and data pipelines Solid SQL and Python fundamentals Understanding of databases, data lakes, ETL/ELT and pipeline monitoring Experience supporting or troubleshooting data solutions in a production, co-op, internship or project environment Familiarity with Git, CI/CD, monitoring or incident-management practices Strong communication, problem-solving skills and willingness to learn This role is best suited to someone who is technically curious, comfortable investigating issues and interested in growing toward data engineering, cloud operations or site reliability engineering. Apply today!.