Experience building and managing CI/CD pipelines using GitHub Actions or similar platforms ... Observability & Reliability * Experience designing and operating monitoring, logging, tracing, and ...
Experience building and managing CI/CD pipelines using GitHub Actions or similar platforms ... Observability & Reliability * Experience designing and operating monitoring, logging, tracing, and ...
Site Reliability Engineer
Toronto, ON · On-site
About the Role We are looking for a Site Reliability Engineer to help design, build, and operate ... Our hiring process is managed in-house and the best way for candidates to express interest is by ...
Site Reliability Engineer
Toronto, ON · On-site
About the Role We are looking for a Site Reliability Engineer to help design, build, and operate ... Our hiring process is managed in-house and the best way for candidates to express interest is by ...
SRE Evaluator - Incident Management
Toronto, ON · Remote
CA$80 - CA$120/hr
Incident management / reliability / SRE Evaluator Type: Contract Compensation: $80-$120/hour Location: Remote Role Responsibilities * Evaluate AI-generated artifacts against domain-specific quality ...
Quick apply
SRE Evaluator - Incident Management
Toronto, ON · Remote
CA$80 - CA$120/hr
Incident management / reliability / SRE Evaluator Type: Contract Compensation: $80-$120/hour Location: Remote Role Responsibilities * Evaluate AI-generated artifacts against domain-specific quality ...
As a SRE, you will implement, measure and gather insights from Operational Level Indicators ... You excel at managing the communication of production releases, and their service availability ...
As a SRE, you will implement, measure and gather insights from Operational Level Indicators ... You excel at managing the communication of production releases, and their service availability ...
Site Reliability Engineer
Toronto, ON · Hybrid
Provide regular reliability reporting, SLO performance metrics, and incident trends to senior management. What You Bring: Technical Proficiency: * CI/CD tools * Cloud platforms (AWS, Azure)
Site Reliability Engineer
Toronto, ON · Hybrid
Provide regular reliability reporting, SLO performance metrics, and incident trends to senior management. What You Bring: Technical Proficiency: * CI/CD tools * Cloud platforms (AWS, Azure)
Senior Site Reliability Developer
Toronto, ON · On-site
CA$107K - CA$157K/yr
Reporting to the Engineering Manager, you will be leading design and development of resilient and ... Streamline CI/CD processes, improve system reliability, and ensure infrastructure scalability and ...
Senior Site Reliability Developer
Toronto, ON · On-site
CA$107K - CA$157K/yr
Reporting to the Engineering Manager, you will be leading design and development of resilient and ... Streamline CI/CD processes, improve system reliability, and ensure infrastructure scalability and ...
... management practices Strong communication, problem-solving skills and willingness to learn This ... reliability engineering. Apply today!
... management practices Strong communication, problem-solving skills and willingness to learn This ... reliability engineering. Apply today!
Manage cross-functional teams and stakeholders to execute upgrade and operational change management * Oversee end-to-end reliability of the ecosystem (hardware, software, network) ensuring 99.9% ...
Manage cross-functional teams and stakeholders to execute upgrade and operational change management * Oversee end-to-end reliability of the ecosystem (hardware, software, network) ensuring 99.9% ...
We are seeking a Senior Manager, Site Reliability Engineering to ensure the availability, performance, and reliability of our containerized business applications. You will bridge the gap between ...
We are seeking a Senior Manager, Site Reliability Engineering to ensure the availability, performance, and reliability of our containerized business applications. You will bridge the gap between ...
The Electric Reliability Engineer is responsible for supporting the overall reliability of ... Reporting to the Engineering Manager, you will also provide expertise on electrical systems to ...
The Electric Reliability Engineer is responsible for supporting the overall reliability of ... Reporting to the Engineering Manager, you will also provide expertise on electrical systems to ...
Site Reliability Engineer
Toronto, ON · On-site +1
CA$125K - CA$250K/yr
We are looking for a Site Reliability Engineer to help build and operate the infrastructure behind ... Improve provisioning, configuration management, testing, and deployment automation * Help plan ...
Site Reliability Engineer
Toronto, ON · On-site +1
CA$125K - CA$250K/yr
We are looking for a Site Reliability Engineer to help build and operate the infrastructure behind ... Improve provisioning, configuration management, testing, and deployment automation * Help plan ...
Lead Site Reliability Administrator
Waterloo, ON · On-site
CA$177K/yr
Hiring Manager: Vaughn Anderson Talent Acquisition Advisor: Smijeet Kurup Job Code Level: IZ-CLD-P4 ... As a Site Reliability Engineer (SRE) at OpenText, you will be responsible for ensuring the ...
Lead Site Reliability Administrator
Waterloo, ON · On-site
CA$177K/yr
Hiring Manager: Vaughn Anderson Talent Acquisition Advisor: Smijeet Kurup Job Code Level: IZ-CLD-P4 ... As a Site Reliability Engineer (SRE) at OpenText, you will be responsible for ensuring the ...
Site Reliability Engineer
Toronto, ON · Hybrid
CA$100K - CA$125K/yr
... supplier management, tax compliance, and treasury. Tipalti partners with leading financial ... Lead reliability-focused practices such as SLO (Service Level Objective) design and implementation ...
Site Reliability Engineer
Toronto, ON · Hybrid
CA$100K - CA$125K/yr
... supplier management, tax compliance, and treasury. Tipalti partners with leading financial ... Lead reliability-focused practices such as SLO (Service Level Objective) design and implementation ...
Site Reliability Engineer
Toronto, ON · Hybrid
CA$100K - CA$125K/yr
... supplier management, tax compliance, and treasury. Tipalti partners with leading financial ... Lead reliability-focused practices such as SLO (Service Level Objective) design and implementation ...
Site Reliability Engineer
Toronto, ON · Hybrid
CA$100K - CA$125K/yr
... supplier management, tax compliance, and treasury. Tipalti partners with leading financial ... Lead reliability-focused practices such as SLO (Service Level Objective) design and implementation ...
Lead Site Reliability Administrator
Mississauga, ON · On-site
CA$177K/yr
Hiring Manager: Vaughn Anderson Talent Acquisition Advisor: Smijeet Kurup Job Code Level: IZ-CLD-P4 ... As a Site Reliability Engineer (SRE) at OpenText, you will be responsible for ensuring the ...
Lead Site Reliability Administrator
Mississauga, ON · On-site
CA$177K/yr
Hiring Manager: Vaughn Anderson Talent Acquisition Advisor: Smijeet Kurup Job Code Level: IZ-CLD-P4 ... As a Site Reliability Engineer (SRE) at OpenText, you will be responsible for ensuring the ...
SRE is part of a global organization that leverages the latest technology to communicate with our ... You'll be a key voice in observability, change management, and service scalability, providing ...
Quick apply
SRE is part of a global organization that leverages the latest technology to communicate with our ... You'll be a key voice in observability, change management, and service scalability, providing ...
Corporate Reliability Engineering Co-Op
CA$22 - CA$25/hr
... Management System (CMMS)with interpretation and trending of highly critical equipment with common ... Supporting Reliability Engineer Data collection and standardization of Failure modes and the ...
Corporate Reliability Engineering Co-Op
CA$22 - CA$25/hr
... Management System (CMMS)with interpretation and trending of highly critical equipment with common ... Supporting Reliability Engineer Data collection and standardization of Failure modes and the ...
Minimize risk of reliability failures related to durability, availability, performance, and ... Experience with incident management processes, on-call rotations, and post-incident review ...
Minimize risk of reliability failures related to durability, availability, performance, and ... Experience with incident management processes, on-call rotations, and post-incident review ...
Support deployments, release activities, change management, and CI/CD pipelines, including GitHub ... Investigate performance, reliability, and availability issues using Datadog, Azure Monitor, Log ...
Support deployments, release activities, change management, and CI/CD pipelines, including GitHub ... Investigate performance, reliability, and availability issues using Datadog, Azure Monitor, Log ...
The team owns reliability and operational excellence for our highly available SaaS platform, a ... Manage and evolve Helm chart definitions and ArgoCD GitOps workflows for multi-region SaaS ...
The team owns reliability and operational excellence for our highly available SaaS platform, a ... Manage and evolve Helm chart definitions and ArgoCD GitOps workflows for multi-region SaaS ...
Reliability Manager information
See Guelph, ON salary details
$74.2K - $79.8K
6% of jobs
$79.8K - $85.5K
7% of jobs
$85.5K - $91.2K
7% of jobs
$93.4K is the 25th percentile. Wages below this are outliers.
$91.2K - $96.8K
10% of jobs
$96.8K - $102.5K
7% of jobs
$102.5K - $108.2K
10% of jobs
The median wage is $108.8K / yr.
$108.2K - $113.9K
20% of jobs
$116K is the 75th percentile. Wages above this are outliers.
$113.9K - $119.5K
18% of jobs
$119.5K - $125.2K
7% of jobs
$125.2K - $130.9K
3% of jobs
$130.9K - $136.5K
3% of jobs
$74.2K
$107K
$136.5K
How much do reliability manager jobs pay per year?
What does a reliability manager do?
A Reliability Manager is responsible for ensuring that equipment, processes, and systems operate efficiently and consistently to minimize downtime and maximize performance. They develop and implement reliability strategies, conduct root cause analyses, and oversee preventive and predictive maintenance programs. Their role involves working closely with maintenance teams, engineers, and production staff to improve asset reliability and extend equipment lifespan. Additionally, they analyze failure data, recommend improvements, and help optimize operational costs through reliability-centered maintenance practices.
What are the key skills and qualifications needed to thrive as a reliability manager?
A Reliability Manager needs strong analytical skills, a solid background in engineering or maintenance, and experience with reliability-centered maintenance methodologies. Familiarity with tools like Failure Mode and Effects Analysis (FMEA), Root Cause Analysis (RCA), and certifications such as Certified Reliability Engineer (CRE) are often required. Leadership, problem-solving, and the ability to communicate complex technical information clearly are crucial soft skills for this role. These skills help ensure equipment uptime, optimize maintenance processes, and foster a culture of continuous improvement within the organization.
What job categories do people searching Reliability Manager jobs in Guelph, ON look for?
The top searched job categories for Reliability Manager jobs in Guelph, ON are:
What cities near Guelph, ON are hiring for Reliability Manager jobs?
Cities near Guelph, ON with the most Reliability Manager job openings:

Full-time
Medical, Retirement
Posted 13 days ago
Job description
This is a hands-on senior engineering role focused on improving production resilience, strengthening security, driving operational excellence, and enhancing the developer experience across the organization.
In this role, you will design, build, and evolve the foundational systems, tooling, and operational practices that enable engineering teams to ship secure, reliable, and scalable software with confidence. You will help establish reliability standards, define service level objectives (SLOs), improve observability, automate operational processes, and drive incident management and post-incident learning practices that strengthen platform stability over time.
Partnering closely with Engineering, Security, Platform, and Product teams, you will architect scalable distributed systems, optimize Kubernetes and AWS-based infrastructure, and build automated delivery pipelines that support rapid and safe software releases. You will play a key role in reducing operational toil, improving system performance, increasing platform reliability, and ensuring that our infrastructure can support continued business growth.
This is a full-time permanent position
This is an existing vacancy
Location: This is a remote location open to candidates legally authorized to work in Canada.
- Drive reliability engineering initiatives and operational excellence for mission-critical services running on AWS and Kubernetes.
- Design, implement, and continuously improve deployment, release, and rollback strategies across complex distributed systems.
- Establish secure-by-default CI/CD pipelines with robust automation, governance, and policy-driven controls.
- Enhance platform observability through metrics, logs, tracing, and actionable alerting to improve system visibility and operational efficiency.
- Define, implement, and mature Service Level Indicators (SLIs), Service Level Objectives (SLOs), and reliability standards across the organization.
- Lead response efforts for high-severity incidents, ensuring timely resolution, effective communication, and meaningful post-incident reviews that drive continuous improvement.
- Partner closely with engineering teams to strengthen platform standards, improve service resilience, optimize runtime performance, and embed reliability best practices.
- Mentor and guide engineers on cloud-native technologies, site reliability engineering principles, and operational excellence practices, fostering a culture of continuous learning and accountability.
- 8+ years of experience in Site Reliability Engineering (SRE), Platform Engineering, DevOps, or related cloud-native engineering roles.
- Deep expertise in AWS services, including EKS, IAM, VPC, Lambda, CloudFront, S3, and cloud networking/security best practices.
- Advanced experience operating and scaling production Kubernetes environments.
- Strong hands-on experience with Istio service mesh, including traffic management, security, observability, and resiliency.
- Proven expertise with Infrastructure as Code (IaC), preferably using AWS CDK.
- Experience building and managing CI/CD pipelines using GitHub Actions or similar platforms.
- Strong troubleshooting, performance optimization, and incident management experience in distributed systems.
- Excellent communication, collaboration, and technical leadership skills.
- Experience designing and operating monitoring, logging, tracing, and alerting solutions for cloud-native platforms.
- Strong knowledge of AWS CloudWatch, OpenTelemetry, AWS X-Ray, and Kubernetes observability tooling.
- Experience defining and operationalizing SLIs, SLOs, alerting strategies, runbooks, and reliability metrics.
- Proven ability to leverage observability data to improve service reliability, reduce incident impact, and optimize operational performance.
- Strong proficiency in TypeScript and Node.js for platform engineering, automation, and operational tooling.
- Experience building and maintaining scalable backend services, APIs, and event-driven systems.
- Deep understanding of Kubernetes architecture, controllers, Gateway API, ingress management, and service networking.
- Experience implementing zero-trust architectures, mTLS, and service-to-service security controls.
- Commitment to high-quality engineering practices, including automated testing, code reviews, and observability-driven development.
- Strong understanding of resilience engineering, including autoscaling, disruption management, failure testing, and safe deployment strategies.
- Experience with progressive delivery practices such as canary, blue/green, and feature-flag-based deployments.
- Experience working in regulated, compliance-driven, or security-sensitive SaaS environments.
- Familiarity with FinOps principles and cost optimization strategies for cloud platforms.
- Experience building internal developer platforms and self-service engineering tooling.
- Cloud-native certifications such as CKA, CKAD, CKS, KCSA, or KCNA.
- Kubestronaut certification or equivalent advanced Kubernetes expertise is highly regarded.
Salary Range:
The annual base salary for this position is between $140,000 CAD and $155,000 CAD per year.
This role is also eligible for discretionary bonus and/or commission, as well as other benefits. Actual pay within the listed range will be determined based on factors such as transferable skills, relevant experience, market conditions, and primary work location. The posted range is subject to change and may be updated periodically.