We are looking to hire a Site Reliability Engineer who will help in building and maintaining the ... Experience managing cloud infrastructure with Infrastructure as Code. Terraform preferred.
We are looking to hire a Site Reliability Engineer who will help in building and maintaining the ... Experience managing cloud infrastructure with Infrastructure as Code. Terraform preferred.
Experience with observability, monitoring, and incident management tools such as Honeycomb.io, New ... Reliability practices are clearer, more consistent, and more measurable across Engineering teams.
New
Experience with observability, monitoring, and incident management tools such as Honeycomb.io, New ... Reliability practices are clearer, more consistent, and more measurable across Engineering teams.
New
Site Reliability Engineer
Toronto, ON · Hybrid
Provide regular reliability reporting, SLO performance metrics, and incident trends to senior management. What You Bring: Technical Proficiency: * CI/CD tools * Cloud platforms (AWS, Azure)
Site Reliability Engineer
Toronto, ON · Hybrid
Provide regular reliability reporting, SLO performance metrics, and incident trends to senior management. What You Bring: Technical Proficiency: * CI/CD tools * Cloud platforms (AWS, Azure)
Experience with observability, monitoring, and incident management tools such as Honeycomb.io, New ... Reliability practices are clearer, more consistent, and more measurable across Engineering teams.
New
Experience with observability, monitoring, and incident management tools such as Honeycomb.io, New ... Reliability practices are clearer, more consistent, and more measurable across Engineering teams.
New
... and access management, cloud computing services/integrations, and data analytics technologies. Responsibilities of the Cloud Service Reliability Engineer: * Establishes technology product ...
Quick apply
... and access management, cloud computing services/integrations, and data analytics technologies. Responsibilities of the Cloud Service Reliability Engineer: * Establishes technology product ...
As a SRE, you will implement, measure and gather insights from Operational Level Indicators ... You excel at managing the communication of production releases, and their service availability ...
As a SRE, you will implement, measure and gather insights from Operational Level Indicators ... You excel at managing the communication of production releases, and their service availability ...
Site Reliability Engineer
Ottawa, ON · On-site
Sectigo's automated, cloud-native CLM platform issues and manages digital certificates across ... The Site Reliability Engineer designs and implements solutions to reduce toil and ensure ...
Quick apply
Site Reliability Engineer
Ottawa, ON · On-site
Sectigo's automated, cloud-native CLM platform issues and manages digital certificates across ... The Site Reliability Engineer designs and implements solutions to reduce toil and ensure ...
Direct and manage all aspects of production support and reliability engineering for large, complex, mission-critical applications. Anticipate, detect, and address production-related issues ...
Direct and manage all aspects of production support and reliability engineering for large, complex, mission-critical applications. Anticipate, detect, and address production-related issues ...
Technical Expertise- Manage and optimize automated application deployments across multiple environments, ensuring consistency and reliability. Monitor system health, performance metrics, and alerts ...
Technical Expertise- Manage and optimize automated application deployments across multiple environments, ensuring consistency and reliability. Monitor system health, performance metrics, and alerts ...
... management practices Strong communication, problem-solving skills and willingness to learn This ... reliability engineering. Apply today!
... management practices Strong communication, problem-solving skills and willingness to learn This ... reliability engineering. Apply today!
Technical Expertise- Manage and optimize automated application deployments across multiple environments, ensuring consistency and reliability. Monitor system health, performance metrics, and alerts ...
Technical Expertise- Manage and optimize automated application deployments across multiple environments, ensuring consistency and reliability. Monitor system health, performance metrics, and alerts ...
Manage cross-functional teams and stakeholders to execute upgrade and operational change management * Oversee end-to-end reliability of the ecosystem (hardware, software, network) ensuring 99.9% ...
Manage cross-functional teams and stakeholders to execute upgrade and operational change management * Oversee end-to-end reliability of the ecosystem (hardware, software, network) ensuring 99.9% ...
We are seeking a Senior Manager, Site Reliability Engineering to ensure the availability, performance, and reliability of our containerized business applications. You will bridge the gap between ...
We are seeking a Senior Manager, Site Reliability Engineering to ensure the availability, performance, and reliability of our containerized business applications. You will bridge the gap between ...
The Electric Reliability Engineer is responsible for supporting the overall reliability of ... Reporting to the Engineering Manager, you will also provide expertise on electrical systems to ...
The Electric Reliability Engineer is responsible for supporting the overall reliability of ... Reporting to the Engineering Manager, you will also provide expertise on electrical systems to ...
Site Reliability Engineer
Toronto, ON · On-site +1
CA$125K - CA$250K/yr
We are looking for a Site Reliability Engineer to help build and operate the infrastructure behind ... Improve provisioning, configuration management, testing, and deployment automation * Help plan ...
Site Reliability Engineer
Toronto, ON · On-site +1
CA$125K - CA$250K/yr
We are looking for a Site Reliability Engineer to help build and operate the infrastructure behind ... Improve provisioning, configuration management, testing, and deployment automation * Help plan ...
Sr. Reliability Engineer - Embedded Products
Markham, ON · On-site
CA$162K - CA$244K/yr
Perform reliability modeling and statistical analysis of energy management and protection/control products to predict field performance and support design decisions. * Develop and execute reliability ...
Sr. Reliability Engineer - Embedded Products
Markham, ON · On-site
CA$162K - CA$244K/yr
Perform reliability modeling and statistical analysis of energy management and protection/control products to predict field performance and support design decisions. * Develop and execute reliability ...
Sr. Reliability Engineer - Embedded Products
Markham, ON · On-site
CA$162K - CA$244K/yr
Perform reliability modeling and statistical analysis of energy management and protection/control products to predict field performance and support design decisions. * Develop and execute reliability ...
Sr. Reliability Engineer - Embedded Products
Markham, ON · On-site
CA$162K - CA$244K/yr
Perform reliability modeling and statistical analysis of energy management and protection/control products to predict field performance and support design decisions. * Develop and execute reliability ...
SRE is part of a global organization that leverages the latest technology to communicate with our ... You'll be a key voice in observability, change management, and service scalability, providing ...
Quick apply
SRE is part of a global organization that leverages the latest technology to communicate with our ... You'll be a key voice in observability, change management, and service scalability, providing ...
Minimize risk of reliability failures related to durability, availability, performance, and ... Experience with incident management processes, on-call rotations, and post-incident review ...
Minimize risk of reliability failures related to durability, availability, performance, and ... Experience with incident management processes, on-call rotations, and post-incident review ...
Corporate Reliability Engineering Co-Op
Brampton, ON · On-site
CA$22 - CA$25/hr
... Management System (CMMS)with interpretation and trending of highly critical equipment with common ... Supporting Reliability Engineer Data collection and standardization of Failure modes and the ...
Corporate Reliability Engineering Co-Op
Brampton, ON · On-site
CA$22 - CA$25/hr
... Management System (CMMS)with interpretation and trending of highly critical equipment with common ... Supporting Reliability Engineer Data collection and standardization of Failure modes and the ...
Reliability Manager information
What is the role of a reliability manager?
Is reliability engineering in demand?
What are the key skills and qualifications needed to thrive as a reliability manager?
A Reliability Manager needs strong analytical skills, a solid background in engineering or maintenance, and experience with reliability-centered maintenance methodologies. Familiarity with tools like Failure Mode and Effects Analysis (FMEA), Root Cause Analysis (RCA), and certifications such as Certified Reliability Engineer (CRE) are often required. Leadership, problem-solving, and the ability to communicate complex technical information clearly are crucial soft skills for this role. These skills help ensure equipment uptime, optimize maintenance processes, and foster a culture of continuous improvement within the organization.
What does a reliability manager do?
A Reliability Manager is responsible for ensuring that equipment, processes, and systems operate efficiently and consistently to minimize downtime and maximize performance. They develop and implement reliability strategies, conduct root cause analyses, and oversee preventive and predictive maintenance programs. Their role involves working closely with maintenance teams, engineers, and production staff to improve asset reliability and extend equipment lifespan. Additionally, they analyze failure data, recommend improvements, and help optimize operational costs through reliability-centered maintenance practices.
What are the most commonly searched types of Reliability jobs in Ontario?
The most popular types of Reliability jobs in Ontario are:
What are popular job titles related to Reliability Manager jobs in Ontario?
For Reliability Manager jobs in Ontario, the most frequently searched job titles are:
What job categories do people searching Reliability Manager jobs in Ontario look for?
The top searched job categories for Reliability Manager jobs in Ontario are:
What cities in Ontario are hiring for Reliability Manager jobs?
Cities in Ontario with the most Reliability Manager job openings:

S&P Global rating
7.3
Based on 10 frontline employees who took The Breakroom Quiz
Job description
We are looking to hire a Site Reliability Engineer who will help in building and maintaining the observability platform across multiple business lines, helping to establish observability best practices.
What you'll be doing:
- Build and improve observability and reliability solutions that help engineering teams operate and support their services with confidence.
- Partner with engineering teams to design monitoring, alerting, dashboards, and service health standards early in the software delivery lifecycle.
- Write and maintain code, infrastructure definitions, and automation that reduce manual work and improve reliability.
- Help engineers instrument services and systems so teams can quickly detect, diagnose, and resolve issues.
- Support the adoption and standardization of telemetry patterns across metrics, logs, and traces, including OpenTelemetry-based instrumentation where appropriate.
- Improve the reliability of our AWS and Kubernetes environments, including EKS, through durable engineering solutions rather than repetitive operational work.
- Participate in incident response and follow-up activities, including troubleshooting, root cause analysis, and the implementation of lasting fixes.
- Identify opportunities to reduce toil and improve the developer experience through automation, reusable patterns, and better engineering practices.
- Continuously evaluate our tooling, reliability practices, and engineering processes for opportunities to improve.
What we're looking for:
- Experience in Site Reliability Engineering, DevOps, Platform Engineering, or Software Engineering roles, with meaningful ownership of reliability-focused solutions.
- Proven experience building, automating, and maintaining engineering solutions, not just operating existing systems.
- Experience writing production-quality code, scripts, or automation. Go is preferred; experience in other languages such as JavaScript/TypeScript or Ruby is also valuable.
- Experience managing cloud infrastructure with Infrastructure as Code. Terraform preferred.
- Experience working with AWS and Kubernetes environments, including EKS.
- Experience with distributed systems and the trade-offs involved in designing for reliability, resiliency, and durability.
- Experience with observability tooling such as Prometheus, Grafana, New Relic, CloudWatch, Google Observability, or similar platforms.
- Familiarity with telemetry standards and instrumentation patterns, including OpenTelemetry, is strongly preferred.
- Experience designing useful monitoring and alerting for applications and infrastructure, with an understanding of how to balance signal, noise, and actionable response.
- Experience with logging and telemetry pipelines at scale. Bindplane experience is a plus.
- Strong troubleshooting skills and the ability to work collaboratively during incidents to restore service and address root causes.
- Strong communication skills, with the ability to document standards, guide engineering teams, and influence reliability best practices.
- A strong bias toward automation, simplification, and reducing toil for yourself and your teammates.
Nice to have:
- Experience working with OpenSearch, Elasticsearch, ELK, or similar logging and search platforms.
- Experience with telemetry pipeline design, routing, sampling, or retention decisions.
- Experience supporting applications written in Go, JavaScript/TypeScript, or Ruby.
If you like wild growth and working with happy, enthusiastic over-achievers, you'll enjoy your career with us!
It is the policy of Mobility to provide equal employment opportunity (EEO) to all persons regardless of age, color, national origin, citizenship status, physical or mental disability, race, religion, creed, gender, sex, sexual orientation, gender identity and/or expression, genetic information, marital status, status with regard to public assistance, veteran status, or any other characteristic protected by federal, state or local law. In addition, Mobility will provide reasonable accommodations for qualified individuals with disabilities.
What S&P Global employees say
Pay
Benefits
Hours and flexibility
Workplace
Get the full story on Breakroom