1

Reliability Engineer Jobs in Ontario (NOW HIRING)

Reliability Engineering Specialist

Brantford, ON ยท On-site

CA$95K - CA$125K/yr

As a Reliability Engineering Specialist, you will play a key role in improving asset reliability, reducing downtime, and optimizing maintenance strategies across the site. Working closely with ...

The OPS Site Reliability Engineer will be a focal role owning and ensuring the fluent operations of Managed Services offerings in the KPMG production cloud environment. The role will be focusing on ...

The OPS Site Reliability Engineer will be a focal role owning and ensuring the fluent operations of Managed Services offerings in the KPMG production cloud environment. The role will be focusing on ...

We follow the Network Reliability Engineer (NRE) model: continuous observation, rapid capacity adaptation, and relentless automation so that network capabilities are delivered as APIs, not tickets.

Job Summary The Director, Reliability Engineering is responsible for implementing and leading strategies to maximize uptime, performance and lifespan of assets, and reducing likelihood of failure by ...

Elevate cloud platform reliability with Manulife as a Senior Platform Reliability Engineer. Lead Azure ecosystem enhancements, focusing on Kubernetes, CI/CD pipelines, and AI technologies. In this ...

Define and mature SRE practices, including Service Level Objectives (SLOs), Service Level Indicators (SLIs), error budgets, production readiness standards, and operational acceptance criteria for ...

Define and mature SRE practices, including Service Level Objectives (SLOs), Service Level Indicators (SLIs), error budgets, production readiness standards, and operational acceptance criteria for ...

Senior Platform Reliability Engineer

Toronto, ON ยท On-site

CA$113K - CA$163K/yr

Senior Platform Reliability Engineer Join Manulife Global Wealth & Asset Management (GWAM) and help power critical cloud platforms that enable data, analytics, and AI across the organization. We're ...

Showing results 41-60

Reliability Engineer information

What are some typical challenges reliability engineers face when implementing preventive maintenance strategies?

Reliability Engineers often encounter challenges such as balancing preventive maintenance schedules with production demands, ensuring buy-in from operations teams, and accurately predicting equipment failures. They must analyze large sets of historical data to identify trends and root causes, which can be complex in facilities with diverse machinery. Collaboration with maintenance, operations, and engineering teams is essential to develop effective strategies that minimize downtime while optimizing resources.

How much do reliability engineers get paid?

Reliability engineers typically earn a median annual salary ranging from $70,000 to $110,000, depending on experience, location, and industry. Senior or specialized reliability engineers with certifications and advanced skills can earn higher salaries, often exceeding $120,000 annually.

What are the key skills and qualifications needed to thrive as a reliability engineer, and why are they important?

To thrive as a Reliability Engineer, you need a solid background in engineering principles, failure analysis, and reliability modeling, typically with a degree in engineering or a related field. Familiarity with tools such as FMEA, Root Cause Analysis (RCA), reliability-centered maintenance (RCM) software, and certifications like Certified Reliability Engineer (CRE) are highly valued. Strong problem-solving abilities, attention to detail, and effective communication are crucial soft skills in this role. These skills ensure systems are dependable, downtime is minimized, and organizational performance and safety are optimized.

What is the difference between Reliability Engineer vs Maintenance Engineer?

AspectReliability EngineerMaintenance Engineer
CredentialsTypically requires engineering degree, certifications in reliability or asset managementOften requires engineering or technical diploma, certifications in maintenance or equipment repair
Work EnvironmentFocuses on analysis, design, and improvement of systems for reliabilityHands-on maintenance, repair, and troubleshooting of equipment
Industry UsageCommon in manufacturing, energy, aerospace, and industrial sectorsPrevalent in manufacturing, facilities, and industrial plants

Reliability Engineers focus on designing and improving systems to prevent failures, using data analysis and modeling. Maintenance Engineers perform hands-on repairs and upkeep of equipment to ensure operational continuity. While both roles aim to optimize equipment performance, Reliability Engineers work proactively on system reliability, whereas Maintenance Engineers handle reactive and scheduled maintenance tasks.

What does a reliability engineer do?

As a reliability engineer, your duties are to test and evaluate the manufacturing of products and components and ensure that the procedures are efficient and do not lead to abnormally high maintenance or operational costs. Your other responsibilities are to find solutions to product reliability risks. You may manage risk in a supply chain, develop loss prevention strategies, and track the entire lifecycle of product development, from building prototypes to moving a product into full-scale production. You analyze information from department heads and recommend strategies to reduce risk and ensure that the product works reliably.

What is a reliability engineer?

Reliability Engineers are professionals responsible for ensuring that systems, equipment, or processes function consistently and efficiently over time. They analyze data, identify potential points of failure, and develop maintenance strategies to improve system reliability and minimize downtime. Their work spans various industries, including manufacturing, energy, and technology, and often involves collaborating with design, operations, and maintenance teams. By implementing reliability-centered maintenance and predictive analysis, they help organizations save costs and increase safety.
What are the most commonly searched types of Reliability Engineer jobs in Ontario? The most popular types of Reliability Engineer jobs in Ontario are:
What job categories do people searching Reliability Engineer jobs in Ontario look for? The top searched job categories for Reliability Engineer jobs in Ontario are:
Infographic showing various Reliability Engineer job openings in Ontario as of July 2026, with employment types broken down into 91% Full Time, 6% Part Time, and 3% Contract. Highlights an 87% Physical, 4% Hybrid, and 9% Remote job distribution.

Senior Site Reliability Engineer, AI Infrastructure

PointClickCare

Mississauga, ON โ€ข Hybrid

Full-time

Medical, Life, Retirement, PTO

Re-posted 17 days ago


Job description

At PointClickCare our mission is simple: to help providers deliver exceptional care. And that starts with our people. As a leading health tech company that's founder-led and privately held, we empower our employees to push boundaries, innovate, and shape the future of healthcare.

With the largest long-term and post-acute care dataset and a Marketplace of 400+ integrated partners, our platform serves over 30,000 provider organizations, making a real difference in millions of lives. We also reinvest a significant percentage of our revenue back into research and development, ensuring our employees have the resources to innovate and make a lasting impact.ย Recognized by Forbes as a top private cloud company and honored as one of Canada's Most Admired Corporate Cultures, we offer flexibility, growth opportunities, and meaningful work.ย 

At PointClickCare, we empower our people to be the architects of a smarter healthcare future; one that is human-first and accelerated by AI to create meaningful and lasting change. Employees harness AI as a catalyst for creativity, productivity, and thoughtful decision-making. By integrating AI tools into our daily workflows, collaboration is enhanced, outcomes are improved, and every team member has the proficiency to maximize their impact. It all starts with our hiring practices where we uncover AI expertise that complements our mission, and we continue to invest in training and development to nurture innovation throughout the employee journey.

Join us in redefining healthcare - so it doesn't just survive, it thrives.ย To learn more about PointClickCare, check out Life at PointClickCareย and connect with us on Glassdoor and LinkedIn.


**Travel to Office expectations**
For Remote Roles: If this role is remote, there will be in-office events that will require travel to and from the Mississauga and/or Salt Lake City office. These will include, but not limited to, onboarding, team events, semi-annual and annual team meetings.

For Hybrid Roles: If this role is Hybrid, there will be an expectation to reside within commutable distance to the office/location specified in the job listing. This will include, but not limited to, weekly/bi-weekly/monthly events in the office with your specific team. This is a requirement for this role.

Team Summary:

The AI SRE team is a focused group of SRE engineers dedicated to making PointClickCare's AI and ML platforms reliable, secure, and operationally excellent - from data processing and ML workspaces to model serving and labeling systems. We treat reliability as a product - prioritizing observability, automation, and safe operations so that data scientists and ML engineers can focus on building AI capabilities that improve patient care. You will spend a significant portion of your time hands-on - building automation, designing guardrails, hardening platforms, and leading incident response across cloud environments like Databricks, Azure AI suites. The team collaborates closely with research, platform, data, and security teams across the AI engineering organization.

Job Summary:

AI SRE exists to ensure PointClickCare's AI platforms run safely, reliably, and efficiently - protecting patient data while enabling teams to move fast with confidence. We solve complex cross-cutting reliability and security problems through well-designed automation, SLOs, and operational guardrails - so that product and research teams can focus on delivering AI-driven value to clinicians and patients. We own the infrastructure operability of AI data processing, ML workspaces, labeling systems, and model serving - and we drive the infrastructure observability, incident response, compliance controls, and cost optimization that keep those platforms healthy through sound SRE practices and a security-first mindset

Key responsibilities:

  • Own service level objectives, error budgets, and reliability targets for the infrastructure underpinning cloud-based platforms - ensuring infrastructure observability (metrics, logs, traces), alert quality, and telemetry completeness across platform components and serving endpoints
  • Design, build, and maintain infrastructure-as-code, operational automation, and change control workflows for AI/ML platforms - with a focus on repeatability, consistency, and toil reduction
  • Implement and maintain platform security controls - including network segmentation, secrets management, encryption, and data protection safeguards - aligned to compliance requirements and partnering with security teams to respond to emerging risks
  • Lead incident response and blameless postmortems; validate backup/restore and disaster recovery processes; conduct game days and resiliency testing to harden platform and infrastructure reliability
  • Mentor engineers, influence design reviews, and collaborate across engineering teams to improve platform resiliency, cost efficiency, capacity planning, and operational standards
Qualification and Skills:
Minimum:
  • 5+ years in SRE, platform engineering, or infrastructure roles supporting production cloud environments and mission-critical applications
  • Strong proficiency with observability - metrics, logging, distributed tracing, SLI/SLO frameworks - and production ownership including incident response, blameless postmortems, and on-call operations
  • Strong proficiency with Infrastructure as Code (Terraform), GitOps practices, and CI/CD for infrastructure and platform changes
  • Working proficiency with cloud platform administration - compute, networking, storage, and operating managed data or AI/ML platform services in production (e.g., Databricks, Azure ML, or Kubernetes-hosted infrastructure)
  • Working proficiency with platform security - network segmentation, secrets management, encryption at rest and in transit, and key management
  • Strong programming skills for automation, operational tooling, and infrastructure management
  • Strong communication and documentation skills - able to write runbooks, lead postmortems, influence operational standards across teams, and translate technical complexity for diverse audiences
Preferred:ย 
  • Experience with disaster recovery planning, multi-region patterns, and capacity or cost optimization (FinOps)
  • ย Working knowledge of container orchestration (Kubernetes), progressive delivery patterns (blue/green, canary), and data lineage tooling
  • Working knowledge of container orchestration (Kubernetes), progressive delivery patterns (blue/green, canary), and data lineage tooling
  • Experience in healthcare, life sciences, or other highly regulated industries with data privacy requirements
$139,000 - $155,000 a year
At PointClickCare, base salary is one of the many components that make up our total rewards package. The CAD base salary range for this position is $139,000-$155,000 + bonus + benefits. Compensation is assessed individually and aligned to experience, skills, and market context. The posted range reflects typical expectations for this role.ย 
PointClickCare Benefits & Perks:

Benefits starting from Day 1!
Retirement Plan Matching
Flexible Paid Time Off
Wellness Support Programs and Resources
Parental & Caregiver Leaves
Fertility & Adoption Support
Continuous Development Support Program
Employee Assistance Program
Allyship and Inclusion Communities
Employee Recognition ... and more!

It is the policy of PointClickCare to ensure equal employment opportunity without discrimination or harassment on the basis of race, religion, national origin, status, age, sex, sexual orientation, gender identity or expression, marital or domestic/civil partnership status, disability, veteran status, genetic information, or any other basis protected by law. PointClickCare welcomes and encourages applications from people with disabilities. Accommodations are available upon request for candidates taking part in all aspects of the selection process. Please contact [emailย protected] should you require any accommodations. As part of our commitment to a streamlined and equitable hiring experience, PointClickCare uses AI tools to assist with candidate screening and assessment.

When you apply for a position, your information is processed and stored with Lever, in accordance with Lever's Privacy Policy. We use this information to evaluate your candidacy for the posted position. We also store this information, and may use it in relation to future positions to which you apply, or which we believe may be relevant to you given your background. When we have no ongoing legitimate business need to process your information, we will either delete or anonymize it.ย  If you have any questions about how PointClickCare uses or processes your information, or if you would like to ask to access, correct, or delete your information, please contact PointClickCare's human resources team: [emailย protected]ย 

PointClickCare is committed to Information Security. By applying to this position, if hired, you commit to following our information security policies and procedures and making every effort to secure confidential and/or sensitive information.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
apply for this job