Mentor and guide engineers on cloud-native technologies, site reliability engineering principles ... Experience building and managing CI/CD pipelines using GitHub Actions or similar platforms.
Mentor and guide engineers on cloud-native technologies, site reliability engineering principles ... Experience building and managing CI/CD pipelines using GitHub Actions or similar platforms.
Senior Site Reliability Engineer
Ottawa, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Quick apply
Senior Site Reliability Engineer
Ottawa, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Senior Site Reliability Engineer
Ottawa, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Quick apply
Senior Site Reliability Engineer
Ottawa, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Site Reliability Engineer
Toronto, ON · On-site +1
CA$125K - CA$250K/yr
Improve provisioning, configuration management, testing, and deployment automation * Help plan ... site reliability engineering, infrastructure engineering, systems engineering, or a related ...
Site Reliability Engineer
Toronto, ON · On-site +1
CA$125K - CA$250K/yr
Improve provisioning, configuration management, testing, and deployment automation * Help plan ... site reliability engineering, infrastructure engineering, systems engineering, or a related ...
Senior Site Reliability Engineer
Kitchener, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Quick apply
Senior Site Reliability Engineer
Kitchener, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Senior Site Reliability Engineer
Kitchener, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Quick apply
Senior Site Reliability Engineer
Kitchener, ON · Remote
$95K - $100K/yr
We are looking for an experienced Senior Site Reliability Engineer for our client. This is a permanent position that is remote to start with later relocation to Calgary or Winnipeg . Our client is a ...
Define and evolve the SRE operating model for the A&D vertical, including incident management, capacity planning, and disaster recovery strategies. Architect resilience patterns and frameworks for ...
Quick apply
Define and evolve the SRE operating model for the A&D vertical, including incident management, capacity planning, and disaster recovery strategies. Architect resilience patterns and frameworks for ...
About the role We are looking for a Site Reliability Engineer to help design and deploy, and ... AKS, GKE, or self-managed Kubernetes * Hands-on Terraform experience for infrastructure ...
About the role We are looking for a Site Reliability Engineer to help design and deploy, and ... AKS, GKE, or self-managed Kubernetes * Hands-on Terraform experience for infrastructure ...
Senior Site Reliability Engineer
Toronto, ON · On-site
CA$90K - CA$132K/yr
We enable the Managed Investment Data (MID), Reference Entity Data (RED), Fixed Income, and Third ... Collaborate with software engineers and data engineers to embed SRE best practices into the ...
Senior Site Reliability Engineer
Toronto, ON · On-site
CA$90K - CA$132K/yr
We enable the Managed Investment Data (MID), Reference Entity Data (RED), Fixed Income, and Third ... Collaborate with software engineers and data engineers to embed SRE best practices into the ...
Senior Site Reliability Engineer
Toronto, ON · On-site
CA$90K - CA$132K/yr
We enable the Managed Investment Data (MID), Reference Entity Data (RED), Fixed Income, and Third ... Collaborate with software engineers and data engineers to embed SRE best practices into the ...
Senior Site Reliability Engineer
Toronto, ON · On-site
CA$90K - CA$132K/yr
We enable the Managed Investment Data (MID), Reference Entity Data (RED), Fixed Income, and Third ... Collaborate with software engineers and data engineers to embed SRE best practices into the ...
Senior Site Reliability Engineer
CA$90K - CA$132K/yr
We enable the Managed Investment Data (MID), Reference Entity Data (RED), Fixed Income, and Third ... Collaborate with software engineers and data engineers to embed SRE best practices into the ...
Senior Site Reliability Engineer
CA$90K - CA$132K/yr
We enable the Managed Investment Data (MID), Reference Entity Data (RED), Fixed Income, and Third ... Collaborate with software engineers and data engineers to embed SRE best practices into the ...
Senior Site Reliability Engineer
Toronto, ON · Hybrid
CA$99K/yr
... Manage infrastructure deployment pipelines and troubleshoot onboarding and operational issues ... in Site Reliability, DevOps, or Cloud Engineering roles Expertise with Microsoft Azure; AWS ...
Senior Site Reliability Engineer
Toronto, ON · Hybrid
CA$99K/yr
... Manage infrastructure deployment pipelines and troubleshoot onboarding and operational issues ... in Site Reliability, DevOps, or Cloud Engineering roles Expertise with Microsoft Azure; AWS ...
You'll have the flexibility to manage your work activities within a hybrid work arrangement where you'll spend 1-3 days per week on-site, while other days will be remote. How you'll succeed
You'll have the flexibility to manage your work activities within a hybrid work arrangement where you'll spend 1-3 days per week on-site, while other days will be remote. How you'll succeed
You'll have the flexibility to manage your work activities within a hybrid work arrangement where you'll spend 1-3 days per week on-site, while other days will be remote. How you'll succeed
You'll have the flexibility to manage your work activities within a hybrid work arrangement where you'll spend 1-3 days per week on-site, while other days will be remote. How you'll succeed
Define and mature SRE practices, including Service Level Objectives (SLOs), Service Level ... Lead incident management for high-severity events, providing incident command, stakeholder ...
Quick apply
Define and mature SRE practices, including Service Level Objectives (SLOs), Service Level ... Lead incident management for high-severity events, providing incident command, stakeholder ...
Define and mature SRE practices, including Service Level Objectives (SLOs), Service Level ... Lead incident management for high-severity events, providing incident command, stakeholder ...
Define and mature SRE practices, including Service Level Objectives (SLOs), Service Level ... Lead incident management for high-severity events, providing incident command, stakeholder ...
Role Summary The Lead, Site Reliability Engineering ensures monitoring and analysis is conducted to ... manage the software development process. The Lead, DEV Platform Support Engineer will play a ...
Role Summary The Lead, Site Reliability Engineering ensures monitoring and analysis is conducted to ... manage the software development process. The Lead, DEV Platform Support Engineer will play a ...
Manage key vendor relationships, contractors, and respective performance management to enable successful execution of SRE and production support services What do you need to succeed? Must-have
Manage key vendor relationships, contractors, and respective performance management to enable successful execution of SRE and production support services What do you need to succeed? Must-have
Senior SRE/AIOps Engineer
Toronto, ON · On-site
Lead problem management and eliminate recurring issues * Develop and maintain runbooks, playbooks ... Must Have: * 3+ years of SRE or Systems Engineering experience with strong technical expertise.
Senior SRE/AIOps Engineer
Toronto, ON · On-site
Lead problem management and eliminate recurring issues * Develop and maintain runbooks, playbooks ... Must Have: * 3+ years of SRE or Systems Engineering experience with strong technical expertise.
We are seeking a Site Reliability Engineer to ensure the availability, performance, and reliability ... Create, manage, and support pipelines that the application support teams will be utilizing to ...
We are seeking a Site Reliability Engineer to ensure the availability, performance, and reliability ... Create, manage, and support pipelines that the application support teams will be utilizing to ...
Site Reliability Engineer Manager information
See Ontario salary details
$130.5K - $136.5K
6% of jobs
$136.5K - $142.5K
6% of jobs
$142.5K - $148.5K
6% of jobs
$152.5K is the 25th percentile. Wages below this are outliers.
$148.5K - $154.5K
9% of jobs
The median wage is $160.5K / yr.
$154.5K - $160.5K
22% of jobs
$160.5K - $166.5K
18% of jobs
$170.7K is the 75th percentile. Wages above this are outliers.
$166.5K - $172.5K
10% of jobs
$172.5K - $178.5K
10% of jobs
$178.5K - $184.5K
6% of jobs
$184.5K - $190.5K
4% of jobs
$190.5K - $196.5K
1% of jobs
$130.5K
$163.3K
$196.5K
How much do site reliability engineer manager jobs pay per year?
What is a site reliability engineer manager?
What are the key skills and qualifications needed to thrive as a site reliability engineer manager?
How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?
What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?
| Aspect | Site Reliability Engineer (SRE) | Site Reliability Engineer Manager |
|---|---|---|
| Responsibilities | Focuses on designing, implementing, and maintaining reliable systems and automation | Oversees SRE teams, manages projects, and aligns reliability goals with business objectives |
| Required Skills | Strong coding, system design, and troubleshooting skills | Leadership, team management, strategic planning |
| Certifications | Google Cloud, AWS certifications, Linux, scripting | Same as SRE, plus management certifications (e.g., PMP) often preferred |
| Work Environment | Technical, hands-on with systems and automation | Managerial, coordinating teams and projects |
The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.
How much do site reliability engineer managers get paid?
Is a Site Reliability Engineer Manager a stressful job?
What are the most commonly searched types of Site Reliability Engineer jobs in Ontario?
The most popular types of Site Reliability Engineer jobs in Ontario are:
What cities in Ontario are hiring for Site Reliability Engineer Manager jobs?
Cities in Ontario with the most Site Reliability Engineer Manager job openings:

Full-time
Medical, Retirement
Posted 18 days ago
Job description
This is a hands-on senior engineering role focused on improving production resilience, strengthening security, driving operational excellence, and enhancing the developer experience across the organization.
In this role, you will design, build, and evolve the foundational systems, tooling, and operational practices that enable engineering teams to ship secure, reliable, and scalable software with confidence. You will help establish reliability standards, define service level objectives (SLOs), improve observability, automate operational processes, and drive incident management and post-incident learning practices that strengthen platform stability over time.
Partnering closely with Engineering, Security, Platform, and Product teams, you will architect scalable distributed systems, optimize Kubernetes and AWS-based infrastructure, and build automated delivery pipelines that support rapid and safe software releases. You will play a key role in reducing operational toil, improving system performance, increasing platform reliability, and ensuring that our infrastructure can support continued business growth.
This is a full-time permanent position
This is an existing vacancy
Location: This is a remote location open to candidates legally authorized to work in Canada.
- Drive reliability engineering initiatives and operational excellence for mission-critical services running on AWS and Kubernetes.
- Design, implement, and continuously improve deployment, release, and rollback strategies across complex distributed systems.
- Establish secure-by-default CI/CD pipelines with robust automation, governance, and policy-driven controls.
- Enhance platform observability through metrics, logs, tracing, and actionable alerting to improve system visibility and operational efficiency.
- Define, implement, and mature Service Level Indicators (SLIs), Service Level Objectives (SLOs), and reliability standards across the organization.
- Lead response efforts for high-severity incidents, ensuring timely resolution, effective communication, and meaningful post-incident reviews that drive continuous improvement.
- Partner closely with engineering teams to strengthen platform standards, improve service resilience, optimize runtime performance, and embed reliability best practices.
- Mentor and guide engineers on cloud-native technologies, site reliability engineering principles, and operational excellence practices, fostering a culture of continuous learning and accountability.
- 8+ years of experience in Site Reliability Engineering (SRE), Platform Engineering, DevOps, or related cloud-native engineering roles.
- Deep expertise in AWS services, including EKS, IAM, VPC, Lambda, CloudFront, S3, and cloud networking/security best practices.
- Advanced experience operating and scaling production Kubernetes environments.
- Strong hands-on experience with Istio service mesh, including traffic management, security, observability, and resiliency.
- Proven expertise with Infrastructure as Code (IaC), preferably using AWS CDK.
- Experience building and managing CI/CD pipelines using GitHub Actions or similar platforms.
- Strong troubleshooting, performance optimization, and incident management experience in distributed systems.
- Excellent communication, collaboration, and technical leadership skills.
- Experience designing and operating monitoring, logging, tracing, and alerting solutions for cloud-native platforms.
- Strong knowledge of AWS CloudWatch, OpenTelemetry, AWS X-Ray, and Kubernetes observability tooling.
- Experience defining and operationalizing SLIs, SLOs, alerting strategies, runbooks, and reliability metrics.
- Proven ability to leverage observability data to improve service reliability, reduce incident impact, and optimize operational performance.
- Strong proficiency in TypeScript and Node.js for platform engineering, automation, and operational tooling.
- Experience building and maintaining scalable backend services, APIs, and event-driven systems.
- Deep understanding of Kubernetes architecture, controllers, Gateway API, ingress management, and service networking.
- Experience implementing zero-trust architectures, mTLS, and service-to-service security controls.
- Commitment to high-quality engineering practices, including automated testing, code reviews, and observability-driven development.
- Strong understanding of resilience engineering, including autoscaling, disruption management, failure testing, and safe deployment strategies.
- Experience with progressive delivery practices such as canary, blue/green, and feature-flag-based deployments.
- Experience working in regulated, compliance-driven, or security-sensitive SaaS environments.
- Familiarity with FinOps principles and cost optimization strategies for cloud platforms.
- Experience building internal developer platforms and self-service engineering tooling.
- Cloud-native certifications such as CKA, CKAD, CKS, KCSA, or KCNA.
- Kubestronaut certification or equivalent advanced Kubernetes expertise is highly regarded.
Salary Range:
The annual base salary for this position is between $140,000 CAD and $155,000 CAD per year.
This role is also eligible for discretionary bonus and/or commission, as well as other benefits. Actual pay within the listed range will be determined based on factors such as transferable skills, relevant experience, market conditions, and primary work location. The posted range is subject to change and may be updated periodically.