1

Site Reliability Engineer Manager Jobs in Ontario

As a Site Reliability Engineer, responsible for the operational stability, reliability, and support ... These mission-critical applications are primarily vendor-managed platforms and support business ...

New

As a Site Reliability Engineer, responsible for the operational stability, reliability, and support ... These mission-critical applications are primarily vendor-managed platforms and support business ...

New

Site Reliability Engineer SALARY: 48,987 - 55,000 LOCATION(S): Edinburgh, Manchester or Leeds HOURS ... Support incident management, root cause analysis and continuous improvement activities, helping to ...

Site Reliability Engineer SALARY: 48,987 - 55,000 LOCATION(S): Edinburgh, Manchester or Leeds HOURS ... Support incident management, root cause analysis and continuous improvement activities, helping to ...

Site Reliability Engineer SALARY: 48,987 - 55,000 LOCATION(S): Edinburgh, Manchester or Leeds HOURS ... Support incident management, root cause analysis and continuous improvement activities, helping to ...

Site Reliability Engineer SALARY: 48,987 - 55,000 LOCATION(S): Edinburgh, Manchester or Leeds HOURS ... Support incident management, root cause analysis and continuous improvement activities, helping to ...

Implement SRE best practices across production support, incident management, problem management, change management, release, and deployment processes. * Build and maintain automation scripts ...

Site Reliability Engineer

Toronto, ON · Hybrid

CA$100K - CA$125K/yr

As a Site Reliability Engineer, you will play a crucial role in enhancing the reliability ... supplier management, tax compliance, and treasury. Tipalti partners with leading financial ...

Senior Site Reliability Engineer SALARY: 72,702 - 82,000 LOCATION(S): Edinburgh, Manchester, Leeds HOURS: Full-time WORKING PATTERN: Our work style is hybrid, which involves spending at least two ...

We are looking to hire a Site Reliability Engineer who will help in building and maintaining the ... Experience managing cloud infrastructure with Infrastructure as Code. Terraform preferred.

Senior Site Reliability Engineer SALARY: 72,702 - 82,000 LOCATION(S): Edinburgh, Manchester, Leeds HOURS: Full-time WORKING PATTERN: Our work style is hybrid, which involves spending at least two ...

Senior Site Reliability Engineer SALARY: 72,702 - 82,000 LOCATION(S): Edinburgh, Manchester, Leeds HOURS: Full-time WORKING PATTERN: Our work style is hybrid, which involves spending at least two ...

Senior Site Reliability Developer

Toronto, ON · On-site

CA$107K - CA$157K/yr

... (SRE) to manage critical cloud infrastructure and site reliability operations for the Autodesk Platform Services and Emerging Technologies organization. The team delivers high-value, exabyte-scale ...

next page

Showing results 1-20

Site Reliability Engineer Manager information

See Ontario salary details

$130.5K

$163.3K

$196.5K

How much do site reliability engineer manager jobs pay per year?

As of Sep 8, 2026, the average yearly pay for site reliability engineer manager in Ontario is $163,333.00, according to ZipRecruiter salary data. Most workers in this role earn between $153,000.00 and $173,500.00 per year, depending on experience, location, and employer.

What is a site reliability engineer manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.

How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How much do site reliability engineer managers get paid?

Site Reliability Engineer Managers typically earn between $120,000 and $180,000 annually, depending on experience, location, and company size. They often oversee teams responsible for system reliability, incident response, and infrastructure automation, requiring strong leadership and technical skills.

Is a Site Reliability Engineer Manager a stressful job?

A Site Reliability Engineer Manager role can be stressful due to the responsibility of maintaining system uptime, managing incident responses, and ensuring reliability across complex infrastructure. The job often involves working under pressure, handling outages, and coordinating teams, but it also offers opportunities for problem-solving and leadership. Stress levels vary depending on company size, team structure, and workload management skills.

What are the most commonly searched types of Site Reliability Engineer jobs in Ontario?

The most popular types of Site Reliability Engineer jobs in Ontario are:

What cities in Ontario are hiring for Site Reliability Engineer Manager jobs?

Cities in Ontario with the most Site Reliability Engineer Manager job openings:

Infographic showing various Site Reliability Engineer Manager job openings in Ontario as of August 2026, with employment types broken down into 84% Full Time, 14% Part Time, and 2% Contract. Highlights an 79% Physical, 3% Hybrid, and 18% Remote job distribution, with an average salary of $163,333 per year, or $78.5 per hour.

Site Reliability Engineer

CIBC US

Toronto, ON • Hybrid

Full-time

Retirement

Posted 3 days ago

New


Job description

We're building a relationship-oriented bank for the modern world. We need talented, passionate professionals who are dedicated to doing what's right for our clients.

At CIBC, we embrace your strengths and your ambitions, so you are empowered at work. Our team members have what they need to make a meaningful impact and are truly valued for who they are and what they contribute.

To learn more about CIBC, please visit CIBC.com

What You'll Be Doing

You'll join CIBC's Wealth Technology team. As a Site Reliability Engineer, responsible for the operational stability, reliability, and support of Tier 1 Wealth Management Trading platforms. These mission-critical applications are primarily vendor-managed platforms and support business operations across trading, investment management, and client servicing functions. You will act as a senior reliability and production support specialist, ensuring high availability, resiliency, and performance of critical applications. You will work closely with business stakeholders, technology teams, and external vendors to drive operational excellence, incident resolution, continuous improvement, and implementation of Site Reliability Engineering (SRE) practices.This role requires participation in 24x7 production support, major incident management, change coordination and implementation, problem management, and recovery coordination activities.


At CIBC we enable the work environment most optimal for you to thrive in your role. Details on your work arrangement (proportion of on-site and remote work) will be discussed at the time of your interview.

How you'll succeed

  • Application Reliability & SRE: Drive the adoption and execution of Site Reliability Engineering (SRE) practices to improve platform availability, resiliency, observability, and operational excellence. Define, monitor, and report reliability metrics, service level objectives (SLOs), operational health indicators, and performance trends. Proactively identify reliability concerns and implement automation, monitoring, alerting, and remediation solutions to minimize service disruptions. Lead root cause investigations and reliability improvement initiatives to reduce recurring incidents and operational inefficiencies. Continuously improve application resiliency through incident reviews, problem management, trend analysis, and operational readiness assessments.
  • Production Support & Service Management: Provide senior-level production support for Tier 1 Wealth Management Trading applications operating in a 24x7 environment. Lead and coordinate resolution of critical production incidents, service disruptions, and business escalations. Apply strong Incident, Problem, and Change Management practices to maintain stable and highly available services. Conduct impact assessments, operational readiness reviews, and post-implementation validation for technology changes and releases. Identify opportunities for process optimization and operational efficiency improvements.
  • Vendor & Stakeholder Management: Act as the primary operational contact for strategic vendor partners supporting Wealth Trading applications. Manage vendor performance, service quality, issue resolution, escalation management, and operational accountability. Collaborate closely with internal technology teams, infrastructure teams, business partners, project teams, and external vendors to ensure effective delivery and support. Facilitate regular service reviews and operational governance discussions with key stakeholders. Ensure vendor-related incidents and service issues are addressed within established service levels and business expectations.
  • Trading Platform & Business Partnership: TDevelop deep understanding of Wealth Management products and trading processes, including Equities, Bonds, GICs, Mutual Funds, and related operational workflows. Translate business requirements and operational challenges into sustainable technology solutions. Partner with business stakeholders to identify opportunities to improve application reliability, user experience, and service delivery. Support strategic initiatives, upgrades, regulatory changes, and business transformation programs impacting trading platforms.
  • Change & Continuous Improvement: Serve as a subject matter expert for production readiness, release management, and operational risk assessments. Support implementation of monitoring, automation, and operational tooling improvements. Develop and maintain support documentation, operational procedures, runbooks, and recovery processes. Drive continuous service improvement initiatives to enhance stability, supportability, and operational maturity. Develop deep understanding of Wealth Management products and trading processes, including Equities, Bonds, GICs, Mutual Funds, and related operational workflows. Translate business requirements and operational challenges into sustainable technology solutions. Partner with business stakeholders to identify opportunities to improve application reliability, user experience, and service delivery. Support strategic initiatives, upgrades, regulatory changes, and business transformation programs impacting trading platforms.


Who you are

  • Experience & Expertise: You can demonstrate 5+years of experience in Application Support, Application Reliability Engineering, Site Reliability Engineering (SRE), Production Operations, or Technology Operations within a financial institution. You have experience supporting Tier 1 mission-critical applications operating in a 24x7 production environment. You possess strong knowledge of Incident, Problem, Change, and Major Incident Management processes. You have experience working with and managing external technology vendors supporting enterprise applications. You have experience implementing or operating SRE practices, including monitoring, observability, automation, reliability measurement, and service resiliency improvements.
  • Business & Industry Knowledge: You have strong knowledge of Wealth Management and Trading platforms supporting Equities, Bonds, GICs, Mutual Funds, and related investment products. You understand the operational and business impact of technology disruptions within trading and investment management environments. You can effectively communicate technical issues and business impacts to both technical and non-technical stakeholders.
  • Technical & Analytical Skills: You're digitally savvy and continuously seek innovative solutions to improve reliability, efficiency, and service quality. You possess strong troubleshooting, analytical, and problem-solving skills with the ability to quickly assess complex production issues. You are skilled in monitoring, observability, automation, and operational support tools. You proactively identify opportunities to reduce manual effort and improve service reliability through automation and continuous improvement.
  • Values Matter: You bring your authentic self to work and live our values of Trust, Teamwork, and Accountability. You put clients first and understand the importance of maintaining highly available systems that support critical business operations and client experiences.

#LI-TA

What CIBC Offers

At CIBC, your goals are a priority. We start with your strengths and ambitions as an employee and strive to create opportunities to tap into your potential. We aspire to give you a career, rather than just a paycheck.

  • We work to recognize you in meaningful, personalized ways including a competitive salary, incentive pay, banking benefits, a benefits program*, defined benefit pension plan*, an employee share purchase plan, a vacation offering, wellbeing support, and MomentMakers, our social, points-based recognition program.

  • Our spaces and technological toolkit will make it simple to bring together great minds to create innovative solutions that make a difference for our clients.

  • We cultivate a culture where you can express your ambition through initiatives like Purpose Day; a paid day off dedicated for you to use to invest in your growth and development.

*Subject to plan and program terms and conditions

What you need to know

  • CIBC is committed to creating an inclusive environment where all team members and clients feel like they belong. We seek applicants with a wide range of abilities and we provide an accessible candidate experience. If you need accommodation, please contact Mailbox.careers-carrieres@cibc.com

  • CIBC is committed to clarity in our hiring process. All roles posted are opportunities we're actively recruiting for, unless stated otherwise.

  • You need to be legally eligible to work at the location(s) specified above and, where applicable, must have a valid work or study permit.

  • We may ask you to complete an attribute-based assessment and other skills test (such as simulation, coding, French proficiency).

  • We use artificial intelligence tools during the recruitment process. Our goal for the application process is to get to know more about you, all that you have to offer, and give you the opportunity to learn more about us.

Job Location

Toronto-81 Bay, 19th Floor

Employment Type

Regular

Weekly Hours

37.5

Skills

Analytical Thinking, Application Production Support, Business Operations, Change Management, Impact Analysis, Implementation Planning, Incident Resolution, IT Operations Support, IT Vendor Management, Operational Efficiency, Problem Management, Reliability Management, Resiliency, Service Levels, Site Reliability Engineering, System Reliability, Technical Knowledge