1

Senior Database Reliability Engineer Jobs in Toronto, ON

Senior Site Reliability Engineer

Toronto, ON · Hybrid

CA$99K/yr

  • Medical

  • Dental

WHY THIS ROLE IS IMPORTANT TO US As a Senior Site Reliability Engineer, you will be embedded within one of our Product Areas, taking ownership of specific responsibility domains where your experience ...

Senior Data Engineer

Mississauga, ON · On-site

CA$130K - CA$140K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

As a Senior Data Engineer , you will ensure scalable and efficient data systems by addressing ... database reliability and functionality. * Mentor junior team members, fostering a culture of ...

As a SRE, you will implement, measure and gather insights from Operational Level Indicators ... You can create SQL Queries with any relational databases. * You possess technical knowledge of ...

The OPS Site Reliability Engineer will be a focal role owning and ensuring the fluent operations of ... Working with UNIX/Linux/BSD systems and/or Windows server systems as an application or database ...

Site Reliability Engineer

Mississauga, ON · Hybrid

  • Medical

  • Life

  • Retirement

  • PTO

We are seeking a Site Reliability Engineer to ensure the availability, performance, and reliability ... Database Basics: Oracle DB queries (updates, inserts, selects, deletes) * Cloud Knowledge:

New

Showing results 21-40

Senior Database Reliability Engineer information

See Toronto, ON salary details

$48.7K

$133.1K

$174.6K

How much do senior database reliability engineer jobs pay per year?

As of Aug 16, 2026, the average yearly pay for senior database reliability engineer in Toronto, ON is $133,098.00, according to ZipRecruiter salary data. Most workers in this role earn between $108,794.00 and $161,283.00 per year, depending on experience, location, and employer.

How much do senior database reliability engineers make?

Senior Database Reliability Engineers typically earn between $110,000 and $160,000 annually, depending on experience, location, and company size. They often require strong skills in database management, scripting, and cloud platforms, with certifications like AWS or Google Cloud enhancing earning potential.

What is a senior database reliability engineer?

A Senior Database Reliability Engineer (DBRE) is responsible for ensuring the reliability, scalability, and performance of database systems. They collaborate with developers, SREs, and database administrators to automate database operations, optimize queries, and implement best practices for high availability and disaster recovery. Additionally, they monitor database performance, troubleshoot incidents, and enhance observability using monitoring tools. Their role often bridges software development and database administration, emphasizing automation and reliability.

What are the key skills and qualifications needed to thrive as a senior database reliability engineer, and why are they important?

To thrive as a Senior Database Reliability Engineer, you need expertise in database administration, performance tuning, automation, and incident response, often supported by a degree in computer science or a related field. Familiarity with tools like MySQL, PostgreSQL, Oracle, scripting languages (such as Python or Bash), and experience with cloud platforms and monitoring systems (e.g., Prometheus, Datadog) are highly valuable. Strong problem-solving abilities, collaboration, and effective communication help in resolving issues and working well within cross-functional teams. These skills ensure reliable, high-performing databases that are essential for critical business operations.

What are some typical challenges faced by senior database reliability engineers in their daily work?

Senior Database Reliability Engineers often deal with challenges such as ensuring high database uptime, responding to unplanned outages, and optimizing performance under heavy loads. They may also need to manage complex migrations, automate repetitive tasks, and maintain data integrity across multiple environments. Collaboration with development, DevOps, and infrastructure teams is common to proactively identify and resolve potential reliability issues. Tackling these challenges requires both technical depth and strong communication skills, making the role both demanding and rewarding for those who enjoy problem-solving in a fast-paced environment.

Infographic showing various Senior Database Reliability Engineer job openings in Toronto, ON as of August 2026, with employment types broken down into 83% Full Time, 12% Part Time, and 5% Contract. Highlights an 86% Physical, 5% Hybrid, and 9% Remote job distribution, with an average salary of $133,098 per year, or $64 per hour.

Full-time

Posted 22 days ago


Job description

Job Description

WHAT IS THE OPPORTUNITY?

This role is responsible for designing, implementing, and maintaining SRE (Site Reliability Engineering) and AIOps (Artificial Intelligence for IT Operations) capabilities to ensure system reliability, proactive monitoring, and automation of self-healing operations. In addition to day-to-day support, the position provides end-to-end operational ownership across systems managed by multiple enterprise teams, including incident coordination, dependency management, and escalation. The role is also responsible for key security and compliance functions such as service ID and certificate management, SSO updates, vulnerability remediation, and lifecycle management of end-of-life components.

Our team supports a portfolio of multi-platform HR data pipelines that move and process data into Snowflake through multiple integrated components, requiring end-to-end monitoring, coordination, and support across systems.
In addition, we support SaaS-based applications that are primarily vendor-managed, while we retain responsibility for integration, access management, monitoring, and operational oversight.

WHAT WILL YOU DO?

Key Responsibilities:

SRE & Reliability Engineering

  • Define and operationalize SLIs, SLOs, and error budgets
  • Own the incident management lifecycle (detection triage resolution RCA prevention)
  • Lead problem management and eliminate recurring issues
  • Develop and maintain runbooks, playbooks, and recovery procedures
  • Drive resilience engineering, including failover testing and capacity planning

AIOps, Observability & Logging

  • Implement and optimize AIOps capabilities using platforms such as Moogsoft
  • Leverage Dynatrace for deep APM insights
  • Integrate alerting and escalation workflows with PagerDuty
  • Utilize synthetic monitoring via Catchpoint
  • Design and maintain centralized logging solutions using the ELK stack (Elasticsearch, Logstash, Kibana) and enterprise Logging as a Service (LaaS) platforms
  • Perform log analysis, correlation, and anomaly detection to support proactive issue identification
  • Drive event noise reduction and intelligent alerting strategies
  • Build dashboards and observability KPIs for operational insights

Automation & Self-Healing Systems

  • Design and implement automation-first solutions using Ansible and scripting (Python, Bash)
  • Enable self-healing capabilities (auto-remediation, restart logic, workflow recovery)
  • Orchestrate workflows using Stonebranch
  • Reduce operational toil through automation and continuous improvement

Application & Data Platform Support

  • Provide L2/L3 support for data pipelines and integration workflows, including Snowflake ingestion and transformation processes
  • Support Snowflake pipelines, ETL workflows, and orchestration dependencies
  • Troubleshoot across distributed systems including:
    • Object storage (e.g., S3)
    • APIs, messaging, and file transfer systems
  • Support containerized workloads on OpenShift

Infrastructure & Platform Expertise

  • Administer and support Windows Server and IIS-based applications
  • Manage certificate lifecycle (TLS, SAML, OAuth)
  • Support SSO integrations via Microsoft Entra ID
  • Work with relational databases (SQL Server, PostgreSQL) for troubleshooting and performance tuning

Disaster Recovery & Resiliency

  • Lead and coordinate Disaster Recovery (DR) planning and execution
  • Validate failover processes across dependent systems
  • Ensure end-to-end DR readiness across integrated platforms
  • Document recovery strategies and participate in DR exercises

Governance, Compliance & Collaboration

  • Ensure adherence to enterprise security, compliance, and audit requirements
  • Collaborate with IAM, PAM, logging, and cloud platform teams
  • Support change management, CAB processes, and release coordination
  • Provide reporting, KPIs, and executive summaries on system health

WHAT DO YOU NEED TO SUCCEED?

Must Have:

  • 3+ years of SRE or Systems Engineering experience with strong technical expertise.
  • Experience with ServiceNow, ITSM processes including incident, problem, change, and release management.
  • Demonstrated ability to work independently, take ownership, and drive projects to completion.
  • Knowledge of containerized platforms such as Kubernetes or OpenShift.
  • Experience working with Ansible Automation Platform or strong willingness to learn.
  • In-depth knowledge of monitoring and observability tools such as Dynatrace, Elasticsearch, Moogsoft, Catchpoint, and PagerDuty.
  • Knowledge and experience with scripting languages such as BASH, Python and PowerShell.
  • Experience with Linux and Windows Server administration.
  • Experience with or strong interest in intelligent monitoring, anomaly detection, and automation technologies
  • Excellent problem-solving skills and attention to detail.
  • Experience with API's and secure file transfer procedures.

Nice to Have:

  • Experience with telemetry standardization (OpenTelemetry) and observability data correlation.
  • Understanding of AI/ML concepts and their application to observability and operations (AIOps).
  • Experience with tools such as Ansible, Stonebranch, Kafka, and their role in system reliability.
  • Experience with CI/CD and developer platform tools such as Jenkins, GitHub, GitHub Actions, Artifactory, and Vault.
  • Familiarity with Snowflake and MongoDB Atlas, including experience with writing and executing basic queries, validating data, and troubleshooting production issues.
  • Understanding of Single Sign-On(SSO) technologies, including Microsoft Entra ID, SAML and OAuth.
  • Experience supporting enterprise applications in banking or financial services industry with understanding of regulatory, security and compliance requirements.

WHAT'S IN IT FOR YOU?

We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.

  • A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable

  • Leaders who support your development through coaching and managing opportunities

  • Ability to make a difference and lasting impact

  • Work in a dynamic, collaborative, progressive, and high-performing team

  • A world-class training program in financial services

  • Opportunities to do challenging work

#LI-POST

#TECHPJ

Job Skills

Critical Thinking, Customer Support Systems, Group Problem Solving, Installation Support, IT Service Level Management, IT Service Management (ITSM), IT Standards, Technical Troubleshooting

Additional Job Details

Address:

20 KING ST W:TORONTO

City:

Toronto

Country:

Canada

Work hours/week:

37.5

Employment Type:

Full time

Platform:

TECHNOLOGY AND OPERATIONS

Job Type:

Regular

Pay Type:

Salaried

Posted Date:

2026-07-23

Application Deadline:

2026-08-23

Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above

Our Employment Opportunities

At RBC, we are guided by living shared values of Client First, Integrity, Collaboration, Respect and Excellence and winning together as One RBC. We believe an inclusive workplace that has diverse perspectives is core to our continued growth as one of the largest and most successful banks in the world. Maintaining a workplace where our employees feel supported to perform at their best, effectively collaborate, drive innovation, and grow professionally helps to bring our Purpose to life and create value for our clients and communities. RBC strives to deliver this through policies and programs intended to foster a workplace based on respect, belonging and opportunity for all.

Join our Talent Community
Stay in-the-know about great career opportunities at RBC. Sign up and get customized info on our latest jobs, career tips and Recruitment events that matter to you.
Expand your limits and create a new future together at RBC. Find out how we use our passion and drive to enhance the well-being of our clients and communities at jobs.rbc.com.

RBC is presently inviting candidates to apply for this existing vacancy. Applying to this posting allows you to express your interest in this current career opportunity at RBC. Qualified applicants may be contacted to review their resume in more detail.

Employment Type: FULL_TIME