1

Senior Observability Engineer Jobs in Ontario (NOW HIRING)

As a Lead Data Platform Engineer, you will be a senior individual contributor and technical lead ... Ingestion, orchestration, and observability * Engineer resilient, observable ingestion patterns ...

Accelerate your career as a Senior Consultant for Data Observability with Acceldata. Engage clients ... with engineering teams for seamless solutions Requirements: • 10+ years in data platforms and ...

New

Observability, SRE & Performance Analysis Observability Architecture: Designs and scales enterprise observability solutions, leveraging metrics, logging, and tracing to provide comprehensive ...

The Role The Senior Systems Engineer will join the Systems Engineering team within the Device Data ... Define and mature systems-level observability requirements so teams can monitor feature performance ...

Drive operational excellence through observability, SRE practices, security- by-default, incident ... Excellent communication and collaboration skills with the ability to influence senior leadership ...

The Role The Senior Systems Engineer will join the Systems Engineering team within the Device Data ... Define and mature systems-level observability requirements so teams can monitor feature performance ...

As a Senior AI Engineer, you'll bring deep technical expertise to that portfolio of work, owning ... Contribute to agent observability and evaluation practices (tracing, automated evals, and quality ...

About the Role We're looking for a Senior Data/ML Engineer to refactor, operationalize, and improve ... You'll reduce technical debt, strengthen testing and observability, and raise the engineering bar ...

Senior Platform Engineer

Toronto, ON · Remote

CA$157K - CA$212K/yr

We are currently seeking a Senior Platform Engineer to join our Systems Engineering Team. This role ... We focus on automation, observability, and reliability, ensuring every engineering team at Clio can ...

Elevate Thumbtack's internal systems as a Senior Platform Engineer, focusing on backend service ... and observability • Create integration patterns that ensure safe system actions • Collaborate ...

We are seeking a Senior Software Engineer to design and build scalable cloud, Big Data, and machine ... Improve platform reliability, scalability, observability, security, and cost efficiency through ...

What We Are Looking For We're looking for a Senior Software Engineer to play a key part in building ... Observability practices: hands-on experience with monitoring tools like Datadog, implementing ...

As a Senior AI Platform Engineer, you will design and build the platform our AI agents and ... Build observability, tracing, and replay capabilities that enable rapid troubleshooting and ...

next page

Showing results 1-20

Senior Observability Engineer information

What is a senior observability engineer?

A Senior Observability Engineer is a seasoned IT professional responsible for designing, implementing, and maintaining systems that monitor and provide insights into the performance, health, and reliability of software applications and infrastructure. They utilize tools for logging, monitoring, tracing, and alerting to ensure that systems are observable and any issues can be quickly detected and resolved. In addition to technical expertise, they often collaborate with development and operations teams to establish best practices, improve incident response, and optimize system performance. Their work is crucial for maintaining uptime, enhancing customer experiences, and supporting the scalability of technology platforms.

How does a senior observability engineer typically collaborate with development and operations teams?

A Senior Observability Engineer works closely with both development and operations teams to ensure robust monitoring, logging, and tracing solutions are in place across all applications and infrastructure. They often participate in architecture discussions to advise on best practices for instrumenting code and systems for observability. By analyzing metrics and alerting patterns, they help teams proactively resolve issues and optimize system performance. This role also involves mentoring engineers on observability tools and fostering a culture of transparency and accountability in incident response.

What are the key skills and qualifications needed to thrive as a senior observability engineer, and why are they important?

To thrive as a Senior Observability Engineer, you need expertise in monitoring, logging, and tracing systems, with a solid background in computer science or a related field. Familiarity with tools like Prometheus, Grafana, ELK stack, and cloud platforms, as well as certifications such as AWS Certified DevOps Engineer, are typically required. Strong problem-solving, collaboration, and communication skills are critical for effectively diagnosing and resolving complex infrastructure issues. These skills ensure reliable system performance, rapid incident response, and continuous improvement of the technology environment.

What is the difference between Senior Observability Engineer vs Site Reliability Engineer?

AspectSenior Observability EngineerSite Reliability Engineer
CredentialsExperience with monitoring tools, scripting, cloud platformsSame as Senior Observability Engineer, often with SRE certifications
Work EnvironmentFocus on monitoring, logging, and tracing systemsFocus on system reliability, automation, and incident response
Industry UsageUsed in tech companies emphasizing system observabilityCommon in large-scale tech and cloud services
Search/Comparison IntentOften compared for monitoring rolesCompared for reliability and system stability roles

While both roles require expertise in cloud platforms and scripting, the Senior Observability Engineer primarily focuses on designing and maintaining monitoring, logging, and tracing systems to ensure system visibility. In contrast, a Site Reliability Engineer emphasizes system reliability, automation, and incident management to maintain service uptime. Both roles are vital in tech environments but serve different core functions related to system health and stability.

How much do senior observability engineers make?

Senior observability engineers typically earn between $110,000 and $160,000 annually, depending on experience, location, and company size. They often work with tools like Prometheus, Grafana, and cloud platforms, and may require advanced knowledge of monitoring, logging, and alerting systems.

What does a senior observability engineer do?

A senior observability engineer designs, implements, and maintains systems to monitor the performance and health of software applications and infrastructure. They utilize tools like Prometheus, Grafana, and ELK stack to analyze metrics, logs, and traces, ensuring system reliability and performance. This role often requires strong scripting skills and knowledge of cloud environments and distributed systems.

What are popular job titles related to Senior Observability Engineer jobs in Ontario?

For Senior Observability Engineer jobs in Ontario, the most frequently searched job titles are:

What job categories do people searching Senior Observability Engineer jobs in Ontario look for?

The top searched job categories for Senior Observability Engineer jobs in Ontario are:

What cities in Ontario are hiring for Senior Observability Engineer jobs?

Cities in Ontario with the most Senior Observability Engineer job openings:

Infographic showing various Senior Observability Engineer job openings in Ontario as of August 2026, with employment types broken down into 85% Full Time, 11% Part Time, and 4% Contract. Highlights an 86% Physical, 5% Hybrid, and 9% Remote job distribution.

Full-time

Re-posted 16 days ago


Job description

Job Description

What is the opportunity?

Are you a hands-on data platform engineer who thrives on building cloud-native, high-scale data platforms and enabling teams on top of them? Come join us!

Global Functions Technology (GFT) partners across RBC to deliver transformative platforms and solutions. In Anti-Money Laundering (AML), we are building a new Data Foundation Hub to ingest enterprise data and power analytics and controls using a medallion architecture. As a Lead Data Platform Engineer, you will be a senior individual contributor and technical lead, owning the design and build of our AWS-based data platform and mentoring other engineers.

You will work 70-80% hands-on across AWS (EKS, S3, RDS, EMR, Glue, Airflow), Snowflake, Spark, and dbt to deliver cloud-native, governed, and reliable data systems.

What will you do?

  • Technical leadership and platform ownership

    • Lead the technical directionfor the AML Data Foundation Hub on AWS.

    • Mentor and coachengineers (techdesign reviews, pair programming, standards), influencing quality and delivery.

  • Cloud-native data platform on AWS (hands-on)

    • Design and build secure, scalable dataplatforms usingAWS S3, Glue, EMR, RDS, and EKS.

    • Define patterns for data lake and warehouseintegration (e.g., S3 + Snowflake) including partitioning, storage classes,encryption, and cost optimization.

    • ImplementInfrastructure-as-Code(e.g., CloudFormation/Terraform) for repeatable environments, networking,IAM roles/policies, and security baselines.

  • Data engineering and architecture (medallion)

    • Design and build batch and incremental pipelinesacrossBronze/Silver/Goldlayers usingSnowflake (Streams, Tasks, Snowpark),Spark on EMR, anddbt.

    • Implementschema evolution, SCD/CDC, partitioning, and performance tuningacross bothcompute and storage (S3, EMR, Snowflake, RDS).

  • Ingestion, orchestration, and observability

    • Engineer resilient, observable ingestion patterns intoS3/Snowflake/RDS.

    • Orchestrate pipelines usingAirflow(or equivalent) and/or AWS-native services (e.g., event triggers), enforcingSLAs, retries, idempotency, and alerting.

    • Build operational dashboardsand alerts for pipeline health, platform capacity, and cost.

  • Reliability, DR, and security

    • Design forhigh availability, resiliency, and disaster recovery(multi-AZ,/regionbackup/restore, RPO/RTO-aware architectures).

    • Implementsecrets management, encryption, IAM least-privilege, and network security inpartnership with Security and Platform/SRE.

    • Participate in incident response and postmortems; drive root-cause fixes and hardening ofthe platform.

  • DevOps for data and platform enablement

    • Own CI/CD for data andplatform components: code review, environment promotion, automated tests (unit, integration, data contract), and versioned artifacts.

    • Partner with Platform/SRE onSLIs/SLOs, capacity planning, and platform standardization across squads.

  • Cross-functional collaboration

    • Translate AML business and control objectives intotechnical roadmaps, platform capabilities, and reusable patterns.

What do you need to succeed?

Must-have

  • Experience depth:7+ years delivering production data pipelines and distributed systems at scale on cloud platforms;demonstrated ability to operate as a senior IC and technical lead influencing architectureand quality across a team.

  • AWS platform depth:Hands-on with S3, Glue, EMR, EKS, and RDS; proficiency with IaC (CloudFormation or Terraform), IAM least-privilege design, VPC/networking, and security baselines.

  • Snowflake expertise:Hands-on with Streams, Tasks, Snowpark, and Snowpipe; strong SQL and warehouse design; performance optimization across compute and storage.

  • Distributed processing:Production experience with Spark (PySpark/Scala) for large-scale batch processing, optimization, and tuning.

  • Data engineering and architecture:Medallion architecturepatterns (Bronze/Silver/Gold), schema evolution, SCD/CDC, partitioning, and end-to-end pipeline performance tuning.

  • Orchestration and automation:Airflow (or equivalent) for DAGs, SLAs, retries, idempotency, and observability; Git-based workflows and CI/CD for data pipelines (e.g., GitHub Actions/Jenkins).

  • Reliability and security:Designingfor HA/DR (multi-AZ, backup/restore, RPO/RTO);encryption, secrets management, and network security in partnership with Platform/SRE.

  • DevOps for data:Ownership of automated testing (unit, integration, data contract), environmentpromotion, and versioned artifacts.

  • Ways of working:Strong ownership, structured problem-solving, and cleartechnical communication; experience with incident response and postmortems.

Nice-to-have

  • dbt proficiency:Development, testing, documentation, and deployment of transformations with dbt.

  • Observability:Metrics, tracing, and logging practices across data pipelines and platform components.

  • Security and privacy:OAuth2/OIDC, data masking/tokenization, PII handling, and regulatory awareness in financial services or AML.

  • Regulated domains:Prior experience in financial services or other highly regulated industries.

  • Cloud depth:AWS certifications (e.g., Solutions Architect, Data Engineer) and hands-on familiarity with SageMaker oradditional AWS-native data services.

  • Data governance:DQ frameworks, source-to-target reconciliation, lineage tooling, and purge/retention strategies.

  • Hadoop ecosystem:Exposure to legacy Hadoop stack where relevant tointegration patterns.

What's in it for you?

As a team, we thrive on the challenge to be our best, encourage progressive thinking for continued growth, and collaborate with one another to deliver trusted advice to help our clients thrive and our communities prosper. We respect and care about all of our team members and support one another in reaching our fullest potential. We work together to make a difference in our communities and to achieve success that is mutual.

This opportunity will provide you with:

  • Work in a dynamic, collaborative, progressive, and high-performing team

  • Opportunities to do challenging work, make a difference and lasting impact

  • Continuous learning and flexibility to work on projects that you are passionate about

  • Leaders who support your development through coaching and managing opportunities

#LI-POST
#TECHPJ

Job Skills

Big Data Management, Cloud Computing, Database Development, Data Mining, Data Warehousing (DW), ETL Processing, Group Problem Solving, Quality Management, Requirements Analysis

Additional Job Details

Address:

RBC CENTRE, 155 WELLINGTON ST W:TORONTO

City:

Toronto

Country:

Canada

Work hours/week:

37.5

Employment Type:

Full time

Platform:

TECHNOLOGY AND OPERATIONS

Job Type:

Regular

Pay Type:

Salaried

Posted Date:

2026-08-06

Application Deadline:

2026-08-31

Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above

Our Employment Opportunities

At RBC, we are guided by living shared values of Client First, Integrity, Collaboration, Respect and Excellence and winning together as One RBC. We believe an inclusive workplace that has diverse perspectives is core to our continued growth as one of the largest and most successful banks in the world. Maintaining a workplace where our employees feel supported to perform at their best, effectively collaborate, drive innovation, and grow professionally helps to bring our Purpose to life and create value for our clients and communities. RBC strives to deliver this through policies and programs intended to foster a workplace based on respect, belonging and opportunity for all.

Join our Talent Community
Stay in-the-know about great career opportunities at RBC. Sign up and get customized info on our latest jobs, career tips and Recruitment events that matter to you.
Expand your limits and create a new future together at RBC. Find out how we use our passion and drive to enhance the well-being of our clients and communities at jobs.rbc.com.

RBC is presently inviting candidates to apply for this existing vacancy. Applying to this posting allows you to express your interest in this current career opportunity at RBC. Qualified applicants may be contacted to review their resume in more detail.

Employment Type: FULL_TIME