This is a full-time permanent position This is an existing vacancy Location: This is a remote ... Mentor and guide engineers on cloud-native technologies, site reliability engineering principles ...
This is a full-time permanent position This is an existing vacancy Location: This is a remote ... Mentor and guide engineers on cloud-native technologies, site reliability engineering principles ...
Site Reliability Engineer
Toronto, ON ยท On-site +1
CA$125K - CA$250K/yr
Based in Toronto or remote, you will work across the systems that enable large-scale AI training ... site reliability engineering, infrastructure engineering, systems engineering, or a related ...
Site Reliability Engineer
Toronto, ON ยท On-site +1
CA$125K - CA$250K/yr
Based in Toronto or remote, you will work across the systems that enable large-scale AI training ... site reliability engineering, infrastructure engineering, systems engineering, or a related ...
The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability, scalability ... Remote Work Environment * Flexible Time Away From Work Policy including PTO, Personal and Sick Days
The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability, scalability ... Remote Work Environment * Flexible Time Away From Work Policy including PTO, Personal and Sick Days
DevOps / SRE Engineer (Remote)
Toronto, ON ยท On-site +1
To do that we are eager to add a highly skilled DevOps / SRE Engineer Engineer to our incredible ... However, we will consider remote applicants +/- 3 hours from eastern time zone. Why work here * We ...
Quick apply
DevOps / SRE Engineer (Remote)
Toronto, ON ยท On-site +1
To do that we are eager to add a highly skilled DevOps / SRE Engineer Engineer to our incredible ... However, we will consider remote applicants +/- 3 hours from eastern time zone. Why work here * We ...
DevOps / SRE Engineer (Remote)
Toronto, ON ยท On-site +1
To do that we are eager to add a highly skilled DevOps / SRE Engineer Engineer to our incredible ... However, we will consider remote applicants +/- 3 hours from eastern time zone. Why work here * We ...
Quick apply
DevOps / SRE Engineer (Remote)
Toronto, ON ยท On-site +1
To do that we are eager to add a highly skilled DevOps / SRE Engineer Engineer to our incredible ... However, we will consider remote applicants +/- 3 hours from eastern time zone. Why work here * We ...
At Newton, you'll work with a remote team spread across Canada, but you'll never feel distant ... Role Overview: We're looking for a Site Reliability Engineer to improve the reliability, resilience ...
At Newton, you'll work with a remote team spread across Canada, but you'll never feel distant ... Role Overview: We're looking for a Site Reliability Engineer to improve the reliability, resilience ...
Cloud Engineer - Fully Remote | Upto $85/hr
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
Quick apply
Cloud Engineer - Fully Remote | Upto $85/hr
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
The Security Engineering Team is at the forefront of protecting Fireflies.ai and its users by ensuring the security of our infrastructure and data. We are looking for a creative hustler who is not ...
The Security Engineering Team is at the forefront of protecting Fireflies.ai and its users by ensuring the security of our infrastructure and data. We are looking for a creative hustler who is not ...
Customer Reliability Engineer
Toronto, ON ยท Remote
$140K/yr
... site reliability engineering, infrastructure or platform services operations. * +4 hands-on ... Remote first work environment * Flexible work hours & location * Paid parental leave options Health ...
Customer Reliability Engineer
Toronto, ON ยท Remote
$140K/yr
... site reliability engineering, infrastructure or platform services operations. * +4 hands-on ... Remote first work environment * Flexible work hours & location * Paid parental leave options Health ...
DevOps Engineer - AI Model Evaluator
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
Quick apply
DevOps Engineer - AI Model Evaluator
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
DevOps Engineer - AI Model Evaluator
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
Quick apply
DevOps Engineer - AI Model Evaluator
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
DevOps Engineer - AI Model Evaluator
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
Quick apply
DevOps Engineer - AI Model Evaluator
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
DevOps Engineer - AI Model Evaluator
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
Quick apply
DevOps Engineer - AI Model Evaluator
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
DevOps Engineer - AI Model Evaluator
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
Quick apply
DevOps Engineer - AI Model Evaluator
Toronto, ON ยท Remote
CA$85/hr
Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience) Type: Contract Compensation: $85/hour Location: Remote Role Responsibilities * Use frontier AI coding agents to complete and evaluate ...
... under an SRE title - this role will not be a fit. * If your observability and reliability ... Remote
New
... under an SRE title - this role will not be a fit. * If your observability and reliability ... Remote
New
... under an SRE title -- this role will not be a fit. * If your observability and reliability ... Remote Thank you for considering this opportunity. Funded.club Senior Recruiters partner ...
New
Quick apply
... under an SRE title -- this role will not be a fit. * If your observability and reliability ... Remote Thank you for considering this opportunity. Funded.club Senior Recruiters partner ...
New
Incident Management Expert - Evaluator
Toronto, ON ยท Remote
CA$80 - CA$120/hr
Incident management / reliability / SRE Evaluator Type: Contract Compensation: $80-$120/hour Location: Remote Role Responsibilities * Evaluate AI-generated artifacts against domain-specific quality ...
Quick apply
Incident Management Expert - Evaluator
Toronto, ON ยท Remote
CA$80 - CA$120/hr
Incident management / reliability / SRE Evaluator Type: Contract Compensation: $80-$120/hour Location: Remote Role Responsibilities * Evaluate AI-generated artifacts against domain-specific quality ...
Senior / Staff Software Engineer (Observability / SRE)
Toronto, ON ยท On-site +1
CA$148K - CA$249K/yr
... full-time employees only). - Unlimited Vacation. - Flexible hours and Work from Home support ... site & virtually. - As we grow, this list continues to evolve! Waabi is a technology start-up ...
Senior / Staff Software Engineer (Observability / SRE)
Toronto, ON ยท On-site +1
CA$148K - CA$249K/yr
... full-time employees only). - Unlimited Vacation. - Flexible hours and Work from Home support ... site & virtually. - As we grow, this list continues to evolve! Waabi is a technology start-up ...
Senior Platform Engineer
Toronto, ON ยท On-site +1
CA$100K - CA$150K/yr
The SPS SRE team is responsible for delivering highly available platform services and deployment ... Location: This role follows a remote work model for candidates who are based in Canada. What We ...
Senior Platform Engineer
Toronto, ON ยท On-site +1
CA$100K - CA$150K/yr
The SPS SRE team is responsible for delivering highly available platform services and deployment ... Location: This role follows a remote work model for candidates who are based in Canada. What We ...
Manager, Network Reliability and Resiliency
Toronto, ON ยท On-site +1
CA$125K - CA$220K/yr
Provide training and support to partner teams that interface with SRE. * Onboarding of new hires to ... Work personas (flexible, remote, or required in office) are categories that are assigned to ...
Manager, Network Reliability and Resiliency
Toronto, ON ยท On-site +1
CA$125K - CA$220K/yr
Provide training and support to partner teams that interface with SRE. * Onboarding of new hires to ... Work personas (flexible, remote, or required in office) are categories that are assigned to ...
Full Time Site Reliability Engineer Remote information
What is a full time site reliability engineer?
What are the typical challenges a remote site reliability engineer faces and how can they be managed?
What are the key skills and qualifications needed to thrive as a full time site reliability engineer, and why are they important?
What is the difference between Full Time Site Reliability Engineer Remote vs Cloud Operations Engineer?
| Aspect | Full Time Site Reliability Engineer Remote | Cloud Operations Engineer |
|---|---|---|
| Credentials | Typically requires SRE certifications, Linux, and cloud platform knowledge | Often requires cloud certifications (AWS, Azure), Linux, and scripting skills |
| Work Environment | Remote, collaborative teams focusing on system reliability and automation | Remote or on-site, managing cloud infrastructure and deployment pipelines |
| Industry Usage | Tech, finance, e-commerce companies emphasizing system uptime | Cloud service providers, SaaS companies, enterprises with cloud infrastructure |
While both roles focus on cloud and infrastructure management, a Full Time Site Reliability Engineer Remote emphasizes system reliability, automation, and incident response, often in a remote setting. A Cloud Operations Engineer concentrates on cloud infrastructure deployment, management, and optimization, which may include on-site or remote work. Both roles require cloud and Linux skills but differ slightly in focus and daily responsibilities.
What are popular job titles related to Full Time Site Reliability Engineer Remote jobs in Toronto, ON?
For Full Time Site Reliability Engineer Remote jobs in Toronto, ON, the most frequently searched job titles are:
- Site Reliability Engineer
- Site Reliability Engineer Remote
- Sre Internship
- Site Reliability Engineer Intern
- Entry Level Site Reliability Engineer
- Part Time Site Reliability Engineer
- Senior Site Reliability Engineer
- Freelance Mechanical Reliability Engineer
- Remote Reliability Engineer
- Remote Site Reliability Engineer Intern
What job categories do people searching Full Time Site Reliability Engineer Remote jobs in Toronto, ON look for?
The top searched job categories for Full Time Site Reliability Engineer Remote jobs in Toronto, ON are:
Full-time
Medical, Retirement
Posted 5 days ago
Job description
This is a hands-on senior engineering role focused on improving production resilience, strengthening security, driving operational excellence, and enhancing the developer experience across the organization.
In this role, you will design, build, and evolve the foundational systems, tooling, and operational practices that enable engineering teams to ship secure, reliable, and scalable software with confidence. You will help establish reliability standards, define service level objectives (SLOs), improve observability, automate operational processes, and drive incident management and post-incident learning practices that strengthen platform stability over time.
Partnering closely with Engineering, Security, Platform, and Product teams, you will architect scalable distributed systems, optimize Kubernetes and AWS-based infrastructure, and build automated delivery pipelines that support rapid and safe software releases. You will play a key role in reducing operational toil, improving system performance, increasing platform reliability, and ensuring that our infrastructure can support continued business growth.
This is a full-time permanent positionย
This is an existing vacancyย
ย Location:ย This is a remote location open to candidates legally authorized to work in Canada.ย ย
- Drive reliability engineering initiatives and operational excellence for mission-critical services running on AWS and Kubernetes.
- Design, implement, and continuously improve deployment, release, and rollback strategies across complex distributed systems.
- Establish secure-by-default CI/CD pipelines with robust automation, governance, and policy-driven controls.
- Enhance platform observability through metrics, logs, tracing, and actionable alerting to improve system visibility and operational efficiency.
- Define, implement, and mature Service Level Indicators (SLIs), Service Level Objectives (SLOs), and reliability standards across the organization.
- Lead response efforts for high-severity incidents, ensuring timely resolution, effective communication, and meaningful post-incident reviews that drive continuous improvement.
- Partner closely with engineering teams to strengthen platform standards, improve service resilience, optimize runtime performance, and embed reliability best practices.
- Mentor and guide engineers on cloud-native technologies, site reliability engineering principles, and operational excellence practices, fostering a culture of continuous learning and accountability.
- 8+ years of experience in Site Reliability Engineering (SRE), Platform Engineering, DevOps, or related cloud-native engineering roles.
- Deep expertise in AWS services, including EKS, IAM, VPC, Lambda, CloudFront, S3, and cloud networking/security best practices.
- Advanced experience operating and scaling production Kubernetes environments.
- Strong hands-on experience with Istio service mesh, including traffic management, security, observability, and resiliency.
- Proven expertise with Infrastructure as Code (IaC), preferably using AWS CDK.
- Experience building and managing CI/CD pipelines using GitHub Actions or similar platforms.
- Strong troubleshooting, performance optimization, and incident management experience in distributed systems.
- Excellent communication, collaboration, and technical leadership skills.
- Experience designing and operating monitoring, logging, tracing, and alerting solutions for cloud-native platforms.
- Strong knowledge of AWS CloudWatch, OpenTelemetry, AWS X-Ray, and Kubernetes observability tooling.
- Experience defining and operationalizing SLIs, SLOs, alerting strategies, runbooks, and reliability metrics.
- Proven ability to leverage observability data to improve service reliability, reduce incident impact, and optimize operational performance.
- Strong proficiency in TypeScript and Node.js for platform engineering, automation, and operational tooling.
- Experience building and maintaining scalable backend services, APIs, and event-driven systems.
- Deep understanding of Kubernetes architecture, controllers, Gateway API, ingress management, and service networking.
- Experience implementing zero-trust architectures, mTLS, and service-to-service security controls.
- Commitment to high-quality engineering practices, including automated testing, code reviews, and observability-driven development.
- Strong understanding of resilience engineering, including autoscaling, disruption management, failure testing, and safe deployment strategies.
- Experience with progressive delivery practices such as canary, blue/green, and feature-flag-based deployments.
- Experience working in regulated, compliance-driven, or security-sensitive SaaS environments.
- Familiarity with FinOps principles and cost optimization strategies for cloud platforms.
- Experience building internal developer platforms and self-service engineering tooling.
- Cloud-native certifications such as CKA, CKAD, CKS, KCSA, or KCNA.
- Kubestronaut certification or equivalent advanced Kubernetes expertise is highly regarded.
Salary Range:ย ย
The annual base salary for this position is between $140,000 CAD and $155,000 CAD per year.ย
This role is also eligible for discretionary bonus and/or commission, as well as other benefits. Actual pay within the listed range will be determined based on factors such as transferable skills, relevant experience, market conditions, and primary work location. The posted range is subject to change and may be updated periodically.