2

Remote Reliability Engineer Jobs in Ontario (NOW HIRING)

... experienced Site Reliability Engineer (SRE)/DevOps Developer to help build, operate, and ... You will report to the Manager, Software Development as part of a remote team (distributed ...

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 3+ years of DevOps/SRE experience * Strong understanding of security best practices

DevOps Engineer

Waterloo, ON ยท On-site +1

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 3+ years of DevOps/SRE experience * Strong understanding of security best practices

DevOps Engineer

Ottawa, ON ยท On-site +1

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 3+ years of DevOps/SRE experience * Strong understanding of security best practices

DevOps Engineer

Ottawa, ON ยท On-site +1

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 3+ years of DevOps/SRE experience * Strong understanding of security best practices

DevOps Engineer

Waterloo, ON ยท On-site +1

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 3+ years of DevOps/SRE experience * Strong understanding of security best practices

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 3+ years of DevOps/SRE experience * Strong understanding of security best practices

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices

This is a permanent position, that can either be remote or in-office at Toronto! Our client is a ... You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices

Remote: Waterloo ON What You'll Do: Team Leadership & Coaching * Lead and mentor a team of DevOps ... , or infrastructure engineering. * 2+ years of leadership experience (formal or informal) in a ...

Work with Development, Security, SRE and Platform teams to turn common delivery needs into reusable ... CIRA embraces a blend of remote and IRL in-office work to keep our team connected and engaged. Our ...

Remote London/Europe - we do need GMT+0 to GMT+4 of overlap with UK team. Zencargo is looking for a ... Build and scale the infrastructure behind our AI workloads, including cost, API reliability and ...

DevOps Developer

Toronto, ON ยท On-site +1

CA$125K - CA$140K/yr

This is a remote location open to candidates legally authorized to work in Canada. What you will be ... SRE, or software engineering roles with strong operational ownership. * Strong TypeScript ...

Showing results 21-40

Remote Reliability Engineer information

What is a remote reliability engineer?

A Remote Reliability Engineer is a professional who works from a remote location to ensure that systems, applications, or infrastructure are reliable, available, and performing well. Their responsibilities typically include monitoring system health, diagnosing issues, implementing preventative measures, and collaborating with teams to improve system reliability. They often use tools for automation, incident response, and performance monitoring, all while working offsite. This role is critical in minimizing downtime and ensuring a smooth user experience, especially for companies with complex technical environments. Remote Reliability Engineers must have strong problem-solving skills and be proficient in cloud technologies, automation, and incident management.

What are the key skills and qualifications needed to thrive as a remote reliability engineer?

To thrive as a Remote Reliability Engineer, you need a strong background in systems engineering, software development, and infrastructure management, often supported by a degree in computer science or a related field. Proficiency with cloud platforms (such as AWS, Azure, or GCP), monitoring tools (like Prometheus, Grafana), and relevant certifications (e.g., AWS Certified DevOps Engineer) is highly valuable. Excellent problem-solving, communication, and collaboration skills are crucial for working effectively across distributed teams and responding to incidents. These abilities ensure system reliability, quick incident resolution, and seamless remote teamwork, which are vital for maintaining high service uptime and user satisfaction.

How do remote reliability engineers typically collaborate with on-site teams to address urgent technical issues?

Remote Reliability Engineers often utilize a combination of video conferencing, instant messaging, and collaborative monitoring tools to stay closely connected with on-site teams. When urgent technical issues arise, they participate in real-time troubleshooting sessions, analyze system logs remotely, and may guide on-site staff through step-by-step resolution procedures. Building strong communication channels and regular check-ins are essential to ensure swift and effective collaboration, even across different time zones. This structure allows Remote Reliability Engineers to contribute significantly to system uptime while working from a distance.

What is the difference between Remote Reliability Engineer vs Remote Site Reliability Engineer?

AspectRemote Reliability EngineerRemote Site Reliability Engineer
CredentialsTypically requires certifications like AWS Certified Solutions Architect, Linux Foundation certificationsSimilar credentials, often with additional focus on site-specific tools and monitoring
Work EnvironmentPrimarily remote, focusing on cloud infrastructure and system reliabilityRemote with some on-site responsibilities, focusing on infrastructure and operational stability
Industry UsageUsed across tech, cloud providers, SaaS companiesCommon in data centers, cloud providers, and large enterprise IT
Search & Comparison IntentOften compared due to overlapping roles in system reliability and cloud infrastructureCompared for on-site vs remote operational responsibilities

The main difference is that Remote Reliability Engineers focus on cloud and system reliability remotely, while Remote Site Reliability Engineers may have some on-site duties related to infrastructure. Both roles require similar skills and certifications but differ in their work environment and specific responsibilities.

What are the most commonly searched types of Reliability Engineer jobs in Ontario?

The most popular types of Reliability Engineer jobs in Ontario are:

What are popular job titles related to Remote Reliability Engineer jobs in Ontario?

For Remote Reliability Engineer jobs in Ontario, the most frequently searched job titles are:

What job categories do people searching Remote Reliability Engineer jobs in Ontario look for?

The top searched job categories for Remote Reliability Engineer jobs in Ontario are:

What cities in Ontario are hiring for Remote Reliability Engineer jobs?

Cities in Ontario with the most Remote Reliability Engineer job openings:

Senior Software Engineer - Site Reliability

Funded.club

Toronto, ON โ€ข Remote

Full-time

Posted 16 days ago


Job description

Windscribe is a leading cyber security and privacy company launched in April 2016 and now with more than 100 million users. We believe that the internet was created so that people across the globe could have access to any type of information, no matter where they are. Our mission is to transform the internet with easy-to-use yet powerful privacy and security tools that allow anyone to circumvent censorship, access geographically restricted content, and minimize their exposure to marketers, criminals, and surveillance dragnets.

Our well-received applications have appeared on Lifehacker, Techradar, and CNET. Headquartered in Toronto, Canada, Windscribe operates two products:

  • Windscribe is a suite of privacy, anonymity and anti-censorship tools for the masses.
  • Control D is a modern and customizable DNS service that blocks threats, unwanted content and ads — on all devices. 

Onboard in minutes, and forget about it. We have infrastructure all over the world, and right now we are looking for a Senior Software Engineer to join our Engineering team. This is a software engineering role first: you will spend most of your time writing software that eliminates operational work, not doing operational work by hand. You do not need to arrive knowing our stack in depth. You will learn it here. 

About the Position

  • Design, write, review, test, and ship software that improves the availability, scalability, latency, and efficiency of our services. This is the majority of the job.
  • Turn operational work into engineering work: every repeated task, checklist, or validation becomes automation. A runbook is a stopgap, not a deliverable.
  • Define SLOs with stakeholders and build the monitoring and alerting that enforces them — and fix or delete every alert that doesn't map to a real expectation.
  • Debug production issues across the whole stack, from application code down through Linux and the network.
  • Lead blameless postmortems and root cause analyses, then land the fix that eliminates the entire class of problem, not just the writeup.
  • Plan capacity and cost with data; document what you build so others can operate it.
  • Join our critical incident response team in an established on-call rotation. To be clear about what that means here: the rotation covers our shared infrastructure, not only the code you wrote. Learn our stack in depth — DNS at global scale, anycast networks, VPN protocols — with the team that runs it. 

Qualifications & Experience 

We hire for the software engineer first. Most of our stack can be learned on the job; the following cannot:

  • Bachelor's degree in Software Engineering or Computer Science or similar.
  • 5+ years of professional software development in one or more general-purpose languages. Go preferred; Python, Rust, and C/C++ are also great. Shell and YAML alone do not qualify.
  •  Automation you personally built — a tool, service, operator, distributed system, or auto-remediation system that eliminated real operational work. Expect us to ask you to walk through its design and trade-offs. 
  • Experience designing, analyzing, and troubleshooting distributed systems in production. Git, testing, code review, and CI as your normal workflow; comfortable in a large codebase you didn't write.
  • Solid Linux fundamentals, containers, and infrastructure as code (Terraform, Ansible, or equivalents). 
  • Observability as a builder: you instrument your own code (Prometheus, Grafana, or equivalents), not just read dashboards.

Who This Role Is Not For 

We want to respect your time, so we will be upfront: 

  • If your experience is primarily operating, configuring, and troubleshooting software that other people built — systems administration, NOC, or infrastructure operations under an SRE title — this role will not be a fit.
  • If your observability and reliability experience is primarily advisory — SLO frameworks, dashboards, and monitoring strategy, without software you built to act on those signals — this role will not be a fit either.

Bonus If You Have

None of these are required. You will learn them here:

  • Deep DNS knowledge (BIND / PowerDNS / Unbound, record types, resolution paths, DNSSEC)
  • Routing experience: unicast, anycast, BGP Linux networking depth: iptables / nftables, eBPF, performance tuning VPN protocols (OpenVPN / WireGuard / IKEv2)
  • Experience with bare metal fleets, hypervisors, or networking hardware (JunOS / VyOS)
  • High-availability databases (MySQL, Postgres, Redis) and load balancers (HAProxy, nginx) 

Compensation:

140K - 180K CAD

Stock options

#LI-JM1

#LI-Remote

Thank you for considering this opportunity.  Funded.club Senior Recruiters partner exclusively with Startups and are in direct communication with hiring managers and founding team members.

Funded.club uses AI-assisted tools as part of our candidate sourcing and screening process. All applications are reviewed by a human recruiter, who makes all decisions about which candidates to progress. If your application seems like a good fit for the position, a real member of our team will contact you soon!