Avez 5 ans ou plus d'experience en infrastructure, DevOps ou ingenierie de la fiabilite des sites ... Site Reliability Engineer (SRE) About Us Toboggan Labs is a boutique consultancy building at the ...
Avez 5 ans ou plus d'experience en infrastructure, DevOps ou ingenierie de la fiabilite des sites ... Site Reliability Engineer (SRE) About Us Toboggan Labs is a boutique consultancy building at the ...
Senior DevOps Engineer
Montreal, QC · On-site +1
You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices * Experience with: * CI/CD * Kubernetes/Docker * Cloud infrastructure (AWS preferably) * Infrastructure ...
Quick apply
Senior DevOps Engineer
Montreal, QC · On-site +1
You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices * Experience with: * CI/CD * Kubernetes/Docker * Cloud infrastructure (AWS preferably) * Infrastructure ...
Senior DevOps Engineer
Montreal, QC · On-site +1
You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices * Experience with: * CI/CD * Kubernetes/Docker * Cloud infrastructure (AWS preferably) * Infrastructure ...
Quick apply
Senior DevOps Engineer
Montreal, QC · On-site +1
You Have: * 5+ years of DevOps/SRE experience * Strong understanding of security best practices * Experience with: * CI/CD * Kubernetes/Docker * Cloud infrastructure (AWS preferably) * Infrastructure ...
We are now looking for an SRE engineer to join our team in Montreal and work on an exciting project ... Have a passion for maintaining operational stability. YOU HAVE : * Extensive Troubleshooting Skills ...
We are now looking for an SRE engineer to join our team in Montreal and work on an exciting project ... Have a passion for maintaining operational stability. YOU HAVE : * Extensive Troubleshooting Skills ...
Spécialiste DevOps Département: Ingénierie Logicielle (CitizenOne) Rapports à : Directeur ... (SRE). Outils de surveillance et de journalisation tels qu'Azure Monitor, Application Insights ...
Quick apply
Spécialiste DevOps Département: Ingénierie Logicielle (CitizenOne) Rapports à : Directeur ... (SRE). Outils de surveillance et de journalisation tels qu'Azure Monitor, Application Insights ...
Salary: $75,000 - $90,000 Spcialiste DevOps Dpartement: Ingnierie Logicielle (CitizenOne) Rapports ... (SRE). * Outils de surveillance et de journalisation tels quAzure Monitor, Application Insights ...
Quick apply
Salary: $75,000 - $90,000 Spcialiste DevOps Dpartement: Ingnierie Logicielle (CitizenOne) Rapports ... (SRE). * Outils de surveillance et de journalisation tels quAzure Monitor, Application Insights ...
Specialiste DevOps Departement: Ingenierie Logicielle (CitizenOne) Rapports a : Directeur ... (SRE). * Outils de surveillance et de journalisation tels qu'Azure Monitor, Application Insights ...
Specialiste DevOps Departement: Ingenierie Logicielle (CitizenOne) Rapports a : Directeur ... (SRE). * Outils de surveillance et de journalisation tels qu'Azure Monitor, Application Insights ...
SRE Associate (Hybrid)
Montreal, QC · Hybrid
We're seeking someone to join our team as a Site Reliability Engineering Specialist to support the ... Work independently to solve ambiguous operational and technical problems while supporting business ...
SRE Associate (Hybrid)
Montreal, QC · Hybrid
We're seeking someone to join our team as a Site Reliability Engineering Specialist to support the ... Work independently to solve ambiguous operational and technical problems while supporting business ...
About the role The SRE & Operational Intelligence team designs and implements practices and platforms that improve the reliability and performance of our on-premises and cloud applications, services ...
About the role The SRE & Operational Intelligence team designs and implements practices and platforms that improve the reliability and performance of our on-premises and cloud applications, services ...
About the role The SRE & Operational Intelligence team designs and implements practices and platforms that improve the reliability and performance of our on-premises and cloud applications, services ...
About the role The SRE & Operational Intelligence team designs and implements practices and platforms that improve the reliability and performance of our on-premises and cloud applications, services ...
About the role The SRE & Operational Intelligence team designs and implements practices and platforms that improve the reliability and performance of our on-premises and cloud applications, services ...
About the role The SRE & Operational Intelligence team designs and implements practices and platforms that improve the reliability and performance of our on-premises and cloud applications, services ...
About the role The SRE & Operational Intelligence team designs and implements practices and platforms that improve the reliability and performance of our on-premises and cloud applications, services ...
About the role The SRE & Operational Intelligence team designs and implements practices and platforms that improve the reliability and performance of our on-premises and cloud applications, services ...
???? DevOps & Infrastructure Developer | Multiplayer FPS ~ ???? Développeur DevOps et infrastruc
Montreal, QC · On-site
We need a DevOps & Infrastructure Developer to own the infrastructure underneath: the Kubernetes ... SRE * Production Kubernetes + Helm * AWS: EKS, VPC, EC2, CloudFront * Terraform (IaC, no ClickOps)
Quick apply
???? DevOps & Infrastructure Developer | Multiplayer FPS ~ ???? Développeur DevOps et infrastruc
Montreal, QC · On-site
We need a DevOps & Infrastructure Developer to own the infrastructure underneath: the Kubernetes ... SRE * Production Kubernetes + Helm * AWS: EKS, VPC, EC2, CloudFront * Terraform (IaC, no ClickOps)
Ingenieur(e) en soutien de production | Programme diplomes Cloud, DevOps et SRE / Production Supp...
Montreal, QC · On-site
... DevOps ou l'ingenierie de fiabilite des sites (SRE). * Des connaissances de base en Linux ou Unix et le desir de developper de solides competences en resolution de problemes. * Une experience avec ...
Ingenieur(e) en soutien de production | Programme diplomes Cloud, DevOps et SRE / Production Supp...
Montreal, QC · On-site
... DevOps ou l'ingenierie de fiabilite des sites (SRE). * Des connaissances de base en Linux ou Unix et le desir de developper de solides competences en resolution de problemes. * Une experience avec ...
We are growing SRE capabilities within our Reliability & Production Engineering (RPE) organization ... Represent the RPE organization in design reviews and operational readiness exercises for new and ...
We are growing SRE capabilities within our Reliability & Production Engineering (RPE) organization ... Represent the RPE organization in design reviews and operational readiness exercises for new and ...
Site Reliability Specialist
Montreal, QC · On-site
As a Site Reliability Specialist at Ubisoft Montréal, you will join the IT Games and Studios team ... Experience with infrastructure engineering, automation, and DevOps practices * Knowledge of GitLab ...
Quick apply
Site Reliability Specialist
Montreal, QC · On-site
As a Site Reliability Specialist at Ubisoft Montréal, you will join the IT Games and Studios team ... Experience with infrastructure engineering, automation, and DevOps practices * Knowledge of GitLab ...
... s/SRE practices to improve quality of work, ease of development/deployment * Relevant experience with payment technologies, fintech, or banking * Experience driving reliability initiatives ...
... s/SRE practices to improve quality of work, ease of development/deployment * Relevant experience with payment technologies, fintech, or banking * Experience driving reliability initiatives ...
Support teams in implementing and evolving SRE practices based on their level of maturity ... operational efficiency. We aim to offer you maximum flexibility to support your quality of life ...
Support teams in implementing and evolving SRE practices based on their level of maturity ... operational efficiency. We aim to offer you maximum flexibility to support your quality of life ...
Support teams in implementing and evolving SRE practices based on their level of maturity ... operational efficiency. We aim to offer you maximum flexibility to support your quality of life ...
Support teams in implementing and evolving SRE practices based on their level of maturity ... operational efficiency. We aim to offer you maximum flexibility to support your quality of life ...
... and operational maturity. This role allows you to have a concrete and lasting impact on the ... SRE practices based on their level of maturity Actively support major incident management by ...
... and operational maturity. This role allows you to have a concrete and lasting impact on the ... SRE practices based on their level of maturity Actively support major incident management by ...
Sre Devops information
What is the difference between Sre Devops vs Devops Engineer?
| Aspect | Sre Devops | Devops Engineer |
|---|---|---|
| Primary Focus | Reliability, automation, and system stability | Deployment, CI/CD, and infrastructure automation |
| Credentials | Certifications like AWS, Linux, or Kubernetes | Similar certifications, often including cloud and scripting skills |
| Work Environment | Operations teams, site reliability, and incident management | Development teams, deployment pipelines, and infrastructure setup |
| Industry Usage | Tech companies, cloud providers, large-scale systems | Startups, software companies, and enterprises adopting DevOps practices |
While both roles involve automation and cloud skills, Sre Devops emphasizes system reliability and incident response, whereas Devops Engineer focuses more on deployment pipelines and infrastructure automation. Understanding these differences helps organizations assign the right responsibilities and professionals for their needs.

Full-time
Medical, Dental, Life, Retirement
Re-posted 22 hours ago
Job description
La version anglaise suivra - English version will follow
Ingenieure en fiabilite des sites
A propos de nous
Toboggan Labs est une firme-conseil boutique qui uvre a l'intersection de l'IA et de la sante. Nous resolvons des problemes humains complexes en appliquant des technologies de pointe combinees a une solide comprehension du domaine.
A propos du poste
Nous sommes a la recherche d'une ingenieure en fiabilite des sites (SRE) pour aider nos clients a concevoir et exploiter des systemes de production fiables, observables et securises.
Dans ce role, vous travaillerez aux cotes des equipes d'ingenierie et d'exploitation de nos clients pour ameliorer la fiabilite des systemes, reduire les taches manuelles repetitives et batir les fondations operationnelles - pipelines de deploiement, surveillance, gestion des incidents, infrastructure - qui maintiennent les systemes en production en bonne sante.
Veuillez noter que, bien que nous soyons specialises dans le secteur de la sante et les industries reglementees, tous nos projets ne relevent pas de ces domaines. Vous pourriez donc etre amenee a travailler sur des projets varies dans differents secteurs, selon les besoins.
Vos responsabilites quotidiennes
- Ingenierie d'infrastructure et de securite - Concevoir et maintenir des infrastructures infonuagiques resilientes et securisees a l'aide d'outils d'infrastructure-as-code; implanter des controles de securite, des standards de durcissement et des guardrails de conformite dans les environnements clients.
- Observabilite et fiabilite - Concevoir et implanter des systemes de surveillance, d'alertes et de journalisation; piloter les processus de reponse aux incidents et de post-mortems; definir et suivre les SLO/SLI.
- Automatisation et operations de plateforme - Automatiser les pipelines de deploiement, le provisionnement d'infrastructure et les runbooks operationnels pour reduire les taches manuelles et ameliorer la resilience des systemes.
- Sur certains mandats, leadership technique - Piloter le volet fiabilite et infrastructure, guider les equipes clients sur les pratiques SRE et contribuer aux decisions architecturales.
- Soutien a l'equipe - Partager votre expertise en fiabilite et operations, contribuer aux outils et a la documentation internes, mentorer des collegues et participer a la communaute Toboggan.
A propos de vous
Nous recherchons des personnes ayant un solide historique en infrastructure infonuagique, DevOps et ingenierie de la fiabilite, qui abordent tout ce qu'elles construisent avec un souci de la securite. La majorite de nos clients utilisent AWS, Terraform, GitHub Actions ou des outils CI/CD similaires. Vous devez etre a l'aise pour travailler a la frontiere entre la fiabilite, la securite et les operations TI.
Quand nous parlons de fiabilite des systemes, nous cherchons quelqu'un qui traite la fiabilite comme une fonctionnalite a part entiere, et non comme une reflexion apres coup - quelqu'un qui ecrit du code pour eliminer les taches repetitives, pense en termes de systemes et construit des infrastructures aussi securisees qu'observables.
Nous vous encourageons a postuler si vous :
- Avez 5 ans ou plus d'experience en infrastructure, DevOps ou ingenierie de la fiabilite des sites;
- Avez une experience pratique avec des infrastructures AWS ou Azure et des outils d'infrastructure-as-code (Terraform, CloudFormation ou equivalents);
- Avez une solide experience avec les pipelines CI/CD (GitHub Actions, ArgoCD, Jenkins ou equivalents) et l'automatisation des deploiements;
- Avez de l'experience avec des outils d'observabilite (Prometheus, Grafana, Datadog, CloudWatch ou equivalents) et les processus de gestion des incidents;
- Etes familierere avec les bonnes pratiques de securite pour l'infrastructure infonuagique, incluant la securite reseau, l'IAM, le chiffrement et la gestion des vulnerabilites;
- Possedez d'excellentes competences en communication et etes capable d'expliquer des concepts d'infrastructure et de fiabilite a des parties prenantes variees;
- Etes adaptable, autonome et a l'aise dans des environnements clients dynamiques;
- Savez expliquer les compromis entre fiabilite et securite et les relier aux besoins d'affaires.
Atouts supplementaires
- Experience dans des roles orientes client (consultation, ingenierie d'implantation, services-conseils);
- Experience dans le secteur de la sante ou d'autres industries fortement reglementees;
- Experience en developpement logiciel au-dela du simple scripting (developpement de fonctionnalites, d'API ou d'applications);
- Experience avec l'orchestration de conteneurs (Kubernetes, ECS) et les outils de securite cloud-native;
- Experience en automatisation d'infrastructure a l'aide de scripts (Python, Bash) ou d'outils de workflow;
- Detention de certifications pertinentes (AWS DevOps Professional, AWS Solutions Architect, CKA ou equivalentes).
Toutes nos offres d'emploi decrivent un peu une licorne. Si vous etes plutot un narval , postulez quand meme ! Il n'est pas necessaire de repondre a toutes les exigences, ni aux criteres bonus. L'experience et les competences sont importantes, mais le potentiel de croissance et l'attitude le sont tout autant. Nous sommes generalement flexibles quant aux niveaux ou pouvons vous orienter vers une offre plus appropriee lorsqu'elle sera ouverte.
Ce que nous offrons
Nous sommes une entreprise en teletravail d'abord, avec un espace de bureau a Montreal. Nous privilegions l'embauche au Quebec, mais sommes ouverts aux candidatures partout au Canada dans les fuseaux horaires EST 2.
Toboggan Labs valorise la diversite des personnes qu'elle embauche et qu'elle sert. Pour nous, la diversite signifie creer un milieu de travail ou les differences de chacune sont reconnues, appreciees, respectees et prises en compte afin de developper et de mettre a profit les talents et les forces de chaque personne.
En plus :
- Budget pour le bureau a domicile et la technologie;
- Budget annuel de developpement professionnel;
- REER avec contribution de l'employeur apres 1 an;
- Des le premier jour :
- Assurance sante et dentaire payee a 100 % par l'employeur, incluant un montant annuel pour les soins complementaires (acupuncture, osteopathie, massotherapie, naturopathie, psychologie, etc.);
- Assurance vie et assurance invalidite de courte et de longue duree;
- Complement de conge parental (8 semaines), disponible pour les employes ayant plus d'un an d'anciennete, quel que soit le chemin vers la parentalite.
Site Reliability Engineer (SRE)
About Us
Toboggan Labs is a boutique consultancy building at the intersection of AI and healthcare. We solve challenging human problems by applying cutting-edge technology and domain understanding.
About the role
We're seeking a Site Reliability Engineer (SRE) to help our clients build reliable, observable, and secure production systems.
In this role, you will work closely with client engineering and operations teams to improve system reliability, reduce toil, and build the operational foundations - deployment pipelines, monitoring, incident management, and infrastructure - that keep production systems running smoothly.
Note that while we specialize in healthcare and regulated industries, not all our projects are in these fields, so you may work across different domains from time to time.
Your work will consist of:
- Infrastructure and security engineering - Design and maintain resilient, secure cloud infrastructure using infrastructure-as-code; implement security controls, hardening standards, and compliance guardrails across client environments.
- Observability and reliability - Design and implement monitoring, alerting, and logging systems; lead incident response and post-mortem processes; define and track SLOs and SLIs.
- Automation and platform operations - Automate deployment pipelines, infrastructure provisioning, and operational runbooks to reduce toil and improve system resilience.
- On some projects, technical leadership - Own the reliability and infrastructure workstream, guide client engineering teams on SRE practices, and contribute to architectural decisions.
- Supporting the team - Share SRE expertise with colleagues, contribute to internal tooling and documentation, mentor team members, and participate in the broader Toboggan community.
About you
We are seeking individuals with a strong background in cloud infrastructure, DevOps, and reliability engineering, who bring security mindedness to everything they build. Most of our clients run AWS, Terraform, GitHub Actions or similar CI/CD tooling. You should be comfortable working at the intersection of reliability, security, and IT operations.
When we say SRE we mean someone who treats reliability as a feature, not an afterthought - someone who writes code to eliminate toil, thinks in systems, and builds infrastructure that is as secure as it is observable.
We want you to apply if you:
- Have 5+ years of experience in infrastructure, DevOps, or site reliability engineering;
- Have hands-on experience with AWS or Azure infrastructure and infrastructure-as-code tools (Terraform, CloudFormation, or equivalents);
- Have strong experience with CI/CD pipelines (GitHub Actions, ArgoCD, Jenkins, or equivalents) and deployment automation;
- Have experience with observability tools (Prometheus, Grafana, Datadog, CloudWatch, or equivalents) and incident management processes;
- Are familiar with security best practices for cloud infrastructure, including network security, IAM, encryption, and vulnerability management;
- Have excellent communication skills and can explain infrastructure and reliability concepts to varied stakeholders;
- Are adaptable, self-directed, and comfortable in dynamic client environments;
- Can explain reliability and security trade-offs and connect them to business needs.
Bonus points if you:
- Have experience in client-facing roles such as consulting, implementation engineering, or advisory work.
- Have worked in healthcare or other heavily regulated industries.
- Have software development experience beyond scripting - experience building features, APIs, or applications.
- Have experience with container orchestration (Kubernetes, ECS) and cloud-native tooling.
- Have built infrastructure automation using scripting (Python, Bash) or workflow tools.
- Hold relevant certifications (AWS DevOps Professional, AWS Solutions Architect, CKA, or similar).
All of our job postings describe a bit of a unicorn. If you're kind of a "narwhal," please apply anyway. You don't need to meet all the requirements, let alone the bonus criteria. While experience and skill sets are valuable, growth potential and attitudes are equally important. We are usually flexible on levels or can advise you when a more relevant posting opens.
What we offer
We are a remote-first company with office space in Montreal. We prefer to hire in Quebec, but we are open to candidates anywhere in the EST2 time zone in Canada.
Toboggan Labs values the diversity of the people it hires and serves. Diversity, for us, means fostering a workplace in which a person's differences are recognized, appreciated, respected and responded to in ways that fully develop and utilize their talents and strengths.
In addition:
- Home office/technology budget;
- Yearly professional development budget;
- Company matching RRSP after 1 year;
- From Day 1:
- 100% employer-paid health & dental insurance including a yearly bank of coverage for complementary medicine (Acupuncture, osteopathy, massage therapy, naturopathy, psychology, etc.);
- Life, long & short-term disability insurance;
- Parental leave top-up (8 weeks), available to employees with 1+ year of tenure, regardless of path to parenthood.