Experience instrumenting application services end point for logging, metrics and events using Azure monitor/Datadog/Prometheus/Grafana. Proven ability to analyze, debug, and troubleshoot large-scale ...
Experience instrumenting application services end point for logging, metrics and events using Azure monitor/Datadog/Prometheus/Grafana. Proven ability to analyze, debug, and troubleshoot large-scale ...
Responsable principal, Exploitation d'entreprise et services d'integration /Lead, Enterprise Oper...
Experience avec des outils de surveillance et d'observabilite tels qu'Azure Monitor, Splunk, Datadog ou Dynatrace. * Solide connaissance des pratiques ITIL, de la gestion des incidents et de la ...
Responsable principal, Exploitation d'entreprise et services d'integration /Lead, Enterprise Oper...
Experience avec des outils de surveillance et d'observabilite tels qu'Azure Monitor, Splunk, Datadog ou Dynatrace. * Solide connaissance des pratiques ITIL, de la gestion des incidents et de la ...
Maîtriser les outils tels que Github, EKS, Hashicorp Vault, TFC, Datadog et Splunk * Connaître les meilleures pratiques de développement logiciel * Posséder une certification AWS * Expérience en ...
New
Maîtriser les outils tels que Github, EKS, Hashicorp Vault, TFC, Datadog et Splunk * Connaître les meilleures pratiques de développement logiciel * Posséder une certification AWS * Expérience en ...
New
Datadog, Prometheus). * Collaborer avec des equipes interfonctionnelles, notamment des chefs de produit, les concepteurs et d'autres ingenieurs, afin de comprendre les besoins et de fournir des ...
Datadog, Prometheus). * Collaborer avec des equipes interfonctionnelles, notamment des chefs de produit, les concepteurs et d'autres ingenieurs, afin de comprendre les besoins et de fournir des ...
Développeur senior DevOps
Montreal, QC · On-site
Maîtriser les outils tels que Github, EKS, Hashicorp Vault, TFC, Datadog et Splunk * Connaître les meilleures pratiques de développement logiciel * Posséder une certification AWS * Expérience en ...
Développeur senior DevOps
Montreal, QC · On-site
Maîtriser les outils tels que Github, EKS, Hashicorp Vault, TFC, Datadog et Splunk * Connaître les meilleures pratiques de développement logiciel * Posséder une certification AWS * Expérience en ...
Senior DevOps Programmer - Unannounced project | Programmeureuse DevOps seniore - Projet non-annonce
Montreal, QC · On-site
Experience with monitoring and observability tools such as Datadog, Prometheus, or similar, to ensure system reliability and performance. * Autonomous and solution-oriented mindset and ability to ...
Senior DevOps Programmer - Unannounced project | Programmeureuse DevOps seniore - Projet non-annonce
Montreal, QC · On-site
Experience with monitoring and observability tools such as Datadog, Prometheus, or similar, to ensure system reliability and performance. * Autonomous and solution-oriented mindset and ability to ...
Practical experience with observability and monitoring platforms such as Prometheus, Grafana, or Datadog. * Strong understanding of networking, security, and distributed systems in cloud-native ...
Practical experience with observability and monitoring platforms such as Prometheus, Grafana, or Datadog. * Strong understanding of networking, security, and distributed systems in cloud-native ...
... de Datadog; definir les SLO/SLI et creer des tableaux de bord exploitables qui generent des resultats de fiabilite. * Developper et favoriser l'automatisation : ameliorer les outils internes, les ...
... de Datadog; definir les SLO/SLI et creer des tableaux de bord exploitables qui generent des resultats de fiabilite. * Developper et favoriser l'automatisation : ameliorer les outils internes, les ...
AWS / Google Cloud, Kubernetes, Docker, Terraform ), , including hands-on IaC, CI/CD pipeline configuration (GitLab CI, GitHub Actions), and cloud monitoring/observability (CloudWatch, Datadog ...
AWS / Google Cloud, Kubernetes, Docker, Terraform ), , including hands-on IaC, CI/CD pipeline configuration (GitLab CI, GitHub Actions), and cloud monitoring/observability (CloudWatch, Datadog ...
Hands-on with observability stacks (CloudWatch, Datadog, Grafana, or equivalents) * Demonstrated experience standing up SRE practices: SLOs, on-call, incident management, blameless postmortems
Hands-on with observability stacks (CloudWatch, Datadog, Grafana, or equivalents) * Demonstrated experience standing up SRE practices: SLOs, on-call, incident management, blameless postmortems
Développeur principal Python
Montreal, QC · On-site
Connaissance d'outils de suivi et de monitoring (JIRA, Datadog, Splunk, Okta) Tes avantages En plus d'une rémunération concurrentielle, nous te proposons, dès ton embauche, une foule d'avantages ...
Développeur principal Python
Montreal, QC · On-site
Connaissance d'outils de suivi et de monitoring (JIRA, Datadog, Splunk, Okta) Tes avantages En plus d'une rémunération concurrentielle, nous te proposons, dès ton embauche, une foule d'avantages ...
Avez de l'experience avec des outils d'observabilite (Prometheus, Grafana, Datadog, CloudWatch ou equivalents) et les processus de gestion des incidents; * Etes familierere avec les bonnes pratiques ...
Avez de l'experience avec des outils d'observabilite (Prometheus, Grafana, Datadog, CloudWatch ou equivalents) et les processus de gestion des incidents; * Etes familierere avec les bonnes pratiques ...
Experience with telemetry and observability tools such as Datadog or Dynatrace * Experience with GraphQL and RESTful APIs * Experience with PostgreSQL and relational databases * Experience designing ...
Experience with telemetry and observability tools such as Datadog or Dynatrace * Experience with GraphQL and RESTful APIs * Experience with PostgreSQL and relational databases * Experience designing ...
Développeur senior DevOps
Laval, QC · On-site
Maîtriser les outils tels que Github, EKS, Hashicorp Vault, TFC, Datadog et Splunk * Connaître les meilleures pratiques de développement logiciel * Posséder une certification AWS * Expérience en ...
New
Développeur senior DevOps
Laval, QC · On-site
Maîtriser les outils tels que Github, EKS, Hashicorp Vault, TFC, Datadog et Splunk * Connaître les meilleures pratiques de développement logiciel * Posséder une certification AWS * Expérience en ...
New
Développeur principal Python
Laval, QC · On-site
Connaissance d'outils de suivi et de monitoring (JIRA, Datadog, Splunk, Okta) Tes avantages En plus d'une rémunération concurrentielle, nous te proposons, dès ton embauche, une foule d'avantages ...
New
Développeur principal Python
Laval, QC · On-site
Connaissance d'outils de suivi et de monitoring (JIRA, Datadog, Splunk, Okta) Tes avantages En plus d'une rémunération concurrentielle, nous te proposons, dès ton embauche, une foule d'avantages ...
New
Connaissance d'outils de suivi et de monitoring (JIRA, Datadog, Splunk, Okta) Tes avantages En plus d'une rémunération concurrentielle, nous te proposons, dès ton embauche, une foule d'avantages ...
New
Connaissance d'outils de suivi et de monitoring (JIRA, Datadog, Splunk, Okta) Tes avantages En plus d'une rémunération concurrentielle, nous te proposons, dès ton embauche, une foule d'avantages ...
New
Practical experience with observability and monitoring platforms such as Prometheus, Grafana, or Datadog. * Strong understanding of networking, security, and distributed systems in cloud-native ...
Practical experience with observability and monitoring platforms such as Prometheus, Grafana, or Datadog. * Strong understanding of networking, security, and distributed systems in cloud-native ...
... Datadog et Splunk Connaître les meilleures pratiques de développement logiciel Posséder une certification AWS Expérience en intégration et déploiement continu Démontrer des aptitudes en ...
... Datadog et Splunk Connaître les meilleures pratiques de développement logiciel Posséder une certification AWS Expérience en intégration et déploiement continu Démontrer des aptitudes en ...
Experience pratique des outils de surveillance (Datadog, Splunk, Prometheus, Grafana) et interet pour le developpement assiste par l'IA. Bilingue (francais et anglais), tant a l'oral qu'a l'ecrit. Ce ...
Experience pratique des outils de surveillance (Datadog, Splunk, Prometheus, Grafana) et interet pour le developpement assiste par l'IA. Bilingue (francais et anglais), tant a l'oral qu'a l'ecrit. Ce ...
DevOps Engineer
Montreal, QC · On-site
DataDog (APM, Logs, Metrics, Dashboards, Monitors, Service Management) * MS SQL Server 2019+ * .NET/C#, PowerShell, bash, Windows command shell * Git What We Offer * Competitive salary based on your ...
Quick apply
DevOps Engineer
Montreal, QC · On-site
DataDog (APM, Logs, Metrics, Dashboards, Monitors, Service Management) * MS SQL Server 2019+ * .NET/C#, PowerShell, bash, Windows command shell * Git What We Offer * Competitive salary based on your ...
Datadog information
What is the difference between Datadog vs Cloud Monitoring Engineer?
| Aspect | Datadog | Cloud Monitoring Engineer |
|---|---|---|
| Primary Role | Monitoring and analytics platform for IT infrastructure and applications | Designing, implementing, and managing cloud monitoring solutions |
| Required Skills | Cloud platforms, monitoring tools, scripting, API integration | Cloud services, monitoring tools, scripting, troubleshooting |
| Certifications | Cloud certifications (AWS, Azure), monitoring tools certifications | Cloud certifications (AWS, Azure), monitoring certifications |
| Work Environment | Using SaaS platform, integrating with various cloud and on-premise systems | Managing cloud infrastructure, configuring monitoring tools, troubleshooting |
While both roles involve cloud monitoring, Datadog focuses on utilizing a specific SaaS platform for analytics and monitoring, whereas a Cloud Monitoring Engineer designs and manages monitoring solutions across cloud environments. The roles often overlap in skills and certifications, but their core responsibilities differ in scope and focus.
What are the key skills and qualifications needed to thrive as a Datadog engineer?
What are some common challenges faced by Datadog engineers when implementing monitoring solutions for large-scale systems?
What is a Datadog engineer?

Full-time
Posted 28 days ago
Morgan Stanley rating
8.4
Based on 155 frontline employees who took The Breakroom Quiz
30th of 150 rated financial services
Job description
We're seeking someone to join our team as a Senior Azure & On-Prem Production Support in Reliability & Production Engineering to ensure stability, resilience, and continuous evolution of critical Settlements systems across hybrid cloud and on-premise environments.
In the Technology division, we leverage innovation to build the connections and capabilities that power our Firm, enabling our clients and colleagues to redefine markets and shape the future of our communities. This is a Lead Software Prod Management & Reliability Engineering position at Director level, which is part of the job family responsible for overseeing the production environment, ensuring the operational reliability of deployed software, and implementing strategies to optimize performance and minimize downtime.
Since 1935, Morgan Stanley is known as a global leader in financial services, always evolving and innovating to better serve our clients and our communities in more than 40 countries around the world.
What you'll do in the role:
Be responsible for the stability and resilience of critical Settlements flows across cloud-based and distributed platforms.
Manage day-to-day production support of application and corresponding infra, handling application and infra alerts and supporting application users, while contributing to technology transformation initiatives, including cloud and observability enhancements
Lead incident management processes, ensuring clear and concise communication with senior stakeholders
Implement observability and telemetry frameworks for application and associated infrastructure to support SLIs, SLOs, and deep operational insights
Review performance, scalability, observability, resiliency, and reliability of systems through collaboration with development and infrastructure teams
Drive automation initiatives, solve complex technical and business challenges using domain expertise across large-scale distributed systems
Lead proof-of-concepts for new technologies and business flows and support their onboarding into production environments
What you'll bring to the role:
Bachelor's degree in Computer Science or a related field
6+ years of experience in IT, including at least 5 years in Production Management across multiple applications and environments
Experience supporting Java-based microservices in both cloud (Azure) and on-premise environments
Technical skills: Terraform or GitHub, Azure Kubernetes or Azure spring apps, Azure Web apps, AzureSQL or MongoDB, Kafka or Azure ServiceBus, unix, python or any other scripting language like bash.
Experience instrumenting application services end point for logging, metrics and events using Azure monitor/Datadog/Prometheus/Grafana.
Proven ability to analyze, debug, and troubleshoot large-scale distributed systems across application, infrastructure, and database layers
Experience working with Python and Java codebases
Strong collaboration and communication skills across technical and business stakeholders
All our positions are located in Montreal, Quebec. We offer a hybrid work environment, combining remote work and attendance in the office.
Knowledge of French and English is required.
WHAT YOU CAN EXPECT FROM MORGAN STANLEY:
At Morgan Stanley, we raise, manage and allocate capital for our clients - helping them reach their goals. We do it in a way that's differentiated - and we've done that for 90 years. Our values - putting clients first, doing the right thing, leading with exceptional ideas, committing to diversity and inclusion, and giving back - aren't just beliefs, they guide the decisions we make every day to do what's best for our clients, communities and more than 80,000 employees in 1,200 offices across 42 countries. At Morgan Stanley, you'll find an opportunity to work alongside the best and the brightest, in an environment where you are supported and empowered. Our teams are relentless collaborators and creative thinkers, fueled by their diverse backgrounds and experiences. We are proud to support our employees and their families at every point along their work-life journey, offering some of the most attractive and comprehensive employee benefits and perks in the industry. There's also ample opportunity to move about the business for those who show passion and grit in their work.
To learn more about our offices across the globe, please copy and paste https://www.morganstanley.com/about-us/global-offices into your browser.
Morgan Stanley is an equal opportunity employer committed to building and maintaining a workforce that is diverse in experience and background. Our recruiting efforts reflect our strong commitment to a culture of inclusion, where individuals are hired, developed, and advanced based on their skills and talents.
Our workforce reflects a broad cross-section of the global communities in which we operate, bringing a variety of backgrounds, talents, perspectives, and experiences.
For more information, please visit: https://www.morganstanley.com/people-opportunities/eeo.
What Morgan Stanley employees say
Pay
Benefits
Hours and flexibility
Workplace
Get the full story on Breakroom