Perform second-level support using visibility tools such as Datadog and Splunk and expert hands on experience with one or more (Datadog, Splunk, Grafana) * Understanding of ITIL processes (Incident ...
Perform second-level support using visibility tools such as Datadog and Splunk and expert hands on experience with one or more (Datadog, Splunk, Grafana) * Understanding of ITIL processes (Incident ...
DevOps Engineer
Austin, TX · On-site
$99K/yr
Monitor performance using tools such as CloudWatch and Datadog * Troubleshoot issues and identify opportunities for proactive improvement * Partner with engineering teams to resolve environment and ...
DevOps Engineer
Austin, TX · On-site
$99K/yr
Monitor performance using tools such as CloudWatch and Datadog * Troubleshoot issues and identify opportunities for proactive improvement * Partner with engineering teams to resolve environment and ...
AWS DevOps Engineer
Plano, TX · On-site
$50.50 - $69.25/hr
Observability - Datadog, Splunk * ALM/Documentation - JIRA, Confluence * Containerizations - Kubernetes, EKS
Quick apply
AWS DevOps Engineer
Plano, TX · On-site
$50.50 - $69.25/hr
Observability - Datadog, Splunk * ALM/Documentation - JIRA, Confluence * Containerizations - Kubernetes, EKS
Architect - Java
San Antonio, TX · On-site
$85K - $115K/yr
Experience in monitoring tools like Datadog, ELK, Grafana and Splunk * Good to have Property & Casualty Insurance domain experience * Should be able to draft architecture and prepare presentation ...
Architect - Java
San Antonio, TX · On-site
$85K - $115K/yr
Experience in monitoring tools like Datadog, ELK, Grafana and Splunk * Good to have Property & Casualty Insurance domain experience * Should be able to draft architecture and prepare presentation ...
Full Stack Engineer
Roanoke, TX · On-site
... Datadog • Strong communication skills with technical and non-technical teammates Qualifications : Required : • Bachelor's / Master's degree or equivalent in Computer Science or Engineering • ...
Full Stack Engineer
Roanoke, TX · On-site
... Datadog • Strong communication skills with technical and non-technical teammates Qualifications : Required : • Bachelor's / Master's degree or equivalent in Computer Science or Engineering • ...
Site Reliability Engineer
$55.25 - $73.50/hr
Observability & Insights - Set up and maintain full observability stacks (logging, metrics, tracing) using tools like Prometheus, Grafana, Datadog, OpenTelemetry, or ELK. - Analyze telemetry and logs ...
Site Reliability Engineer
$55.25 - $73.50/hr
Observability & Insights - Set up and maintain full observability stacks (logging, metrics, tracing) using tools like Prometheus, Grafana, Datadog, OpenTelemetry, or ELK. - Analyze telemetry and logs ...
Site Reliability Engineer
Houston, TX · On-site
$55.25 - $73.50/hr
Observability & Insights - Set up and maintain full observability stacks (logging, metrics, tracing) using tools like Prometheus, Grafana, Datadog, OpenTelemetry, or ELK. - Analyze telemetry and logs ...
Site Reliability Engineer
Houston, TX · On-site
$55.25 - $73.50/hr
Observability & Insights - Set up and maintain full observability stacks (logging, metrics, tracing) using tools like Prometheus, Grafana, Datadog, OpenTelemetry, or ELK. - Analyze telemetry and logs ...
Fircosoft Support Consultant
Plano, TX · On-site
Familiarity with cloud platforms (AWS/Azure) and monitoring tools (e.g., Datadog). * Familiarity with Linux scripting.
Fircosoft Support Consultant
Plano, TX · On-site
Familiarity with cloud platforms (AWS/Azure) and monitoring tools (e.g., Datadog). * Familiarity with Linux scripting.
Site Reliability Engineer
$55.25 - $73.50/hr
Observability & Insights - Set up and maintain full observability stacks (logging, metrics, tracing) using tools like Prometheus, Grafana, Datadog, OpenTelemetry, or ELK. - Analyze telemetry and logs ...
Site Reliability Engineer
$55.25 - $73.50/hr
Observability & Insights - Set up and maintain full observability stacks (logging, metrics, tracing) using tools like Prometheus, Grafana, Datadog, OpenTelemetry, or ELK. - Analyze telemetry and logs ...
Evaluate and evolve the observability toolset -- making principled decisions on how Datadog, Honeycomb, SumoLogic, OpenTelemetry, and Bugsnag work together as a unified platform * Build self-service ...
Evaluate and evolve the observability toolset -- making principled decisions on how Datadog, Honeycomb, SumoLogic, OpenTelemetry, and Bugsnag work together as a unified platform * Build self-service ...
Evaluate and evolve the observability toolset - making principled decisions on how Datadog, Honeycomb, SumoLogic, OpenTelemetry, and Bugsnag work together as a unified platform * Build self-service ...
Evaluate and evolve the observability toolset - making principled decisions on how Datadog, Honeycomb, SumoLogic, OpenTelemetry, and Bugsnag work together as a unified platform * Build self-service ...
Observability Engineer
Houston, TX · On-site
Experience with Datadog, AppDynamics, NewRelic or similar APM system
Observability Engineer
Houston, TX · On-site
Experience with Datadog, AppDynamics, NewRelic or similar APM system
Evaluate and evolve the observability toolset - making principled decisions on how Datadog, Honeycomb, SumoLogic, OpenTelemetry, and Bugsnag work together as a unified platform * Build self-service ...
Evaluate and evolve the observability toolset - making principled decisions on how Datadog, Honeycomb, SumoLogic, OpenTelemetry, and Bugsnag work together as a unified platform * Build self-service ...
Evaluate and evolve the observability toolset - making principled decisions on how Datadog, Honeycomb, SumoLogic, OpenTelemetry, and Bugsnag work together as a unified platform * Build self-service ...
Evaluate and evolve the observability toolset - making principled decisions on how Datadog, Honeycomb, SumoLogic, OpenTelemetry, and Bugsnag work together as a unified platform * Build self-service ...
Senior Golang AWS Developer
Plano, TX · On-site
$115K - $149K/yr
The right candidate brings advanced knowledge of Golang microservices, MongoDB, Kubernetes, observability tools (OpenTelemetry and Datadog), and modern CI/CD practices. Key Responsibilities * Design ...
Quick apply
Senior Golang AWS Developer
Plano, TX · On-site
$115K - $149K/yr
The right candidate brings advanced knowledge of Golang microservices, MongoDB, Kubernetes, observability tools (OpenTelemetry and Datadog), and modern CI/CD practices. Key Responsibilities * Design ...
Cloud Software Engineer
Westlake, TX · On-site
$57.50 - $75/hr
Java and Python coding AWS EKS Terraform Oracle, DB2, write SQL queries against oracle DB Kafka Jenkins Datadog
Quick apply
Cloud Software Engineer
Westlake, TX · On-site
$57.50 - $75/hr
Java and Python coding AWS EKS Terraform Oracle, DB2, write SQL queries against oracle DB Kafka Jenkins Datadog
Full Stack Java Developer
Plano, TX · On-site
$50.25 - $64.75/hr
Splunk, Datadog, Dynatrace, or Grafana
Full Stack Java Developer
Plano, TX · On-site
$50.25 - $64.75/hr
Splunk, Datadog, Dynatrace, or Grafana
Manager, Global Database Administration
Spring, TX · On-site
$262K/yr
Oversee architecture standards, monitoring tooling (Oracle Enterprise Manager, Microsoft System Center Operations Manager, Datadog), and the custom-built DB360 platform for automated password resets ...
Manager, Global Database Administration
Spring, TX · On-site
$262K/yr
Oversee architecture standards, monitoring tooling (Oracle Enterprise Manager, Microsoft System Center Operations Manager, Datadog), and the custom-built DB360 platform for automated password resets ...
Cloud Engineer
Plano, TX · On-site
$53.25 - $71.25/hr
Configure and monitor cloud resources using Prometheus, Grafana, Dynatrace, Splunk, Datadog , and similar tools. * Collaborate with networking teams to support service-to-service communication and ...
Quick apply
Cloud Engineer
Plano, TX · On-site
$53.25 - $71.25/hr
Configure and monitor cloud resources using Prometheus, Grafana, Dynatrace, Splunk, Datadog , and similar tools. * Collaborate with networking teams to support service-to-service communication and ...
System Operations Engineer - Cloud Infra
Austin, TX · On-site
$60 - $65/hr
Develop and maintain Datadog dashboards for real-time system monitoring, performance metrics, and alerting. * Implement logging and monitoring best practices to ensure high availability and rapid ...
Quick apply
System Operations Engineer - Cloud Infra
Austin, TX · On-site
$60 - $65/hr
Develop and maintain Datadog dashboards for real-time system monitoring, performance metrics, and alerting. * Implement logging and monitoring best practices to ensure high availability and rapid ...
Datadog information
What is the difference between Datadog vs Cloud Monitoring Engineer?
| Aspect | Datadog | Cloud Monitoring Engineer |
|---|---|---|
| Primary Role | Monitoring and analytics platform for IT infrastructure and applications | Designing, implementing, and managing cloud monitoring solutions |
| Required Skills | Cloud platforms, monitoring tools, scripting, API integration | Cloud services, monitoring tools, scripting, troubleshooting |
| Certifications | Cloud certifications (AWS, Azure), monitoring tools certifications | Cloud certifications (AWS, Azure), monitoring certifications |
| Work Environment | Using SaaS platform, integrating with various cloud and on-premise systems | Managing cloud infrastructure, configuring monitoring tools, troubleshooting |
While both roles involve cloud monitoring, Datadog focuses on utilizing a specific SaaS platform for analytics and monitoring, whereas a Cloud Monitoring Engineer designs and manages monitoring solutions across cloud environments. The roles often overlap in skills and certifications, but their core responsibilities differ in scope and focus.
What are the key skills and qualifications needed to thrive as a Datadog engineer?
What are some common challenges faced by Datadog engineers when implementing monitoring solutions for large-scale systems?
What is a Datadog engineer?

Full-time
Re-posted 7 days ago
Fidelity Investments rating
8.7
Based on 271 frontline employees who took The Breakroom Quiz
15th of 150 rated financial services
Job description
Note: Fidelity is not providing immigration sponsorship for this position.
The Role
We are seeking a technical problem solver with a strong background in production support to join our Major Incident Management team. This role combines Major Incident Management responsibilities with hands-on technical expertise across cloud, infrastructure, and application environments.
The ideal candidate is passionate about troubleshooting, thrives in high-pressure situations, and is eager to grow their technical and coordination skills in a dynamic environment.
The Expertise and Skills You Bring
- Requires Bachelors or equivalent with 2+ years of experience or Masters with 0+ years of experience
- A minimum of 2 + years of hybrid experience in Production Support, Development or SRE Experience. Hands-On experience developing or supporting highly distributed multi-tiered systems at scale
- Ability to triage while leading an incident call, perform root cause analysis, and be decisive under pressure
- A self-starter and team player who can independently manage multiple responsibilities in a dynamic environment
- Solid understanding of Cloud Computing and DevOps concepts including CI/CD Pipelines
- Perform second-level support using visibility tools such as Datadog and Splunk and expert hands on experience with one or more (Datadog, Splunk, Grafana)
- Understanding of ITIL processes (Incident, Problem, Change Management)
- Effective business communication and influencing skills
- Relevant certifications (ITIL, AWS, Azure) are a plus
- Bonus: Retail trading , Asset Trading, or Financial Services Technology experience is a plus
- Cloud Platforms: AWS, Azure
- Operating Systems: Unix, Linux, Windows Server
- Scripting and Development: Shell Script, .NET
- Middleware and Integration: Tomcat, Apache
- Monitoring Tools: Splunk, Datadog, Grafana,
- Databases: Oracle DB, MS-SQL, Sybase
- ITSM Tools: JIRA, ServiceNow
The Team
The Fidelity Support Center (FSC) is Fidelity's centralized enterprise monitoring and incident management hub. FSC supports a wide range of environments including application, distributed systems, mainframe, network, cloud, end-user computing, and security. As the first line of defense for production incidents, FSC provides detection, coordination, escalation, communication, and mitigation across the enterprise.
Fidelity's Onsite Working Model
Fidelity is transitioning to a full-time onsite working model through a phased rollout across regions and roles. Currently, some roles and locations require 100% onsite presence, while others require less. Onsite expectations are likely to evolve as the rollout continues. This transition does not apply to fully remote roles.
Certifications:
Category:
Information Technology
Please be advised that Fidelity's business is governed by the provisions of the Securities Exchange Act of 1934, the Investment Advisers Act of 1940, the Investment Company Act of 1940, ERISA, numerous state laws governing securities, investment and retirement-related financial activities and the rules and regulations of numerous self-regulatory organizations, including FINRA, among others. Those laws and regulations may restrict Fidelity from hiring and/or associating with individuals with certain Criminal Histories.
What Fidelity Investments employees say
Pay
Benefits
Hours and flexibility
Workplace
Get the full story on Breakroom
About Fidelity
Sourced by ZipRecruiter