1

Senior Observability Engineer Jobs in Florida (NOW HIRING)

Lead Systems Engineer

Tampa, FL ยท On-site

  • Medical

  • Life

  • Retirement

  • PTO

The Senior Associate Lead Systems Engineer - Endpoint Management Observability Engineer will serve as a technical lead responsible for supporting and operating the observability model that executes ...

Sr Software Engineer - AI DevEx

Aventura, FL ยท On-site

$176K - $196K/yr

As a Senior Engineer within the Developer Intelligence pillar, you will be an autonomous contributor focused on key components of our platform - working across observability, engineering telemetry ...

Enterprise Observability Architect

Tampa, FL ยท On-site

$62.75 - $81/hr

  • Medical

  • Life

  • Retirement

  • PTO

... senior stakeholders across engineering, infrastructure, SRE, security, risk, and business ... Define multi-year roadmaps for observability modernization, including OpenTelemetry adoption ...

Observability Systems Associate

Lake Mary, FL ยท On-site

$50K/yr

  • Retirement

  • PTO

Perform daily monitoring reviews across observability and security platforms. * Identify, triage ... Support senior engineers by testing and refining scripts used for proactive remediation and system ...

Senior Software Engineer (SRE)

Miami, FL ยท On-site

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

1. General Information Job Title Senior Software Engineer Department Engineering Location Miami, FL ... Design and implement robust monitoring, alerting, and observability systems across all services and ...

Back Senior Security Engineer (Akamai) Infrastructure & Security Bay Lake , FL Contract Hybrid Aug ... Participate in ongoing improvement of observability and resilience standards Required ...

New

Senior AI Engineer

Orlando, FL ยท On-site

$97K - $134K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Senior AI Engineer Orlando, FL (Hybrid) Position Summary Thales is looking for a Senior AI Engineer ... observability, and continuous optimization within production environments. * Architect, develop ...

Senior Security Engineer (Akamai)

Orlando, FL ยท Hybrid

$106K - $146K/yr

Senior Security Engineer - Akamai (Remote/Hybrid) Location: Need to be located in hubs (Orlando ... Participate in ongoing improvement of observability and resilience standards Required ...

Senior Software Engineer, AI Platform

Orlando, FL ยท On-site

$140.54 - $223.21/hr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Senior Software Engineer, AI Platform Location: Orlando HQ - Remote, FL Job Id: 677 # of Openings ... Build and maintain observability, telemetry, tracing, centralized logging, governance, and policy ...

Senior Software Engineer (Sre)

Miami, FL ยท On-site

$140 - $200/hr

  • Life

  • Retirement

  • PTO

As a Senior Software Engineer in SRE at eMed, you will play a key role in ensuring our platform is ... Design and implement robust monitoring, alerting, and observability systems across all services and ...

next page

Showing results 1-20

Senior Observability Engineer information

What is a senior observability engineer?

A Senior Observability Engineer is a seasoned IT professional responsible for designing, implementing, and maintaining systems that monitor and provide insights into the performance, health, and reliability of software applications and infrastructure. They utilize tools for logging, monitoring, tracing, and alerting to ensure that systems are observable and any issues can be quickly detected and resolved. In addition to technical expertise, they often collaborate with development and operations teams to establish best practices, improve incident response, and optimize system performance. Their work is crucial for maintaining uptime, enhancing customer experiences, and supporting the scalability of technology platforms.

How does a senior observability engineer typically collaborate with development and operations teams?

A Senior Observability Engineer works closely with both development and operations teams to ensure robust monitoring, logging, and tracing solutions are in place across all applications and infrastructure. They often participate in architecture discussions to advise on best practices for instrumenting code and systems for observability. By analyzing metrics and alerting patterns, they help teams proactively resolve issues and optimize system performance. This role also involves mentoring engineers on observability tools and fostering a culture of transparency and accountability in incident response.

What are the key skills and qualifications needed to thrive as a senior observability engineer, and why are they important?

To thrive as a Senior Observability Engineer, you need expertise in monitoring, logging, and tracing systems, with a solid background in computer science or a related field. Familiarity with tools like Prometheus, Grafana, ELK stack, and cloud platforms, as well as certifications such as AWS Certified DevOps Engineer, are typically required. Strong problem-solving, collaboration, and communication skills are critical for effectively diagnosing and resolving complex infrastructure issues. These skills ensure reliable system performance, rapid incident response, and continuous improvement of the technology environment.

What is the difference between Senior Observability Engineer vs Site Reliability Engineer?

AspectSenior Observability EngineerSite Reliability Engineer
CredentialsExperience with monitoring tools, scripting, cloud platformsSame as Senior Observability Engineer, often with SRE certifications
Work EnvironmentFocus on monitoring, logging, and tracing systemsFocus on system reliability, automation, and incident response
Industry UsageUsed in tech companies emphasizing system observabilityCommon in large-scale tech and cloud services
Search/Comparison IntentOften compared for monitoring rolesCompared for reliability and system stability roles

While both roles require expertise in cloud platforms and scripting, the Senior Observability Engineer primarily focuses on designing and maintaining monitoring, logging, and tracing systems to ensure system visibility. In contrast, a Site Reliability Engineer emphasizes system reliability, automation, and incident management to maintain service uptime. Both roles are vital in tech environments but serve different core functions related to system health and stability.

What are the most commonly searched types of Observability Engineer jobs in Florida?

The most popular types of Observability Engineer jobs in Florida are:

What are popular job titles related to Senior Observability Engineer jobs in Florida?

For Senior Observability Engineer jobs in Florida, the most frequently searched job titles are:

What job categories do people searching Senior Observability Engineer jobs in Florida look for?

The top searched job categories for Senior Observability Engineer jobs in Florida are:

What cities in Florida are hiring for Senior Observability Engineer jobs?

Cities in Florida with the most Senior Observability Engineer job openings:

Infographic showing various Senior Observability Engineer job openings in Florida as of August 2026, with employment types broken down into 90% Full Time, 5% Part Time, and 5% Contract. Highlights an 86% Physical, 5% Hybrid, and 9% Remote job distribution.

Lead Systems Engineer

DTCC

Tampa, FL โ€ข On-site

Other

Medical, Life, Retirement, PTO

Re-posted 16 days ago


Job description

Are you ready to make an impact at DTCC?
Do you want to work on innovative projects, collaborate with a dynamic and supportive team, and receive investment in your professional development? At DTCC, we are at the forefront of innovation in the financial markets. We are committed to helping our employees grow and succeed. We believe that you have the skills and drive to make a real impact. We foster a thriving internal community and are committed to creating a workplace that looks like the world that we serve.
The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted infrastructure of the global capital markets. The team delivers high-quality information through activities that include development of essential, building infrastructure capabilities to meet client needs and implementing data standards and governance.
Pay and Benefits:
  • Competitive compensation, including base pay and annual incentive
  • Comprehensive health and life insurance and well-being benefits, based on location
  • Pension / Retirement benefits
  • Paid Time Off and Personal/Family Care, and other leaves of absence when needed to support your physical, financial, and emotional well-being.
  • DTCC offers a flexible/hybrid model of 3 days onsite and 2 days remote (onsite Tuesdays, Wednesdays and a third day unique to each team or employee).

The Impact you will have in this role:
The Senior Associate Lead Systems Engineer - Endpoint Management Observability Engineer will serve as a technical lead responsible for supporting and operating the observability model that executes accelerated enterprise patch management. The role focuses on measuring patch management success, monitoring endpoint health during rapid update cycles, identifying user-impacting conditions early, and improving the overall Digital Employee Experience (DEX). This position will correlate telemetry from endpoint management, monitoring, and experience platforms to provide actionable visibility into deployment progress, endpoint stability, remediation needs, and risk exposure during standard, accelerated, and critical patch events. The role partners closely with Endpoint Engineering, Security, Vulnerability Management, Application Packaging, VDI, Service Desk, Operations, and vendor teams to ensure patch execution is measurable, auditable, and aligned with enterprise risk and control expectations.
Your Primary Responsibilities:
  • Lead the design and continuous improvement of an Endpoint Management Observability capability supporting accelerated patch management across Windows endpoints, virtual desktops, and managed endpoint platforms.

  • Define and report patch success measures, including deployment coverage, install success and failure rates, reboot compliance, device check-in health, remediation progress, and vulnerability exposure reduction.

  • Monitor user impact throughout patch waves by analyzing endpoint performance, system stability, application health, reboot behavior, login experience, device responsiveness, and service desk signals.

  • Correlate patch deployment rings, pilot groups, phased rollouts, and accelerated response timelines with telemetry to identify anomalies, failing device populations, and emerging DEX issues.

  • Support 24-hour, 48-hour, and other accelerated patch response models by providing command-center visibility into deployment status, risk areas, user impact, and remediation actions.

  • Partner with engineering teams to establish pre-deployment readiness checks, post-deployment validation, regression indicators, and operational go/no-go criteria for patch waves.

  • Drive remediation coordination for failed patch installs, unhealthy devices, repeated reboot failures, performance degradation, application compatibility issues, and other endpoint conditions that may reduce patch completion or user experience.

  • Translate observability insights into clear executive, operational, and audit-ready reporting that communicates patch progress, residual risk, user impact, and required actions.

  • Work with Security and Vulnerability Management teams to align patch observability with vulnerability remediation priorities, risk exposure, and evidence requirements.

  • Collaborate with Application Packaging, VDI, Service Desk, Operations, and vendor teams to resolve application or platform issues identified during patch cycles, including high CPU, memory, disk, or application failure patterns.

  • Create, maintain, and improve documentation for observability controls, dashboards, alert thresholds, escalation workflows, remediation tracking, and evidence collection.

  • Provide advanced technical support and guidance to desktop support and endpoint operations teams during patch deployment, monitoring, and remediation activities.

  • Align risk and control processes into day-to-day responsibilities to monitor and mitigate risk; escalates appropriately.

  • Improved visibility into patch deployment progress, endpoint health, remediation status, and user-impacting conditions during patch events.

  • Higher patch completion and remediation effectiveness through telemetry-driven identification of failures, blockers, and unhealthy device populations.
    Reduced risk exposure through faster detection, prioritization, and escalation of patch gaps aligned to accelerated response requirements.
    Improved DEX during patch cycles by proactively identifying performance degradation, application instability, reboot friction, and service-impacting trends.
    Consistent, audit-ready evidence that demonstrates patch management execution, monitoring controls, remediation actions, and risk-based decision-making.

Qualifications:
  • Minimum of 6 years of related experience
  • Bachelor's degree preferred or equivalent experience

Talents Needed for Success:
  • 6+ years of experience in systems engineering, endpoint management, infrastructure operations, or digital workplace engineering, with demonstrated technical leadership responsibility.

  • Hands-on experience supporting enterprise patch management at scale using SCCM/MECM, Microsoft Intune, co-management, Windows Update for Business, or comparable endpoint management platforms.

  • Experience building or operating observability, monitoring, and reporting capabilities for endpoint health, patch compliance, application stability, and user experience.

  • Strong understanding of vulnerability management, accelerated remediation practices, risk prioritization, and enterprise control expectations.

  • Ability to analyze telemetry, identify trends, isolate device or application failure patterns, and convert findings into actionable remediation plans.

  • Experience with PowerShell, automation, KQL/Log Analytics, Power BI, or related reporting and analytics technologies preferred.

  • Familiarity with Digital Employee Experience frameworks and endpoint experience platforms such as SysTrack/Lakeside, ZDX, or comparable monitoring solutions.

  • Experience with Azure, AVD, Windows 365, VDI, cloud-managed endpoints, or hybrid endpoint environments preferred.

  • Strong written and verbal communication skills with the ability to prepare executive-ready, operational, and audit-ready reporting.

  • ITIL alignment and experience working in regulated enterprise environments preferred.

The salary range is indicative for roles at the same level within DTCC across all US locations. Actual salary is determined based on the role, location, individual experience, skills, and other considerations. We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, sex, gender, gender expression, sexual orientation, age, marital status, veteran status, or disability status. We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.
About Us
With over 50 years of experience, DTCC is the premier post-trade market infrastructure for the global financial services industry. From 20 locations around the world, DTCC, through its subsidiaries, automates, centralizes, and standardizes the processing of financial transactions, mitigating risk, increasing transparency, enhancing performance and driving efficiency for thousands of broker/dealers, custodian banks and asset managers. Industry owned and governed, the firm innovates purposefully, simplifying the complexities of clearing, settlement, asset servicing, transaction processing, trade reporting and data services across asset classes, bringing enhanced resilience and soundness to existing financial markets while advancing the digital asset ecosystem. In 2024, DTCC's subsidiaries processed securities transactions valued at U.S. $3.7 quadrillion and its depository subsidiary provided custody and asset servicing for securities issues from over 150 countries and territories valued at U.S. $99 trillion. DTCC's Global Trade Repository service, through locally registered, licensed, or approved trade repositories, processes more than 25 billion messages annually. To learn more, please visit us at or connect with us on LinkedIn , X , YouTube , Facebook and Instagram .
DTCC proudly supports Flexible Work Arrangements favoring openness and gives people freedom to do their jobs well, by encouraging diverse opinions and emphasizing teamwork. When you join our team, you'll have an opportunity to make meaningful contributions at a company that is recognized as a thought leader in both the financial services and technology industries. A DTCC career is more than a good way to earn a living. It's the chance to make a difference at a company that's truly one of a kind.
Learn more about Clearance and Settlement by clicking here .
About the Team
Serves as a dedicated technology resource for advancing DTCC's business opportunities and providing industry thought leadership for leveraging new technology. The goal of this new department is to partner internally with IT, our business and regulatory divisions and externally with clients, regulators, and fintech vendors, to help build new platforms and business models to advance DTCC's mission to support the financial markets.