1

Cloud Observability Engineer Jobs in Texas (NOW HIRING)

Strong Linux and cloud platform administration skills. * Experience with scripting and automation ... Observability Engineering * Platform Engineering * Linux Administration * AWS Cloud Services

Grafana and Observability Engineer

Coppell, TX ยท On-site

$53 - $70.50/hr

Administer and support Grafana Cloud and on prem Grafana environments. * Build and deploy ... Champion observability as a core engineering discipline across the organization. Required ...

Senior Observability Engineer

Coppell, TX ยท On-site

$100K - $137K/yr

Job#: 3046566 Senior Observability Engineer Location(s): Dallas, TX | Tampa, FL | Jersey City, NJ ... Administer and support Grafana Cloud and on-premises Grafana deployments. * Design and implement ...

Agentic AI & Observability Engineer

Plano, TX ยท On-site

$50.75 - $69.50/hr

Role: Agentic AI & Observability Engineer Location: Plano, TX or Mc Lean VA Prefer Local Role ... Design and integrate solutions with AWS services and cloud-native architectures . * Build and ...

Senior ITSMA Observability Engineer

Dallas, TX ยท On-site +1

$103K - $142K/yr

The Senior ITSMA Observability Engineer is responsible for the design and development of the ... Experience with Elastic Cloud and AWS Managed Prometheus * Knowledge of installation, system tasks ...

Senior ITSMA Observability Engineer

Dallas, TX ยท On-site

$103K - $142K/yr

The Senior ITSMA Observability Engineer is responsible for the design and development of the ... Experience with Elastic Cloud and AWS Managed Prometheus * Knowledge of installation, system tasks ...

Reliability Observability Engineer 2

Irving, TX

$54.75 - $72.75/hr

Expertise in Observability Engineering , Site Reliability Engineering (SRE) , or Reliability ... Strong understanding of distributed systems, microservices, cloud platforms, and Kubernetes

Senior Observability Engineers

Dallas, TX ยท On-site

$103K - $142K/yr

Senior Observability Engineers Location: Dallas, TX (Hybrid 3 days/week) I hope you're having a ... Linux administration and cloud platform experience. Automation/scripting with Python, PowerShell ...

We're looking for OSS Telemetry Observability Engineer to grow our team. This role will be hands on ... Experience with cloud platforms (AWS, GCP, Azure) * Understanding of SLO/SLI/SLA concepts

We're looking for OSS Telemetry Observability Engineer to grow our team. This role will be hands on ... Experience with cloud platforms (AWS, GCP, Azure) * Understanding of SLO/SLI/SLA concepts

Cloud Engineer, Specialist

Dallas, TX ยท On-site

$55.25 - $73.75/hr

Design, build, and support cloud-based observability platforms, including logging, metrics, and ... Partner with engineering teams to onboard applications and promote observability best practices

next page

Showing results 1-20

Cloud Observability Engineer information

What does a cloud observability engineer do?

A Cloud Observability Engineer is responsible for ensuring the visibility, monitoring, and performance of cloud-based systems and applications. They implement tools and processes to collect metrics, logs, and traces, allowing organizations to detect issues, optimize resources, and maintain system health. Their role often involves collaborating with development and operations teams to troubleshoot problems and improve the overall reliability of cloud services.

What are the key skills and qualifications needed to thrive as a cloud observability engineer?

To thrive as a Cloud Observability Engineer, you need a solid background in cloud platforms (such as AWS, Azure, or GCP), monitoring frameworks, and scripting or programming languages, often supported by a degree in computer science or related certifications. Familiarity with observability tools like Datadog, Prometheus, Grafana, and log aggregation systems, as well as CI/CD pipelines, is typically required. Strong problem-solving, analytical thinking, and communication skills help you proactively detect issues and collaborate across teams. These skills and qualities are crucial for maintaining system reliability, optimizing performance, and ensuring rapid incident response in dynamic cloud environments.

How does a cloud observability engineer typically collaborate with development and operations teams?

Cloud Observability Engineers work closely with both development and operations teams to ensure system reliability and performance. They are often responsible for setting up monitoring tools, defining key performance metrics, and proactively identifying issues before they impact users. Regular collaboration involves participating in incident response, sharing insights from logs and metrics, and providing recommendations for architectural improvements. This cross-functional teamwork helps bridge gaps between software deployment and infrastructure management, fostering a culture of shared responsibility for system health.

What cities in Texas are hiring for Cloud Observability Engineer jobs?

Cities in Texas with the most Cloud Observability Engineer job openings:

Infographic showing various Cloud Observability Engineer job openings in Texas as of August 2026, with employment types broken down into 89% Full Time, 6% Part Time, and 5% Contract. Highlights an 85% Physical, 5% Hybrid, and 10% Remote job distribution.

Observability Engineer

Irving, TX โ€ข On-site

Indotronix International Corporation
Recruiting and Staffing Servicesย โ€ขย 1 - 5K employees

$53 - $70.25/hr

Full-time

Re-posted 3 days ago


Job description

Job Summary: Senior Observability Engineer - Irving, TX (Onsite)
About the Role
Join a dynamic technology team as a Senior Observability Engineer based in Irving, TX. You will architect and lead end-to-end observability solutions across complex, hybrid environments, transforming platform telemetry into actionable insights that drive reliability and performance. This is an opportunity to shape observability strategy, work with cutting-edge tools, and collaborate with engineering, operations, and leadership. Advance your career by building and standardizing monitoring frameworks at scale.
Responsibilities
- Architect and implement comprehensive observability frameworks across cloud, on-premises, networking, databases, middleware, and applications
- Evaluate, select, and integrate observability and monitoring tools, establishing reference architectures
- Design and maintain standardized Grafana dashboards for platform and workload health (OCP, AKS, GKE)
- Define golden signals and platform health KPIs tied to availability, performance, and reliability
- Serve as an advanced Splunk user: develop complex SPL queries, dashboards, and root-cause investigations
- Correlate logs, metrics, and events across Grafana and Splunk to drive rapid incident resolution (MTTR reduction)
- Implement and tune platform-specific observability for Kubernetes platforms (OCP, AKS, GKE)
- Configure and manage ThousandEyes for synthetic monitoring and network intelligence
- Administer BigPanda for AIOps-driven event correlation and noise reduction
- Integrate ServiceNow for automated incident creation and enriched alerting
- Document observability standards, dashboards, and onboarding processes
Required Skills and Experience
- 7+ years in IT operations, SRE, systems engineering, or infrastructure
- 5+ years designing and implementing observability/monitoring solutions across distributed systems
- 5+ years production experience with Kubernetes platforms
- Expertise in logging, tracing, and enhanced monitoring
- Strong hands-on experience with Grafana (dashboards, alerts, data sources)
- Advanced Splunk SPL, dashboarding, and investigation capabilities
- Proficiency with querying languages (SQL, PromQL)
- Deep understanding of Kubernetes internals, OpenShift (OCP), AKS, GKE
- Experience with Prometheus, OpenTelemetry, Kubernetes exporters
- OS-level monitoring (Linux/Windows) and network fundamentals
- Experience with ThousandEyes, BigPanda, and ServiceNow ITSM workflows
Preferred Skills
- Experience designing SLOs/SLIs, reliability scorecards
- Familiarity with Istio, service mesh metrics, mTLS
- Capacity planning and trend analysis using observability data
- Exposure to multi-cloud observability strategies
- Monitoring for databases, message brokers, middleware
- Familiarity with AIOps or ML-driven anomaly detection
Benefits
- Work in a collaborative, technology-driven environment
- Direct impact on reliability, performance, and operational excellence
- Exposure to the latest observability and cloud-native technologies
- Career growth opportunities in a high-visibility engineering role
How to Apply
Ready to drive observability excellence and elevate platform reliability? Submit your resume today to join our Irving, TX team and advance your career as a Senior Observability Engineer.

Indotronix logo

About Indotronix

Sourced by ZipRecruiter

In 1986, Indotronix established itself in the staffing space. 22 years later, Avani entered the scene, offering consulting and technology development. Finally, in 2016, the two joined forces to begin delivering talent across all areas, from Staffing to Consulting to unique platform development.

Industry

Recruiting and staffing services

Company size

1,001 - 5,000 Employees

Headquarters location

Rochester, NY, US