1

Observability Site Reliability Engineer Jobs in Austin, TX

SRE/DevOps Engineer - GenAI

Austin, TX · On-site

$56.50 - $75/hr

Core SRE / DevOps Skills * Strong experience in DevOps / SRE roles supporting production systems ... Experience with monitoring and observability tools such as Prometheus, Grafana, and Splunk * Strong ...

Site Reliability Engineer

Austin, TX · On-site

$56.50 - $75/hr

... a Site Reliability Engineer to design, build, and operate the platforms that power AI Co-Workers ... and improve observability across monitoring, logging, and alerting • Partner closely with ...

Sr Site Reliability Engineer

Austin, TX · On-site

$56.50 - $75/hr

This role will contribute to the reliability, observability, and operational excellence of our platform infrastructure serving millions of users. As a Senior SRE, you will be a strong technical ...

Site Reliability Engineer

Austin, TX · On-site

$56.50 - $75/hr

They are seeking a Site Reliability Engineer to design, build, and operate platforms that support ... and improve observability across monitoring, logging, and alerting • Partner closely with ...

SRE ENGINEER

Austin, TX · On-site

$100K - $115K/yr

Role: SRE Engineer Location: Austin, TX Salary: $100000 to $115000/Annum Who are we looking for? We ... Monitor system health, set up ing and observability (e.g., Prometheus, Grafana), and proactively ...

Sr Site Reliability Engineer

Austin, TX · On-site

$56.50 - $75/hr

This role will contribute to the reliability, observability, and operational excellence of our platform infrastructure serving millions of users. As a Senior SRE, you will be a strong technical ...

Senior Site Reliability Engineer This role contributes to the reliability, observability, and operational excellence of our platform infrastructure serving millions of users. As a Senior SRE, you ...

Site Reliability Engineer

Austin, TX · On-site

$56.50 - $75/hr

... and improve observability across monitoring, logging, and alerting • Partner closely with ... Ruby • Direct Site Reliability Engineer experience or equivalent, including reliability ...

Site Reliability Engineer

Austin, TX · On-site

$56.50 - $75/hr

Build and improve observability across monitoring, logging, and alerting * Partner closely with ... professional experience in Site Reliability Engineering or DevOps Engineering * Kubernetes ...

Site Reliability Engineer

Austin, TX · On-site

$130 - $180/hr

Site Reliability Engineer Department: Infrastructure Employment Type: Full Time Location: Austin ... Observability: Apply best practices for monitoring and alerting using tools such as Prometheus ...

Senior Site Reliability Engineer

Austin, TX · On-site

$56.50 - $75/hr

Build and maintain CI/CD pipelines, observability stacks, and incident response workflows * Define ... SRE, DevOps, or infrastructure engineering * Hands-on experience across at least two major cloud ...

next page

Showing results 1-20

Observability Site Reliability Engineer information

See Austin, TX salary details

$10

$63

$91

How much do observability site reliability engineer jobs pay per hour?

As of Aug 11, 2026, the average hourly pay for observability site reliability engineer in Austin, TX is $63.18, according to ZipRecruiter salary data. Most workers in this role earn between $54.33 and $72.21 per hour, depending on experience, location, and employer.

What is the difference between Observability Site Reliability Engineer vs Monitoring Engineer?

AspectObservability Site Reliability EngineerMonitoring Engineer
FocusEnsuring system reliability through observability, automation, and incident responseImplementing and managing monitoring tools and dashboards
SkillsCloud platforms, scripting, incident management, observability toolsMonitoring tools, alerting systems, data analysis
Work EnvironmentDevOps teams, cloud infrastructure, large-scale systemsOperations teams, infrastructure monitoring

While both roles involve system health, the Observability Site Reliability Engineer focuses on comprehensive system reliability using observability practices, whereas Monitoring Engineers primarily manage monitoring tools and alerts. The SRE role emphasizes automation, incident response, and system resilience, making it broader in scope.

What are popular job titles related to Observability Site Reliability Engineer jobs in Austin, TX? For Observability Site Reliability Engineer jobs in Austin, TX, the most frequently searched job titles are:
What job categories do people searching Observability Site Reliability Engineer jobs in Austin, TX look for? The top searched job categories for Observability Site Reliability Engineer jobs in Austin, TX are:
What cities near Austin, TX are hiring for Observability Site Reliability Engineer jobs? Cities near Austin, TX with the most Observability Site Reliability Engineer job openings:

Senior Site Reliability Engineer ( SRE)

Charles Schwab Inc.

Austin, TX • On-site

$130K - $175K/yr

Full-time

Posted 19 days ago


Job description

Your Opportunity
At Schwab, you're empowered to make an impact on your career. Here, innovative thought meets creative problem solving, helping us challenge the status quo and transform the finance industry together. We believe in the importance of in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location(s).
As a Senior Reliability Engineer, you will help shape the reliability, scalability, and operational excellence of mission-critical enterprise platforms that support our clients and business operations. This role blends software engineering, systems engineering, and operational expertise to drive resilient, high-performing technology solutions in complex distributed environments. Working closely with engineering, product, scrum, and operations teams, you will apply Site Reliability Engineering (SRE) principles to improve system availability, accelerate delivery, reduce operational complexity, and strengthen platform observability.
You will play a key role in designing and implementing automation solutions that minimize manual effort, improve operational efficiency, and enhance system reliability at scale. Leveraging modern observability practices, AI/ML-enabled operational capabilities, and proactive monitoring strategies, you will help identify risks, detect anomalies, improve incident response, and support data-driven decision making. Your work will influence platform stability through automation, predictive insights, deployment optimization, rollout validation, capacity planning, and continuous improvement initiatives.
Success in this role requires strong problem-solving skills, sound technical judgment, and the ability to collaborate across teams to address complex operational challenges. You will contribute to the evolution of reliability practices, champion automation-first approaches, support continuous delivery initiatives, and help build systems that enable teams to operate with greater confidence, agility, and efficiency. As part of a highly collaborative environment, you will also participate in incident response and on-call support while helping drive long-term improvements that enhance customer and business outcomes.
What you have
Required Qualifications
  • 8+ years of experience supporting and administering enterprise-scale applications, platforms, or infrastructure environments.
  • 6+ years of experience developing automation solutions, operational tooling, monitoring dashboards, and alerting frameworks.
  • 6+ years of experience applying Software Development Life Cycle (SDLC) practices and driving process improvement initiatives.
  • Experience with Site Reliability Engineering (SRE), production operations, system monitoring, deployment management, and operational excellence practices.
  • Experience administering and supporting Linux and Windows Server environments, including troubleshooting, performance tuning, and system optimization.
  • Experience deploying, supporting, configuring, or migrating cloud-based applications and platforms.
  • Knowledge of networking fundamentals including DNS, DHCP, firewalls, routing, and related infrastructure technologies.
  • Experience supporting large-scale distributed systems and highly available application architectures.
  • Proficiency in one or more programming or scripting languages such as Python, Java, PowerShell, Bash, or .NET.
  • Experience working with relational or NoSQL database technologies including SQL Server, Oracle, or MongoDB.
  • Knowledge of messaging and event-streaming technologies such as Kafka, RabbitMQ, IBM MQ, or Solace.
  • Experience with observability and monitoring platforms such as Splunk, AppDynamics, or similar tools.
  • Experience applying AI/ML-powered operational practices, including anomaly detection, predictive alerting, AIOps, or ML-assisted observability capabilities.
  • Bachelor's degree in computer science, Information Technology, Engineering, or a related field.

Preferred Qualifications
  • 8+ years of experience in enterprise technology operations, platform engineering, or reliability engineering.
  • Experience within the financial services industry.
  • Experience working in Agile environments and cross-functional product teams.
  • Hands-on experience with AIOps platforms, intelligent automation solutions, or ML-driven observability tools.
  • Experience integrating AI/ML capabilities into operational automation, deployment workflows, or continuous delivery processes.
  • Experience with CI/CD technologies such as Jenkins, Harness, GitHub Actions, or similar platforms.
  • Experience implementing GitOps practices and infrastructure automation strategies.
  • Experience with containerization and orchestration platforms such as Kubernetes or OpenShift.
  • Experience supporting cloud platforms such as Google Cloud Platform (GCP), AWS, or Microsoft Azure.
  • Demonstrated ability to influence technical strategy, drive operational improvements, and lead reliability-focused initiatives across teams.

In addition to the salary range, this role is eligible for bonus or incentive opportunities.