1

Sr Site Reliability Engineer Jobs in Texas (NOW HIRING)

Role: Senior SRE Engineer Location: Washington DC - Hybrid We are seeking a high-caliber Senior SRE Engineer to join a premier client in Washington, DC , to spearhead the evolution of their ...

Engineer III, Site Reliability

Austin, TX ยท On-site

$130 - $185/hr

You'll work sideโ€‘byโ€‘side with a Senior SRE in a true playerโ€‘coach model, gaining handsโ€‘on mentorship while helping define how reliability, observability, and incident response are built at ...

Sr Site Reliability Engineer

Austin, TX ยท On-site

$56.50 - $75/hr

We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization, reporting to the Director, Operations Excellence. This role will contribute to the ...

Sr Site Reliability Engineer

Austin, TX ยท On-site

$130 - $160/hr

Senior Site Reliability Engineer This role contributes to the reliability, observability, and operational excellence of our platform infrastructure serving millions of users. As a Senior SRE, you ...

Sr Site Reliability Engineer

Austin, TX ยท On-site

$56.50 - $75/hr

We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization, reporting to the Director, Operations Excellence. This role will contribute to the ...

Sr Site Reliability Engineer

Austin, TX ยท On-site

$56.50 - $75/hr

We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence organization, reporting to the Director, Operations Excellence. This role will contribute to the ...

Senior Site Reliability Engineer

Austin, TX ยท On-site +1

$56.50 - $75/hr

The role We're looking for a Senior SRE to own the reliability, scalability, and operational posture of Satsuma's multi-cloud infrastructure. You'll be the person who keeps things running, builds the ...

Senior Site Reliability Engineer

Plano, TX ยท On-site

$54.50 - $72.50/hr

The Senior Site Reliability Engineer acts as an advanced senior individual contributor responsible for designing, implementing, and maturing reliability engineering capabilities. The role focuses on ...

Sr. SRE Engineer (IAM focus)

Plano, TX ยท On-site

$54.50 - $72.50/hr

Position: Sr. SRE Engineer (IAM focus) (26-06661) Location: Plano, TX / Charlotte, NC / Pennington, NJ (3 days onsite, 2 days remote) Duration: 12-month contract (extension possible up to 36 months ...

Senior Site Reliability Engineer

Austin, TX

$56.50 - $75/hr

The role We're looking for a Senior SRE to own the reliability, scalability, and operational posture of Satsuma's multi-cloud infrastructure. You'll be the person who keeps things running, builds the ...

Senior Site Reliability Engineer

Austin, TX ยท On-site

$140 - $210/hr

The role We're looking for a Senior SRE to own the reliability, scalability, and operational posture of Satsuma's multi-cloud infrastructure. You'll be the person who keeps things running, builds the ...

Senior Site Reliability Engineer

Coppell, TX ยท On-site

$53 - $70.50/hr

Senior Site Reliability Engineer -- combination of deep operational expertise and hands-on engineering ability. The majority of your time (~70%) will be focused on owning the reliability ...

Senior Site Reliability Engineer

Plano, TX ยท On-site

$54.50 - $72.50/hr

As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products. As part of a new SRE team supporting ...

Senior Site Reliability Engineer

Plano, TX

$54.50 - $72.50/hr

As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products. As part of a new SRE team supporting ...

next page

Showing results 1-20

Sr Site Reliability Engineer information

See Texas salary details

$10

$59

$85

How much do sr site reliability engineer jobs pay per hour?

As of Sep 5, 2026, the average hourly pay for sr site reliability engineer in Texas is $59.39, according to ZipRecruiter salary data. Most workers in this role earn between $51.06 and $67.84 per hour, depending on experience, location, and employer.

What is a Sr Site Reliability Engineer?

Sr Site Reliability Engineers (SREs) are experienced IT professionals who ensure that software systems are reliable, scalable, and efficient. They bridge the gap between software development and IT operations by automating processes, monitoring system health, and responding to incidents. Their responsibilities include designing robust infrastructure, optimizing performance, and implementing best practices for system reliability. Additionally, Sr SREs mentor junior engineers and often lead initiatives to improve system resilience across the organization.

What skills and qualifications are needed to be a Sr Site Reliability Engineer?

To thrive as a Sr Site Reliability Engineer, you need deep expertise in software engineering, systems administration, and cloud infrastructure, typically supported by a degree in computer science or related field and several years of relevant experience. Proficiency with tools such as Kubernetes, Docker, Terraform, monitoring systems like Prometheus, and scripting languages like Python or Go, as well as certifications like AWS Certified DevOps Engineer, are common requirements. Strong problem-solving skills, effective communication, and the ability to work under pressure set outstanding SREs apart. These skills and qualities ensure systems are reliable, scalable, and resilient, directly supporting business continuity and user satisfaction.

What challenges do Sr Site Reliability Engineers face when balancing reliability and rapid software delivery?

Sr Site Reliability Engineers (SREs) often need to strike a balance between ensuring system stability and supporting fast-paced software releases. A primary challenge is implementing automation and monitoring solutions that reduce manual intervention without hindering developer agility. SREs regularly collaborate with development and operations teams to establish service-level objectives (SLOs), prioritize incident response, and address technical debt, all while advocating for best practices in reliability. This balancing act requires strong communication skills and a proactive mindset to anticipate potential risks and minimize downtime.

What is the difference between Sr Site Reliability Engineer vs Site Reliability Engineer?

AspectSr Site Reliability EngineerSite Reliability Engineer
Required CredentialsTypically requires 5+ years experience, advanced knowledge of cloud platforms, scripting, and monitoring toolsUsually 2-4 years experience, foundational knowledge of systems, scripting, and cloud services
Work EnvironmentLeads complex projects, mentors junior staff, and handles high-availability systemsSupports system reliability, automates deployment, and monitors infrastructure
Employer & Industry UsageCommon in large tech companies, finance, and cloud providersWidely used across tech startups, mid-sized firms, and cloud services

The main difference is experience level and responsibility. Sr Site Reliability Engineers handle more complex systems, lead projects, and mentor others, while Site Reliability Engineers focus on maintaining system stability and automation at a foundational level.

Is a senior site reliability engineer a stressful job?

A senior site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often involves on-call duties, troubleshooting complex issues, and working with monitoring tools, which can contribute to work-related stress. However, it also offers opportunities for problem-solving and skill development in a dynamic environment.

What does a senior site reliability engineer do?

A senior site reliability engineer (SRE) is responsible for maintaining and improving the reliability, availability, and performance of large-scale systems and services. They develop automation tools, monitor system health, troubleshoot issues, and implement best practices to ensure continuous operation, often using skills in scripting, cloud platforms, and incident management. Their role combines software engineering with systems administration to enhance system resilience and efficiency.

What cities in Texas are hiring for Sr Site Reliability Engineer jobs?

Cities in Texas with the most Sr Site Reliability Engineer job openings:

Infographic showing various Sr Site Reliability Engineer job openings in Texas as of August 2026, with employment types broken down into 1% As Needed, 85% Full Time, 12% Part Time, and 2% Contract. Highlights an 93% Physical, 3% Hybrid, and 4% Remote job distribution, with an average salary of $123,522 per year, or $59.4 per hour.

Senior Site Reliability Engineer

Electric Power Engineers

Austin, TX โ€ข Remote

$56.50 - $75/hr

Full-time

Medical, Dental, Vision, Retirement, PTO

Posted 11 days ago


Job description

Overview

We are designing the grid of the future!ย 

We are seeking an experienced Senior Site Reliability Engineer (SRE) to join our engineering organization. The ideal candidate will combine strong software engineering and cloud infrastructure expertise to improve the reliability, scalability, security, and operational efficiency of our systems. This role will focus on building resilient AWS and Kubernetes platforms, defining and measuring service reliability, automating operational work, and improving how we detect, respond to, and learn from production incidents.ย 

The Senior SRE will work closely with software engineering, platform, security, QA, and product teams to establish reliability standards and ensure production services can scale safely as the business grows.ย 

Responsibilities

Howย you can make an impact:

  • Service Reliability & Availability: Define, measure, and improve service reliability using service-level indicators (SLIs), service-level objectives (SLOs), error budgets, availability targets, and capacity planning.
  • Observability: Build and maintain monitoring, logging, tracing, dashboards, and alerting using tools such as CloudWatch, Prometheus, Grafana, New Relic, or similar platforms. Ensure alerts are actionable and aligned to customer and service impact.
  • Incident Response: Participate in and improve production incident response, including on-call practices, troubleshooting, escalation, communication, root-cause analysis, and blameless post-incident reviews.
  • Operational Excellence: Improve runbooks, documentation, production readiness reviews, change management, operational standards, and engineering practices that reduce risk and improve system maintainability.
  • Cost & Efficiency: Optimize infrastructure for reliability and performance while maintaining responsible cloud spend and supporting FinOps initiatives.
  • Performance & Capacity: Analyze system performance, resource utilization, latency, throughput, and growth trends; identify bottlenecks and implement scalable solutions before they become production issues.
  • Automation & Toil Reduction: Identify repetitive operational work and replace it with reliable automation using Python, Bash, CI/CD tooling, and platform APIs.
  • CI/CD & Release Reliability: Build and improve deployment pipelines that support safe, repeatable releases through automated testing, validation, progressive delivery, rollback strategies, and deployment observability.
  • Security & Compliance: Apply cloud security and operational controls including least-privilege IAM, encryption, network security, secrets management, patching, auditability, and compliance requirements.
  • Infrastructure as Code: Design, develop, review, and maintain reusable Terraform modules and infrastructure-as-code patterns for AWS environments.
  • Kubernetes Reliability: Operate and improve Kubernetes-based platforms, including EKS clusters, workloads, Helm deployments, autoscaling, upgrades, resource management, and workload resilience.
  • Cloud Platform Engineering: Architect, operate, and optimize AWS services such as EC2, S3, RDS, EKS, Lambda, VPC, IAM, Route 53, and related services.
  • Resilience & Disaster Recovery: Design and validate fault-tolerant architectures, backup strategies, recovery procedures, and disaster recovery capabilities. Conduct reliability testing and failure exercises where appropriate.
  • Cross-Functional Collaboration: Partner with development teams to improve application operability, reliability, instrumentation, deployment patterns, and production readiness.
Qualifications

Bring your passion, here's what's needed:

Required Skills and Qualificationsย 

  • Experience: 5+ years of experience in Site Reliability Engineering, DevOps, platform engineering, cloud infrastructure, or a closely related discipline, including significant production ownership.
  • AWS: Strong hands-on experience designing and operating production workloads in AWS, including networking, IAM, compute, storage, databases, DNS, and managed Kubernetes.
  • Terraform: Advanced proficiency with Terraform, including reusable modules, remote state, dependency management, environment design, code review, and infrastructure lifecycle management.
  • Observability: Experience with metrics, logs, traces, dashboards, alerting, and production telemetry using platforms such as Prometheus, Grafana, CloudWatch, New Relic, ELK/OpenSearch, or similar tools.
  • Reliability Engineering: Practical knowledge of SRE concepts such as SLIs, SLOs, error budgets, capacity planning, fault tolerance, graceful degradation, and reducing operational toil.
  • Kubernetes: Deep understanding of Kubernetes architecture and operations, including EKS, Helm, workload scheduling, networking, storage, autoscaling, upgrades, and troubleshooting.
  • Linux & Systems: Strong Linux systems knowledge and the ability to diagnose issues involving CPU, memory, disk, networking, processes, DNS, and application dependencies.
  • Programming & Automation: Proficiency in Python, Bash, Go, or another general-purpose language used to build operational tooling and automation.
  • Incident Management: Experience troubleshooting complex production incidents and contributing to incident response, root-cause analysis, postmortems, and corrective-action tracking.
  • CI/CD: Experience designing or operating CI/CD systems such as GitHub Actions, Jenkins, GitLab CI, Argo CD, or comparable tooling.
  • Security: Working knowledge of cloud security best practices, including IAM, encryption, secrets management, network segmentation, vulnerability management, and audit controls.
  • Communication: Strong written and verbal communication skills, with the ability to collaborate effectively across engineering and business teams.

Preferred Qualificationsย 

  • AWS certifications such as AWS Certified Solutions Architect - Professional or AWS Certified DevOps Engineer - Professional.
  • Experience with GitOps practices and tools such as Argo CD or Flux.
  • Experience designing or participating in formal on-call rotations and incident management programs.
  • Familiarity with chaos engineering, resilience testing, or game-day exercises.
  • Experience with service meshes, distributed systems, and microservice architectures.
  • Knowledge of database operations for technologies such as Amazon RDS, DynamoDB, and PostgreSQL.
  • Experience with multi-account AWS environments, landing zones, governance, or large-scale cloud platform design.
  • Experience with infrastructure cost optimization, cloud financial management, or FinOps practices.
  • Familiarity with security and compliance frameworks such as SOC 2, ISO 27001, PCI DSS, or similar standards.

What Success Looks Likeย 

  • Production services become more measurable, reliable, and resilient over time.
  • Operational toil and recurring incidents are reduced through engineering and automation.
  • Teams have clear SLOs, useful dashboards, actionable alerts, and well-understood operational ownership.
  • Infrastructure and deployment changes are repeatable, observable, secure, and low risk.
  • Incidents produce meaningful learning and durable improvements rather than recurring fixes.
  • Cloud capacity and cost are proactively managed without compromising reliability.

Why Join Us?ย 

  • Work on modern cloud and reliability engineering challenges in a fast-paced, innovative environment.
  • Help shape reliability standards and engineering practices for mission-critical systems.
  • Collaborate with talented engineers across software, cloud, security, and product disciplines.
  • Competitive salary, comprehensive benefits, and opportunities for professional growth.
  • Flexible remote or hybrid work options.

Be a part of an innovative team shaping the grid of the future through advanced energy intelligence.ย  For more than half a century, Electric Power Engineers (EPE) has partnered with power and energy clients across the globe, providing consulting expertise and energy intelligence software solutions for complex engineering and grid modeling challenges. As leaders in the renewables space, we are focused on building a modern, secure, and resilient grid. ย ย Join us in making an impact on the communities we serve and the environment in which we live. Together we can transform the future of energy. ย 

ย 

How we support you:

  • Comprehensive health and wellness benefits including medical, dental, and vision with 100% premium coverage forย you
  • Generous PTO and paid holidays
  • MyShare Employee Ownership Program
  • Work with industry leaders
  • 401K, up to a 4% match (100% vested from day 1)

Location: This position will be located in City, State

Travel:ย  Occasional travel may be needed (10% or less)

ย 

EPE is an equal opportunity/AA/Disability/Veteran employer. The EEO is the Law poster, and its supplement are available using the following links:ย EEOC is the Law Poster

ย 

Third-Party Recruiting Notification

EPE does not accept unsolicited resumes fromย third-party recruiters. Any unsolicited third-party resumes forwarded by recruiters to EPE via our career page or to any of our managers or employees will be considered public information, may be treated as a direct application from the person identified in the resume, and will not be eligible for placement fee payment to the agency.ย EPE will not pay a fee to a third-party recruiter or agencyย without a previously signed third-party agreementย and has not coordinated their recruiting activity with the appropriate member of the Talent Acquisition team.ย 

Employment Type: FULL_TIME