1

Reliability Engineer Manager Jobs in Texas (NOW HIRING)

Reliability Engineer

Houston, TX · On-site

$97K - $123K/yr

Plan and implement reliability engineering strategy, including implementation practices and ... Prepare technical documentation and presentations for peers, management, and external customers and ...

Reliability Engineer

Ennis, TX · On-site

$94K - $119K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Role: Reliability Engineer *This is a Safety Sensitive position. * Job Summary: The Reliability ... Troubleshooting equipment, Design for Reliability, select and managing qualified contractors to ...

Reliability Engineer

Ennis, TX · On-site

$94K - $119K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Role: Reliability Engineer *This is a Safety Sensitive position. * Job Summary: The Reliability ... Troubleshooting equipment, Design for Reliability, select and managing qualified contractors to ...

Site Reliability Engineer (SRE)

Austin, TX · On-site

$56.50 - $75/hr

Austin, TX Job Type: Full Time Technical Skills: * 6+ years of professional engineering experience developing, managing, or supporting distributed systems * 4+ SRE experience managing multi-cloud ...

Reliability Engineer

Houston, TX · On-site

$91K - $114K/yr

Reliability Engineer Department: Manufacturing Job Status: Full Time FLSA Status: Salary, Exempt ... Quality Manager Location: Houston, TX (Channelview / East Houston) Amount of Travel Required: 30 ...

Reliability Engineer

Houston, TX · On-site

$91K - $114K/yr

Reliability Engineer Department: Manufacturing Job Status: Full Time FLSA Status: Salary, Exempt ... Quality Manager Location: Houston, TX (Channelview / East Houston) Amount of Travel Required: 30 ...

Site Reliability Engineer (SRE)

Austin, TX · On-site

$56.50 - $75/hr

Site Reliability Engineer (SRE) Location: Austin, TX Job Type: Full Time Job Summary - Seasoned ... Highly skilled in managing production failures, conducting root cause analysis, and driving ...

Reliability Engineer

Dallas, TX

$101K - $127K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

Premier PV is seeking a Reliability Engineer to support the development, validation, and continuous ... Ability to manage multiple priorities in a fast-paced manufacturing environment. PREFERRED ...

Reliability Engineer

Houston, TX · On-site

$91K - $114K/yr

  • Medical

  • Life

  • Retirement

  • PTO

... process, and management systems, as well as assist the sites in the development of these ... reliability engineer position. Governance & Standards • Drive continuous improvement in the ...

Reliability Engineer

Houston, TX · On-site

$97K - $123K/yr

Plan and implement reliability engineering strategy, including implementation practices and ... Prepare technical documentation and presentations for peers, management, and external customers and ...

Reliability Engineer

Houston, TX

$91K - $114K/yr

  • Medical

  • Life

  • Retirement

  • PTO

General knowledge of project management, reliability engineering, and root cause analysis. Must demonstrate a working knowledge of applicable engineering and inspection codes/standards and their ...

Site Reliability Engineer

Houston, TX · On-site

$54.50 - $72.25/hr

With guidance from the SRE Manager, evaluate technology strategy, solutions, resources, process, and compliance Minimum Qualifications * Bachelor's degree in computer science, computer engineering or ...

Site Reliability Engineer (SRE)

Decatur, TX · On-site

$51 - $67.75/hr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

Join Our Team as a Site Reliability Engineer (SRE)! About Us At Energy Worldnet, Inc. (EWN), we ... management) Experience with distributed systems, API-driven platforms, and service-oriented ...

Site Reliability Engineer (SRE)

Decatur, TX

$51 - $67.75/hr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

Join Our Team as a Site Reliability Engineer (SRE)! About Us At Energy Worldnet, Inc. (EWN), we ... management) Experience with distributed systems, API-driven platforms, and service-oriented ...

Site Reliability Engineer

Houston, TX · On-site

$52.75 - $70.25/hr

With guidance from the SRE Manager, evaluate technology strategy, solutions, resources, process, and compliance Minimum Qualifications * Bachelor's degree in computer science, computer engineering or ...

Reliability Engineer

Sherman, TX · On-site

$88K - $111K/yr

Present reliability test results and recommendations to management, customers, and cross-functional teams Education & Experience * B.S. or M.S. in Physics, Engineering, Materials Science, Statistics ...

Reliability Engineer

Sherman, TX · On-site

$90 - $120/hr

Experience planning, executing, and analyzing engineering experiments Skills * Reliability analysis ... Ability to manage multiple projects simultaneously and meet aggressive deadlines * Strong technical ...

Asset Reliability Engineer

Houston, TX · On-site

$97K - $123K/yr

Job Summary The Asset Reliability Engineer supports the reliability, performance, and longterm ... Exposure to fleetlevel asset management or performance analytics, preferred * Strong analytical ...

Showing results 21-40

Reliability Engineer Manager information

See Texas salary details

$50.9K

$108.1K

$129.8K

How much do reliability engineer manager jobs pay per year?

As of Aug 18, 2026, the average yearly pay for reliability engineer manager in Texas is $108,075.00, according to ZipRecruiter salary data. Most workers in this role earn between $94,500.00 and $118,300.00 per year, depending on experience, location, and employer.

What does a reliability engineer manager do?

A Reliability Engineer Manager oversees teams responsible for improving the reliability and performance of systems, machinery, or processes within an organization. They develop maintenance strategies, lead root cause analyses of failures, and implement best practices to minimize downtime and costs. Additionally, they collaborate with other departments to ensure that reliability goals align with business objectives and compliance standards. Their role is crucial in industries such as manufacturing, energy, and technology, where system uptime and safety are critical.

What are the key skills and qualifications needed to thrive as a reliability engineer manager?

To thrive as a Reliability Engineer Manager, you need a strong background in engineering principles, reliability analysis, and maintenance strategies, typically supported by a degree in engineering and experience in reliability roles. Familiarity with reliability-centered maintenance (RCM), failure mode and effects analysis (FMEA), and asset management software such as SAP or Maximo is common, along with certifications like Certified Reliability Engineer (CRE). Leadership, problem-solving, and effective communication are vital soft skills for managing teams and driving cross-functional initiatives. These competencies are crucial for minimizing downtime, optimizing equipment performance, and ensuring long-term operational efficiency.

What are some common challenges reliability engineer managers face when balancing long-term reliability improvements with immediate operational demands?

Reliability Engineer Managers often need to prioritize urgent maintenance issues while also driving long-term reliability initiatives. Balancing these competing demands can be challenging, as immediate equipment failures may require quick fixes that temporarily interrupt ongoing improvement projects. Effective managers work closely with operations, maintenance, and engineering teams to communicate priorities, allocate resources, and implement sustainable solutions that address root causes rather than just symptoms. This role typically involves using data-driven decision-making and fostering a culture of proactive maintenance and continuous improvement.

What is the difference between Reliability Engineer Manager vs Reliability Engineer?

AspectReliability EngineerReliability Engineer Manager
Required CredentialsBachelor's in Engineering or related field; certifications like CRC, CRESame as Reliability Engineer, plus leadership experience
Work EnvironmentDesign, analyze, and improve system reliability; often in teamsOversees Reliability Engineers; manages projects and teams
Employer & Industry UsageManufacturing, aerospace, energy, automotiveSame industries, with added managerial responsibilities
Common Search & ComparisonFocuses on technical skills and hands-on reliability tasksFocuses on leadership, team management, and strategic planning

The main difference between a Reliability Engineer and a Reliability Engineer Manager lies in their responsibilities. The Reliability Engineer focuses on technical analysis and system improvements, while the Reliability Engineer Manager oversees teams, manages projects, and develops strategies to enhance reliability across the organization.

What are the most commonly searched types of Reliability Engineer jobs in Texas?

The most popular types of Reliability Engineer jobs in Texas are:

What job categories do people searching Reliability Engineer Manager jobs in Texas look for?

The top searched job categories for Reliability Engineer Manager jobs in Texas are:

What cities in Texas are hiring for Reliability Engineer Manager jobs?

Cities in Texas with the most Reliability Engineer Manager job openings:

Infographic showing various Reliability Engineer Manager job openings in Texas as of August 2026, with employment types broken down into 88% Full Time, and 12% Contract. Highlights an 88% In-person, 6% Hybrid, and 6% Remote job distribution, with an average salary of $108,075 per year, or $52 per hour.

Senior Manager, Site Reliability Engineering - Paylo Platform

PDI Technologies

Dallas, TX

Full-time

Posted 4 days ago


PDI Technologies rating

7.8

Company rating: 7.8 out of 10

Based on 5 frontline employees who took The Breakroom Quiz

137th of 245 rated software companies


Job description

At PDI Technologies, we empower some of the world's leading convenience retail and petroleum brands with cutting-edge technology solutions that drive growth and operational efficiency. By “Connecting Convenience” across the globe, we empower businesses to increase productivity, make more informed decisions, and engage faster with customers through loyalty programs, shopper insights, and unmatched real-time market intelligence via mobile applications, such as GasBuddy.  We’re a global team committed to excellence, collaboration, and driving real impact. Explore our opportunities and become part of a company that values diversity, integrity, and growth.

Role Overview

PDI Technologies is looking for a Senior Manager, Site Reliability Engineering to lead the SRE organization supporting Paylo, PDI’s payments, loyalty, and fuel-pricing product suite. This role owns the reliability, infrastructure, and operational strategy for a portfolio of high-traffic, customer- and partner-facing platforms that power payment transactions, fuel pricing, loyalty and rewards, and offer/coupon redemption for convenience retail and fuel customers around the world. 

This is a hands-on, leadership-first role. You will manage a team of three SRE Managers/Leads who together lead approximately 20 engineers, while staying technically engaged yourself — reviewing architecture, unblocking hard infrastructure problems, and setting the technical bar across the organization. You will bring strong, current, hands-on expertise across AWS, Azure, Kubernetes, Helm, Argo CD, Terraform/OpenTofu, Jenkins, and Datadog, and you will be a strong, visible people leader who can coach managers and represent SRE to senior engineering and business stakeholders.

Key Responsibilities
  • Directly manage and develop 3 SRE Managers/Leads and own the overall health, growth, and performance of an ~20-person SRE organization supporting the Paylo product suite.

  • Set the vision, priorities, and operating cadence for the SRE function; translate business and product priorities into a reliability roadmap your managers can execute against.

  • Build a strong bench by hiring, coaching, and developing managers and senior engineers while creating clear career paths and succession plans.

  • Foster a blameless, learning-oriented culture around incidents, on-call, and operational excellence.

  • Partner closely with engineering directors, product managers, and business stakeholders across the Paylo organization to align reliability investments with business risk and customer impact.

  • Stay technically engaged day to day by participating in architecture and design reviews, troubleshooting complex production issues, and directly contributing to infrastructure-as-code, Kubernetes manifests/Helm charts, and CI/CD pipelines when needed.

  • Set and enforce engineering standards for multi-cloud infrastructure across AWS and Azure and for container orchestration on Kubernetes at scale.

  • Own adoption and standards for GitOps-based continuous delivery using Argo CD/Argo Workflows, including deployment strategy, rollout policy, and multi-cluster promotion.

  • Own the Infrastructure-as-Code strategy across teams (Terraform, OpenTofu), including module standards, state management, drift detection, and remediation.

  • Own CI/CD pipeline architecture and standards built on Jenkins, driving build/deploy automation, pipeline reliability, and progressive delivery practices such as blue-green/canary deployments and automated rollback.

  • Evaluate and guide adoption of new infrastructure tooling and patterns as the platform evolves across AWS and Azure.

  • Own the observability strategy across all supported products, with deep, hands-on expertise in Datadog (APM, infrastructure monitoring, log management, dashboards, and alerting) as the standard platform for metrics, tracing, and alerting.

  • Define and drive adoption of SLIs/SLOs, error budgets, and reliability KPIs across the organization, holding managers and teams accountable to them.

  • Own the incident management program end to end, including on-call structure, escalation paths, severity definitions, postmortems, and follow-through on remediation actions.

  • Drive root-cause analysis and long-term reliability investments that reduce Sev1/Sev2 frequency and recurrence.

  • Ensure appropriate resilience, disaster recovery, and capacity planning practices are in place given the sensitivity of payment- and transaction-related systems.

  • Partner with Security and Compliance to maintain awareness of PCI DSS and related compliance requirements and ensure the SRE organization supports audit and compliance readiness.

  • Track and report cost, capacity, and operational KPIs to senior leadership.

Required Qualifications
  • 8+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure/Platform Engineering, including 4+ years in a people-leadership role. 

  • Proven experience managing managers — you have directly led team leads/managers, not just individual contributors, and are comfortable operating at the scale of ~20 total reports. 

  • Strong, hands-on expertise across AWS and Azure — you can architect, troubleshoot, and operate multi-cloud infrastructure yourself, not just direct others to do so. 

  • Strong, hands-on expertise with Kubernetes and Helm — cluster operations, troubleshooting at scale, and chart design/maintenance. 

  • Strong, hands-on expertise with Argo CD/Argo Workflows for GitOps-based continuous delivery. 

  • Strong, hands-on expertise with Infrastructure as Code (Terraform, OpenTofu), including module design and state management. 

  • Strong, hands-on expertise with Jenkins for CI/CD pipeline design, administration, and automation. 

  • Strong, hands-on expertise with Datadog (or equivalent enterprise observability platform), including designing monitoring/alerting strategy, dashboards, and APM/tracing at scale. 

  • Demonstrated track record of driving incident management, on-call, and postmortem programs for high-traffic, customer-facing systems. 

  • Excellent communication and stakeholder-management skills; able to represent SRE to engineering leadership and business partners with equal credibility. 

  • A strong, visible leadership style — someone who sets clear direction, holds teams accountable, and builds trust across the organization. 

  • Applicants must be legally authorized to work in the United States without the need for employer sponsorship, now or in the future. PDI Technologies is unable to offer visa sponsorship for this role.
Preferred Qualifications
  • Experience supporting payments, fuel/retail, or loyalty platforms, or other systems with PCI DSS or similar compliance obligations. 

  • Relevant certifications such as CKA/CKAD, AWS Certified Solutions Architect, Microsoft Certified: Azure Solutions Architect, or HashiCorp Terraform Associate. 

  • Experience with messaging systems (Kafka/SQS/SNS), PagerDuty (or similar), and multi-region/multi-AZ resilience patterns. 

  • Prior experience consolidating or standardizing SRE and DevOps practices across multiple product lines or recently-integrated/acquired teams. 

  • Experience partnering with product and business stakeholders to translate reliability investments into business outcomes. 

What Success Looks Like
  • A stable, well-led SRE organization with clear ownership, career paths, and low regrettable attrition among your managers and their teams. 

  • Consistent, Datadog-driven observability and SLOs in place across the organization, with measurable reduction in Sev1/Sev2 incidents and mean time to detect/resolve. 

  • Modern, standardized infrastructure practices — GitOps delivery via Argo, IaC via Terraform/OpenTofu, and reliable CI/CD via Jenkins — adopted consistently across teams and clouds. 

  • A mature, blameless incident-management culture with strong postmortem follow-through. 

  • Strong cross-functional trust with engineering, product, and security/compliance stakeholders. 

PDI is committed to offering a well-rounded benefits program, designed to support and care for you, and your family throughout your life and career.  This includes a competitive salary, market-competitive benefits, and a quarterly perks program. We encourage a good work-life balance with ample time off [time away] and, where appropriate, hybrid working arrangements.  Employees have access to continuous learning, professional certifications, and leadership development opportunities. Our global culture fosters diversity, inclusion, and values authenticity, trust, curiosity, and diversity of thought, ensuring a supportive environment for all.

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.


What PDI Technologies employees say

Hours and flexibility

Workplace

Get the full story on Breakroom