1

Reliability Engineer Jobs in McKinney, TX (NOW HIRING)

SRE Lead

Plano, TX · On-site

$54.50 - $72.50/hr

Hi , This is Rachael from SidRam Tech, We have an urgent position SRE Lead @ Plano, TX - Onsite Job Title: SRE Lead Location: Plano, TX - Onsite We are seeking an experienced 13 to 18 years of ...

.NET Software Engineer / SRE II

Irving, TX · On-site

$54.75 - $72.75/hr

.NET Software Engineer / SRE II DETAILS Location : Location : Arlington, TX 76014 (hybrid onsite 2-days per week) Position Type : Direct-Hire Hourly / Salary : to $145K (based on experience) + 10% ...

Senior Site Reliability Engineer

Plano, TX · On-site

$54.50 - $72.50/hr

The Senior Site Reliability Engineer acts as an advanced senior individual contributor responsible for designing, implementing, and maturing reliability engineering capabilities. The role focuses on ...

Cloud Service Reliability Engineer

Plano, TX · On-site

$54.50 - $72.50/hr

The Cloud Service Reliability Engineer will be responsible for effective design, execution, and maintenance of systems implemented on premise or in the cloud, primarily focused on identity and access ...

Lead SRE

Plano, TX · On-site

$150 - $200/hr

We are seeking a Delivery SRE leader who will ensure security applications are delivered with strong SDLC discipline and measurable reliability. This role partners closely with Product Owners and ...

Senior Director - Observability | SRE

Coppell, TX · On-site

$52.50 - $70/hr

About the Role The Senior Director - Observability and SRE, is a strategic leader accountable for ensuring the reliability, availability, and performance of the enterprise technology ecosystem. This ...

Senior Site Reliability Engineer

Coppell, TX · On-site

$53 - $70.50/hr

Senior Site Reliability Engineer -- combination of deep operational expertise and hands-on engineering ability. The majority of your time (~70%) will be focused on owning the reliability ...

Lead Site Reliability Engineer

Plano, TX · On-site

$53.25 - $70.75/hr

As a Lead Site Reliability Engineering at JPMorgan Chase within the Chief Technology Office, Identity & Access Management team, you are the non-functional requirement owner and champion for the ...

Lead Site Reliability Engineer

Plano, TX

$54.50 - $72.50/hr

As a Lead Site Reliability Engineering at JPMorgan Chase within the Chief Technology Office, Identity & Access Management team, you are the non-functional requirement owner and champion for the ...

Senior Site Reliability Engineer

Plano, TX

$54.50 - $72.50/hr

As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products. As part of a new SRE team supporting ...

Lead Site Reliability Engineer

Plano, TX · On-site

$53.25 - $70.75/hr

As a Lead Site Reliability Engineering at JPMorgan Chase within the Chief Technology Office, Identity & Access Management team, you are the non-functional requirement owner and champion for the ...

Senior Site Reliability Engineer

Plano, TX · On-site

$54.50 - $72.50/hr

As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products. As part of a new SRE team supporting ...

SRE Lead

Dallas, TX · On-site

$56.50 - $75/hr

Role- SRE Lead Location-Dallas TX Onsite Term : W2 JD: Required Skills & Experience As a Senior SRE Lead, you will lead the implementation, optimization, and maintenance of production systems at the ...

Showing results 41-60

Reliability Engineer information

See McKinney, TX salary details

$56.6K

$109.5K

$130.9K

How much do reliability engineer jobs pay per year?

As of Sep 7, 2026, the average yearly pay for reliability engineer in McKinney, TX is $109,483.00, according to ZipRecruiter salary data. Most workers in this role earn between $95,100.00 and $119,700.00 per year, depending on experience, location, and employer.

What is a reliability engineer?

Reliability Engineers are professionals responsible for ensuring that systems, equipment, or processes function consistently and efficiently over time. They analyze data, identify potential points of failure, and develop maintenance strategies to improve system reliability and minimize downtime. Their work spans various industries, including manufacturing, energy, and technology, and often involves collaborating with design, operations, and maintenance teams. By implementing reliability-centered maintenance and predictive analysis, they help organizations save costs and increase safety.

What does a reliability engineer do?

As a reliability engineer, your duties are to test and evaluate the manufacturing of products and components and ensure that the procedures are efficient and do not lead to abnormally high maintenance or operational costs. Your other responsibilities are to find solutions to product reliability risks. You may manage risk in a supply chain, develop loss prevention strategies, and track the entire lifecycle of product development, from building prototypes to moving a product into full-scale production. You analyze information from department heads and recommend strategies to reduce risk and ensure that the product works reliably.

What are the key skills and qualifications needed to thrive as a reliability engineer, and why are they important?

To thrive as a Reliability Engineer, you need a solid background in engineering principles, failure analysis, and reliability modeling, typically with a degree in engineering or a related field. Familiarity with tools such as FMEA, Root Cause Analysis (RCA), reliability-centered maintenance (RCM) software, and certifications like Certified Reliability Engineer (CRE) are highly valued. Strong problem-solving abilities, attention to detail, and effective communication are crucial soft skills in this role. These skills ensure systems are dependable, downtime is minimized, and organizational performance and safety are optimized.

What are some typical challenges reliability engineers face when implementing preventive maintenance strategies?

Reliability Engineers often encounter challenges such as balancing preventive maintenance schedules with production demands, ensuring buy-in from operations teams, and accurately predicting equipment failures. They must analyze large sets of historical data to identify trends and root causes, which can be complex in facilities with diverse machinery. Collaboration with maintenance, operations, and engineering teams is essential to develop effective strategies that minimize downtime while optimizing resources.

What is the difference between Reliability Engineer vs Maintenance Engineer?

AspectReliability EngineerMaintenance Engineer
CredentialsTypically requires engineering degree, certifications in reliability or asset managementOften requires engineering or technical diploma, certifications in maintenance or equipment repair
Work EnvironmentFocuses on analysis, design, and improvement of systems for reliabilityHands-on maintenance, repair, and troubleshooting of equipment
Industry UsageCommon in manufacturing, energy, aerospace, and industrial sectorsPrevalent in manufacturing, facilities, and industrial plants

Reliability Engineers focus on designing and improving systems to prevent failures, using data analysis and modeling. Maintenance Engineers perform hands-on repairs and upkeep of equipment to ensure operational continuity. While both roles aim to optimize equipment performance, Reliability Engineers work proactively on system reliability, whereas Maintenance Engineers handle reactive and scheduled maintenance tasks.

Are reliability engineers in demand?

Reliability engineers are in high demand across industries such as manufacturing, energy, and aerospace due to their role in improving system performance and reducing downtime. Employers seek professionals with skills in data analysis, failure modes, and maintenance strategies, often requiring certifications like Certified Reliability Engineer (CRE). The job outlook is positive, with steady growth expected as companies prioritize operational efficiency and risk management.

How much do reliability engineers get paid?

Reliability engineers typically earn a median annual salary ranging from $70,000 to $110,000, depending on experience, location, and industry. Senior or specialized reliability engineers with certifications and advanced skills can earn higher salaries, often exceeding $120,000 annually.

What are the most commonly searched types of Reliability Engineer jobs in McKinney, TX?

The most popular types of Reliability Engineer jobs in McKinney, TX are:

What are popular job titles related to Reliability Engineer jobs in McKinney, TX?

For Reliability Engineer jobs in McKinney, TX, the most frequently searched job titles are:

What job categories do people searching Reliability Engineer jobs in McKinney, TX look for?

The top searched job categories for Reliability Engineer jobs in McKinney, TX are:

What cities near McKinney, TX are hiring for Reliability Engineer jobs?

Cities near McKinney, TX with the most Reliability Engineer job openings:

Infographic showing various Reliability Engineer job openings in McKinney, TX as of August 2026, with employment types broken down into 100% Full Time. Highlights an 74% In-person, and 26% Remote job distribution, with an average salary of $109,483 per year, or $52.6 per hour.

Engineering - SRE Platforms - Site Reliability Engineer - Vice President - Dallas

Goldman Sachs, Inc.

Dallas, TX • On-site

$56.50 - $75/hr

Full-time

Re-posted 27 days ago


Goldman Sachs rating

7.8

Company rating: 7.8 out of 10

Based on 28 frontline employees who took The Breakroom Quiz

88th of 175 rated banks


Job description

Site Reliability Engineer - Vice President

Site Reliability Engineering (SRE) is an engineering discipline that combines software and systems engineering to build and run scalable, massively distributed, fault-tolerant systems. At Goldman Sachs, SRE is responsible for improving the availability and reliability of the firm's most critical platform services and ensures they meet the requirements of our internal and external users. It is also responsible for firmwide policies and standards focused on firm's digital resilience. We are looking for engineers who are motivated to collaborate with our businesses to build and run sustainable production systems, which can evolve and adapt to changes in our fast-paced, global business environment.

The SRE team develops and maintains platforms and tools which help other Engineering teams in Goldman Sachs to build and operate reliable and resilient systems. These systems span on-premises datacenters and multiple public cloud environments.   The platforms we offer include central logging, monitoring, agents and alerting and we provide tools to drive adoption and improvements to capacity planning, operational readiness assessments, production incident postmortems, SLIs / SLOs, and deployment automation including canary releases.

The products and services we provide to our internal customers are used by thousands of engineers every day. We believe that reliability is the most important feature of any system, and we are devoted to giving our engineers the platforms and tools they need to build and operate reliable products.

  Role Overview

As a Site Reliability Engineer (SRE) at Goldman Sachs, you will be a pivotal leader in ensuring the availability, reliability, and scalability of the firm's most critical platform applications and services. You will combine deep software and systems engineering expertise to architect, build, and run large-scale, massively distributed, fault-tolerant systems. This role involves providing technical leadership, mentoring senior engineers, and collaborating closely with internal teams and executive stakeholders to build and operate sustainable production systems that can adapt to our dynamic global business environment. You will drive a culture of continuous improvement, championing the adoption of advanced SRE principles and best practices across the organization.

 Responsibilities

  • Strategic Reliability & Performance: Drive the strategic direction for availability, scalability, and performance of mission-critical applications and platform services, ensuring alignment with firm-wide objectives.
  • Architectural Leadership: Lead the design, build, and implementation of highly available, resilient, and scalable infrastructure and application architectures.
  • Advanced Automation & Tooling: Architect and develop sophisticated platforms, tools, and automation solutions to eliminate toil, optimize operational workflows, and enhance deployment processes across the enterprise.
  • Complex Incident Management & Post-Mortem Analysis: Lead critical incident response, conduct in-depth root cause analysis for systemic issues, and implement long-term preventative measures to significantly enhance system stability and resilience.
  • System Design & Capacity Planning: Partner with development teams to embed reliability into application design from inception, provide expert system design consulting, and lead comprehensive capacity planning initiatives for future growth.
  • Observability & Insights: Define and implement advanced monitoring, high volume logging with multi-user query capabilities, and tracing strategies to provide deep, actionable insights into application performance, infrastructure health, and user experience.
  • Technical Vision & Mentorship: Provide technical vision, lead complex technical projects, conduct rigorous code reviews, enforce SDLC best practices, and actively mentor and develop senior and staff-level engineers.
  • Technology Evaluation & Adoption: Stay at the forefront of industry trends and advancements, evaluating and integrating cutting-edge tools and frameworks to significantly improve operational efficiency and reliability.
  • On-Call Leadership: Participate in and lead on-call rotations, providing expert guidance and hands-on support for critical system incidents.
Qualifications
  • Experience: Minimum of 6+ years of hands-on experience in Site Reliability Engineering, with a proven track record in architecting, designing, building, and maintaining highly available, scalable, and fault-tolerant systems at an enterprise level.
  • Technical Proficiency:
    • Exceptional programming skills in one or more major languages such as Java, Python, Go with a focus on building robust, scalable software.
    • Extensive hands-on experience with cloud platforms (e.g., AWS, GCP) and deep expertise in containerization and orchestration technologies (e.g., Docker, Kubernetes).
    • Mastery of Infrastructure as Code (IaC) tools (e.g., Terraform, CloudFormation) and configuration management tools (e.g., Puppet, Chef, Ansible).
    • Advanced proficiency in Prompt Engineering and Retrieval-Augmented Generation (RAG) architectures to automate complex SRE workflows, such as the generation of Infrastructure as Code (IaC), dynamic runbooks, and incident response summaries.
    • Profound understanding of Linux internals, networking, distributed systems, and advanced system performance tuning.
    • Expertise in designing and implementing comprehensive monitoring, alerting, logging and tracing solutions (e.g., Prometheus, Grafana, ELK stack, Datadog, PagerDuty).
    • Deep experience with CI/CD tools and practices (e.g., Jenkins, GitLab, Maven).
    • Strong foundation in databases and distributed systems.
    • Exceptional problem-solving abilities and analytical skills, with a track record of resolving complex technical challenges.
  • Preferred Experience:
    • Experience with Distributed Databases like Elastic Search
    • Experience with working on GCP Big Query
    • Experience with messaging Systems Like Kafka
  • Education: Advanced degree (Bachelor's or Mas ter's or PhD) in Computer Science or a related technical field involving coding and/or systems engineering, or equivalent practical experience.
  • Soft Skills: Superior communication, collaboration, and interpersonal skills, with the ability to influence technical direction, lead cross-functional initiatives, and effectively engage with global teams and executive leadership. Proven ability to work independently, manage multiple complex stakeholders, and drive significant organizational change.

What Goldman Sachs employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Goldman Sachs logo

About Goldman Sachs

Sourced by ZipRecruiter

At Goldman Sachs, we commit our people, capital and ideas to help our clients, shareholders and the communities we serve to grow. Founded in 1869, we are a leading global investment banking, securities and investment management firm. Headquartered in New York, we maintain offices around the world. We believe who you are makes you better at what you do. We're committed to fostering and advancing diversity and inclusion in our own workplace and beyond by ensuring every individual within our firm has a number of opportunities to grow professionally and personally, from our training and development opportunities and firmwide networks to benefits, wellness and personal finance offerings and mindfulness programs.

Industry

Finance and insurance

Company size

10,000+ Employees

Headquarters location

New York, NY, US

Year founded

1869