1

Software Reliability Engineer Jobs in Rockwall, TX

Site Reliability Engineer

Plano, TX · On-site

$54.50 - $72.50/hr

This role requires a strong foundation in software development, infrastructure automation, and reliability engineering. You will be responsible for designing, implementing, and maintaining high ...

Senior Site Reliability Engineer

Plano, TX

$54.50 - $72.50/hr

As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable ... You will combine software engineering and production operations to deploy, run, monitor, improve ...

Meet the Team The SRE Fleet team is responsible for maintaining the stability, scalability, and ... Experience leveraging AI-assisted development tools to improve software development, automation ...

Meet the Team The SRE Fleet team is responsible for maintaining the stability, scalability, and ... Experience leveraging AI-assisted development tools to improve software development, automation ...

next page

Showing results 1-20

Software Reliability Engineer information

See Rockwall, TX salary details

$37

$62

$82

How much do software reliability engineer jobs pay per hour?

As of Jul 26, 2026, the average hourly pay for software reliability engineer in Rockwall, TX is $62.36, according to ZipRecruiter salary data. Most workers in this role earn between $55.00 and $69.28 per hour, depending on experience, location, and employer.

What engineer makes $500,000 a year?

Software Reliability Engineers with extensive experience, advanced skills in automation and testing, and leadership roles can earn salaries approaching or exceeding $500,000 annually, especially in high-cost-of-living areas or large tech companies. Such compensation often includes bonuses, stock options, and other incentives.

What are the key skills and qualifications needed to thrive as a Software Reliability Engineer, and why are they important?

To thrive as a Software Reliability Engineer, you need a strong background in software development, system architecture, and incident response, often supported by a degree in computer science or related field. Familiarity with monitoring tools (like Prometheus), cloud platforms (AWS, GCP), automation frameworks, and certifications such as AWS Certified DevOps Engineer are highly valuable. Excellent problem-solving, collaboration, and communication skills help you coordinate effectively during high-pressure situations and with cross-functional teams. These abilities are crucial for maintaining system uptime, quickly resolving outages, and ensuring the overall reliability of critical software services.

What are Software Reliability Engineers?

Software Reliability Engineers (SREs) are IT professionals who focus on ensuring that software systems are reliable, scalable, and maintain high availability. They work at the intersection of software development and IT operations, often automating processes, monitoring system performance, and responding to incidents. SREs use engineering principles to solve operational problems, aiming to reduce downtime and improve user experience. Their responsibilities can include building tools, managing infrastructure, and collaborating with development teams to implement best practices for reliability.

How does a Software Reliability Engineer typically interact with development and operations teams to improve system stability?

Software Reliability Engineers (SREs) work closely with both development and operations teams to ensure that systems are reliable, scalable, and maintainable. They often participate in design reviews, provide input on architectural decisions, and help define service-level objectives. SREs also collaborate with developers to automate deployment processes and create monitoring solutions, and they partner with operations staff to manage incident response and root cause analysis. This collaborative environment enables them to proactively identify potential issues and drive cross-functional improvements.

How much do SRE get paid?

Software Reliability Engineers (SREs) typically earn between $90,000 and $150,000 annually, depending on experience, location, and company size. Senior SREs with specialized skills in automation, monitoring, and cloud platforms can earn higher salaries, often exceeding $160,000.

Will AI replace SRE jobs?

AI is unlikely to fully replace Software Reliability Engineers (SREs), as their role involves complex problem-solving, system design, and incident management that require human judgment. Instead, AI tools are increasingly used to automate routine tasks, enhance monitoring, and improve system reliability, allowing SREs to focus on more strategic issues. SREs with skills in automation, scripting, and cloud environments will continue to be valuable in managing and optimizing complex systems.

What is the difference between Software Reliability Engineer vs Software Test Engineer?

AspectSoftware Reliability EngineerSoftware Test Engineer
Primary FocusEnsuring software reliability, stability, and performance over timeDesigning and executing tests to identify bugs and verify functionality
Skills & CertificationsKnowledge of reliability engineering, scripting, monitoring toolsTesting methodologies, automation tools, scripting
Work EnvironmentCollaborates with development and operations teams, often in DevOpsWorks primarily in QA/testing teams, often in dedicated testing phases
Industry UsageCommon in software companies focusing on product stabilityWidely used in software development and QA departments

The main difference is that Software Reliability Engineers focus on maintaining long-term software stability and performance, while Software Test Engineers concentrate on identifying bugs through testing. Both roles require technical skills and often collaborate, but their core objectives differ: reliability versus defect detection.

What does a software reliability engineer do?

A software reliability engineer focuses on ensuring software systems are dependable and perform consistently by analyzing failure data, developing testing strategies, and implementing automation tools. They often work with monitoring systems, perform root cause analysis, and collaborate with development teams to improve software quality and stability.
What are popular job titles related to Software Reliability Engineer jobs in Rockwall, TX? For Software Reliability Engineer jobs in Rockwall, TX, the most frequently searched job titles are:
Infographic showing various Software Reliability Engineer job openings in Rockwall, TX as of July 2026, with employment types broken down into 94% Full Time, 3% Part Time, and 3% Contract. Highlights an 89% Physical, 4% Hybrid, and 7% Remote job distribution, with an average salary of $129,700 per year, or $62.4 per hour.
Site Reliability Engineer

Site Reliability Engineer

MM International

Plano, TX • On-site

$54.50 - $72.50/hr

Contractor

Posted 24 days ago


Job description

Job Title: Site Reliability Engineer

Hybrid 3 times a week in Iselin, NJ
OR
Hybrid 3 times a week in PLANO, TX
 

Interview Process:
Virtual- 30 min round

Onsite- 1-2 hours

Needs:
Openshift
Kubernetes

Development Experience(Java, Python, Golang)
SRE Skills
Nice to Haves:
Baremetal
Cloud
Job Description:
Client Job Description:

We are looking for a highly skilled Site Reliability and operations Engineer (SRE) with extensive experience in Kubernetes-based distributed caching and compute grid solutions. This role requires a strong foundation in software development, infrastructure automation, and reliability engineering. You will be responsible for designing, implementing, and maintaining high-performance distributed systems, ensuring reliability, scalability, and efficiency.

Development & Implementation:

• Design, develop, and optimize distributed caching and compute grid solutions on Kubernetes/OpenShift

• Understanding of microservices and containerized workloads using Kubernetes, Docker, and Helm.

• Implement high-throughput compute grid solutions using IBM Spectrum Symphony, Tibco Grid Server or similar technologies.

• Optimize application performance by leveraging parallel compute strategies, load balancing, and efficient data distribution.

Site Reliability Engineering (SRE):

• Ensure high availability, scalability, and reliability of distributed systems.

• Implement observability, logging, and monitoring using tools like Prometheus, Grafana, ELK, or OpenTelemetry.

• Automate infrastructure provisioning and deployments using Ansible, and Helm Charts.

• Understanding of CI/CD pipelines for seamless software deployment.

• Troubleshoot and resolve incidents related to platform, infrastructure and distributed compute platforms, ensuring minimal downtime.

Required Skills & Qualifications:

• Strong experience in Kubernetes (OpenShift and on-prem/cloud clusters).•

• Understanding of programming languages like Java, Go, or Python. – this will be the difference maker of the L4 vs L5

• Experience with containerization technologies (Docker, Helm, etc.).

• Strong knowledge of CI/CD pipelines (Jenkins, ArgoCD, GitHub Actions).

• Hands-on experience with observability tools (Prometheus, Grafana, Loki, Jaeger).

• Understanding of networking, service meshes (Istio/Linkerd), and security best practices in Kubernetes.

• Experience with multi-cluster and hybrid cloud Kubernetes deployments.