1

Site Reliability Engineer Jobs in Spring, TX (NOW HIRING)

Site Reliability Engineer

Houston, TX · On-site

$55.25 - $73.50/hr

As a Site Reliability Engineer, you will be responsible for: Operational Excellence & Incident Management - Maintain and monitor production systems for availability, latency, and performance. - Lead ...

Site Reliability Engineer II

Houston, TX · Hybrid

$54.50 - $72.25/hr

Site Reliability Engineer II About PROS: PROS, Inc. is the leading offer management provider to the airline industry, helping airlines deliver seamless retail experiences designed to maximize revenue ...

Senior Site Reliability Engineer NEX

Houston, TX · On-site

$54.50 - $72.25/hr

Required Knowledge, Skills, and Abilities • Three or more years of experience in Site Reliability Engineering, platform engineering, DevOps, cloud engineering, production software engineering, or a ...

Senior Site Reliability Engineer NEX

Houston, TX · On-site

$54.50 - $72.25/hr

Provide technical guidance and coaching on SRE, cloud, Kubernetes, observability, and incident-management practices. Required Knowledge, Skills, and Abilities Three or more years of experience in ...

DevOps & Site Reliability Engineer

Houston, TX

$47.25 - $62.75/hr

... SRE ENGINEER Location: HOUSTON, TX FLSA Class: EXEMPT Responsible to: Directo of Software Engineering Position Summary: DevOps / Site Reliability Engineer to implement and evolve the infrastructure ...

DevOps & Site Reliability Engineer

Houston, TX · On-site

$47.25 - $62.75/hr

... SRE ENGINEER Location: HOUSTON, TX FLSA Class: EXEMPT Responsible to: Directo of Software Engineering Position Summary: DevOps / Site Reliability Engineer to implement and evolve the infrastructure ...

DevOps & Site Reliability Engineer

Houston, TX · On-site

$54.50 - $72.25/hr

... SRE ENGINEER Location: HOUSTON, TX FLSA Class: EXEMPT Responsible to: Directo of Software Engineering Position Summary: DevOps / Site Reliability Engineer to implement and evolve the infrastructure ...

DevOps & Site Reliability Engineer

Houston, TX · On-site

$54.50 - $72.25/hr

... SRE ENGINEER Location: HOUSTON, TX FLSA Class: EXEMPT Responsible to: Directo of Software Engineering Position Summary: DevOps / Site Reliability Engineer to implement and evolve the infrastructure ...

Systems Engineering Manager, SRE & DevOps

Houston, TX · On-site

$47.25 - $62.75/hr

... SRE & DevOps Location : HOUSTON, TX FLSA Class : EXEMPT Responsible to : Director of Software Engineering Position Summary: Manager of Systems Engineering to lead a small, high-impact DevOps/SRE ...

Showing results 21-40

Site Reliability Engineer information

See Spring, TX salary details

$9

$56

$81

How much do site reliability engineer jobs pay per hour?

As of Sep 14, 2026, the average hourly pay for site reliability engineer in Spring, TX is $56.72, according to ZipRecruiter salary data. Most workers in this role earn between $48.75 and $64.81 per hour, depending on experience, location, and employer.

What is a site reliability engineer?

A site reliability engineer specializes in site reliability engineering, or SRE, a specific branch of operations first pioneered by Google. You are responsible for ensuring that when a website decides to scale a particular feature for various users to access, it does not break the underlying software or website functions. This means you need to use analytical problem-solving skills to determine how to make specific features on a new software release work on top of existing source code.

What is a site reliability engineer?

A Site Reliability Engineer (SRE) is a professional who applies software engineering principles to infrastructure and operations problems. Their primary goal is to create scalable and highly reliable software systems, often bridging the gap between development and IT operations. SREs automate tasks, monitor system health, respond to incidents, and work to improve system reliability and performance. They also help define service level objectives (SLOs) and ensure systems meet customer expectations for uptime and availability.

What are the key skills and qualifications needed to thrive as a site reliability engineer?

To thrive as a Site Reliability Engineer, you need a strong background in computer science, systems administration, and software engineering, often supported by a degree in a technical field. Familiarity with cloud platforms (like AWS or GCP), container orchestration (such as Kubernetes), infrastructure as code (Terraform or Ansible), and monitoring tools (Prometheus, Grafana) is typically expected. Strong problem-solving skills, effective communication, and a proactive mindset help SREs excel at incident management and cross-functional collaboration. These skills are crucial for maintaining system reliability, minimizing downtime, and driving continuous improvement in complex technical environments.

What are some of the most common challenges site reliability engineers face when balancing system reliability with rapid software delivery?

Site Reliability Engineers (SREs) often navigate the challenge of maintaining highly reliable systems while supporting fast-paced software releases. This involves managing incidents, automating processes to reduce manual toil, and working closely with development teams to embed reliability into the software development lifecycle. SREs must carefully prioritize their efforts between proactive improvements and urgent, reactive fire-fighting. Effective communication and collaboration with both operations and development teams are crucial to ensuring service uptime without slowing down innovation.

What is the difference between Site Reliability Engineer vs DevOps Engineer?

AspectSite Reliability EngineerDevOps Engineer
CredentialsTypically requires a computer science degree, certifications like AWS, Google Cloud, or KubernetesSimilar credentials, often with cloud certifications and scripting skills
Work EnvironmentFocuses on maintaining and improving system reliability, often in large-scale production environmentsWorks on automation, CI/CD pipelines, and deployment processes across development and operations teams
Industry UsageCommon in tech, cloud services, and large-scale enterprise companiesWidely used in software development, cloud, and IT organizations

Both roles require strong technical skills and cloud knowledge, but SREs focus more on system reliability and uptime, while DevOps engineers emphasize automation and deployment processes. They often collaborate but have distinct primary responsibilities.

Is a site reliability engineer a stressful job?

A site reliability engineer (SRE) role can be stressful due to the responsibility of maintaining system uptime, handling incidents, and ensuring reliability under tight deadlines. The job often requires strong problem-solving skills, familiarity with monitoring tools, and the ability to work in high-pressure situations, but it also offers opportunities for skill development and process improvements.

What are the most commonly searched types of Site Reliability Engineer jobs in Spring, TX?

The most popular types of Site Reliability Engineer jobs in Spring, TX are:

What are popular job titles related to Site Reliability Engineer jobs in Spring, TX?

For Site Reliability Engineer jobs in Spring, TX, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer jobs in Spring, TX look for?

The top searched job categories for Site Reliability Engineer jobs in Spring, TX are:

What cities near Spring, TX are hiring for Site Reliability Engineer jobs?

Cities near Spring, TX with the most Site Reliability Engineer job openings:

Infographic showing various Site Reliability Engineer job openings in Spring, TX as of August 2026, with employment types broken down into 1% As Needed, 83% Full Time, 14% Part Time, and 2% Contract. Highlights an 93% Physical, 3% Hybrid, and 4% Remote job distribution, with an average salary of $117,984 per year, or $56.7 per hour.

Site Reliability Engineer

Houston, TX • On-site

NOV, Inc.
Oil and Gas Extraction • 10K+ employees

$55.25 - $73.50/hr

Full-time

Re-posted 23 days ago


NOV rating

8.1

Company rating: 8.1 out of 10

Based on 72 frontline employees who took The Breakroom Quiz


Job description

As a Site Reliability Engineer, you will be responsible for: Operational Excellence & Incident Management

- Maintain and monitor production systems for availability, latency, and performance.

- Lead incident response efforts, including communication, resolution, and postmortem documentation.

- Design and implement health checks, alerting systems, and automated remediation workflows.

- Drive root cause analysis and implement permanent resolutions for recurring issues.

Observability & Insights

- Set up and maintain full observability stacks (logging, metrics, tracing) using tools like Prometheus, Grafana, Datadog, OpenTelemetry, or ELK.

- Analyze telemetry and logs to identify trends, anomalies, and opportunities for improvement.

- Conduct post-incident reviews and use insights to inform future engineering investments.

Performance & Systems Optimization

- Tune and optimize distributed systems, including AKKA.NET actors, for performance and resource efficiency.

- Work with developers to evolve architecture and improve system throughput, latency, and stability.

- Optimize PostgreSQL performance, queries, and maintenance strategies.

CI/CD & Automation

- Design and maintain modern CI/CD pipelines using GitHub Actions, Azure Pipelines, or GitLab CI.

- Automate deployment, testing, and rollback processes to reduce friction and increase deployment frequency.

- Standardize infrastructure as code practices across environments.

We'd love to talk to you if you have:

- 5+ years of experience in SRE, DevOps, or Infrastructure Engineering roles.

- Expertise in Kubernetes and container orchestration at scale.

- Strong experience with AKKA.NET or similar actor-based frameworks.

- Proficiency with scripting and automation (Bash, PowerShell, Python).

- Experience with observability tools (Phobos,Datadog, Prometheus, Grafana, OpenTelemetry, ELK).

- Hands-on experience with cloud platforms (AWS, Azure, or GCP).

- Strong PostgreSQL knowledge-performance tuning, query optimization, maintenance.

- Proven ability to lead incident management and drive postmortem processes.

- A builder's mindset with high standards for operational excellence and technical ownership.

Preferred Tools & Ecosystem Experience

- CI/CD: GitHub Actions, Azure Pipelines, GitLab CI

- Infrastructure: Kubernetes, Docker, Terraform

- Monitoring: Phobos (AKKA.NET), Datadog, Prometheus

- Source Control: GitHub, GitLab, Azure DevOps

- Programming: C#, Python, Bash, PowerShell

Every day, the oil and gas industry's best minds put more than 150 years of experience to work to help our customers achieve lasting success.
We Power the Industry that Powers the World
Throughout every region in the world and across every area of drilling and production, our family of companies has provided the technical expertise, advanced equipment, and operational support necessary for success-now and in the future.
Global Family
We are a global family of thousands of individuals, working as one team to create a lasting impact for ourselves, our customers, and the communities where we live and work.
Purposeful Innovation
Through purposeful business innovation, product creation, and service delivery, we are driven to power the industry that powers the world better.
Service Above All
This drives us to anticipate our customers' needs and work with them to deliver the finest products and services on time and on budget.
Corporate
Our family of companies is supported by our global Corporate teams, providing expert knowledge from functions including Human Resources, Information Technology, Compliance, Finance, QHSE, Marketing and Legal centers of expertise.  We are structured to provide guidance and service above all to all our business operations.

What NOV employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


NOV logo

About NOV

Sourced by ZipRecruiter

Throughout every region in the world and across every area of drilling and production, our family of companies has provided the technical expertise, advanced equipment and operational support necessary for success. We have the people, capabilities and vision to serve the needs of a challenging and evolving industry. One the world can’t live without. We are a global family of thousands of individuals, working as one team to create lasting impact for ourselves, our customers and the communities where we live and work. We take responsibility for each other and our company’s future, knowing that personal ownership leads to broader success. We believe in purposeful innovation because we see what others do not and we act. Through business innovation, product creation and service delivery, we are driven to power the industry that powers the world better.

Industry

Oil and gas extraction

Company size

10,000+ Employees

Headquarters location

Houston, TX, US

Year founded

1841