1

Network Reliability Engineer Jobs in Arlington, TX

Senior Site Reliability Engineer

Plano, TX · On-site

$54.50 - $72.50/hr

Position Summary: The Senior Site Reliability Engineer acts as an advanced senior individual ... network observability, GenAI platform health monitoring, and enterprise dashboard automation.

Lead SRE

Plano, TX · On-site

$150 - $200/hr

... networking, messaging, automation (CloudFormation, Terraform), and data services. * Build a culture ... , and Security. Required Qualifications, Capabilities, and Skills * Formal training or ...

Senior Site Reliability Engineer

Coppell, TX · On-site

$53 - $70.50/hr

Senior Site Reliability Engineer -- combination of deep operational expertise and hands-on ... Strong understanding of networking, Linux/Windows administration, distributed systems, and cloud ...

Senior Site Reliability Engineer

Plano, TX · On-site

$54.50 - $72.50/hr

Position Summary: The Senior Site Reliability Engineer acts as an advanced senior individual ... network observability, GenAI platform health monitoring, and enterprise dashboard automation.

AI with SRE

Irving, TX · On-site

$54.75 - $72.75/hr

Hi Role : SRE AI Engineer Location : Austin TX We are currently seeking a highly skilled SRE hands ... network etc.,) Telemetry data collection using Dynatrace APM, SolarWinds, CISCO Switches, F5 ...

SRE Consultant

Dallas, TX · On-site

$56.50 - $75/hr

... Networking, or Security) Willingness to work onsite and participate in a 24/7 on-call rotation as ... , DevOps, or Performance Engineering are a plus

Senior Site Reliability Engineer

Plano, TX · On-site

$54.50 - $72.50/hr

As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable ... networking, DNS, and distributed systems * Experience with observability platforms such as Splunk ...

Lead Site Reliability Engineer

Plano, TX · On-site

$53.25 - $70.75/hr

As a Lead Site Reliability Engineering at JPMorgan Chase within the Chief Technology Office ... Strong troubleshooting skills across common networking technologies and issues * Proficiency in at ...

Lead Site Reliability Engineer

Plano, TX

$54.50 - $72.50/hr

As a Lead Site Reliability Engineering at JPMorgan Chase within the Chief Technology Office ... Strong troubleshooting skills across common networking technologies and issues * Proficiency in at ...

Senior Site Reliability Engineer

Plano, TX

$54.50 - $72.50/hr

As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable ... networking, DNS, and distributed systems * Experience with observability platforms such as Splunk ...

Senior Site Reliability Engineer

Plano, TX · On-site

$54.50 - $72.50/hr

Design and implement advanced reliability patterns for Azure landing zones, private networking, DNS ... Mentor SRE engineers and raise the technical bar for automation, troubleshooting, documentation ...

Lead Site Reliability Engineer

Plano, TX · On-site

$53.25 - $70.75/hr

As a Lead Site Reliability Engineering at JPMorgan Chase within the Chief Technology Office ... Strong troubleshooting skills across common networking technologies and issues * Proficiency in at ...

Lead Site Reliability Engineer

Plano, TX · On-site

$53.25 - $70.75/hr

As a Lead Site Reliability Engineering at JPMorgan Chase within the Chief Technology Office ... Strong troubleshooting skills across common networking technologies and issues * Proficiency in at ...

Lead SRE

Plano, TX · On-site

$53.25 - $70.75/hr

... networking, messaging, automation (CloudFormation, Terraform), and data services. * Build a culture ... , and Security. Required Qualifications, Capabilities, and Skills * Formal training or ...

Lead SRE

Plano, TX · On-site

$200 - $250/hr

... networking, messaging, automation (CloudFormation, Terraform), and data services. * Build a culture ... , and Security. Required Qualifications, Capabilities, and Skills * Formal training or ...

Lead Site Reliability Engineer

Plano, TX · On-site

$53.25 - $70.75/hr

As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Technology team , you ... Experience with troubleshooting common networking technologies and issues * Advanced knowledge of ...

Showing results 21-40

Network Reliability Engineer information

See Arlington, TX salary details

$54.9K

$106.2K

$126.9K

How much do network reliability engineer jobs pay per year?

As of Sep 7, 2026, the average yearly pay for network reliability engineer in Arlington, TX is $106,169.00, according to ZipRecruiter salary data. Most workers in this role earn between $92,200.00 and $116,100.00 per year, depending on experience, location, and employer.

What is a network reliability engineer?

A Network Reliability Engineer (NRE) is an IT professional responsible for ensuring the reliability, performance, and scalability of network systems. They combine skills in networking, software engineering, and automation to proactively detect and resolve potential network issues before they affect users. NREs often design and implement monitoring tools, automate network management tasks, and work to improve the overall stability of network infrastructure. Their goal is to minimize downtime and ensure seamless connectivity across an organization’s network.

What are the key skills and qualifications needed to thrive as a network reliability engineer, and why are they important?

To thrive as a Network Reliability Engineer, you need a strong background in computer networking, network protocols, troubleshooting, and often a degree in computer science or a related field. Familiarity with tools like Wireshark, Nagios, Cisco IOS, and certifications such as CCNA or CCNP are commonly required. Analytical thinking, proactive problem-solving, and effective communication are standout soft skills in this role. These skills are crucial to maintaining reliable network operations, minimizing downtime, and ensuring seamless communication across organizational systems.

What are some common challenges faced by network reliability engineers, and how are they typically addressed?

Network Reliability Engineers often encounter challenges such as diagnosing intermittent connectivity issues, managing network upgrades with minimal downtime, and maintaining high availability during peak traffic. These are typically addressed by leveraging robust monitoring tools, implementing automation for routine tasks, and collaborating closely with software engineers, network administrators, and incident response teams. Staying current with evolving network technologies and best practices is essential for effectively identifying and resolving problems before they impact users.

What is the difference between Network Reliability Engineer vs Network Operations Center (NOC) Technician?

AspectNetwork Reliability EngineerNetwork Operations Center (NOC) Technician
CertificationsCCNA, CCNP, Network+CCNA, Network+
Work EnvironmentDesign, analyze, and improve network infrastructureMonitor, troubleshoot, and maintain networks in real-time
Employer & Industry UsageTelecom, large enterprises, cloud providersISPs, data centers, enterprise networks
Common Search & ComparisonFocus on network reliability and designFocus on network monitoring and incident response

The main difference is that Network Reliability Engineers focus on designing and improving network systems to ensure long-term reliability, while NOC Technicians monitor and troubleshoot networks in real-time to resolve issues quickly. Both roles require relevant certifications and are essential in maintaining network performance, but they serve different functions within network management.

What are popular job titles related to Network Reliability Engineer jobs in Arlington, TX?

For Network Reliability Engineer jobs in Arlington, TX, the most frequently searched job titles are:

What job categories do people searching Network Reliability Engineer jobs in Arlington, TX look for?

The top searched job categories for Network Reliability Engineer jobs in Arlington, TX are:

What cities near Arlington, TX are hiring for Network Reliability Engineer jobs?

Cities near Arlington, TX with the most Network Reliability Engineer job openings:

Infographic showing various Network Reliability Engineer job openings in Arlington, TX as of August 2026, with employment types broken down into 1% As Needed, 81% Full Time, 12% Part Time, and 6% Contract. Highlights an 93% Physical, 2% Hybrid, and 5% Remote job distribution, with an average salary of $106,169 per year, or $51 per hour.

Site Reliability Engineer II (SRE II) - Data & Intelligence

Noblesoft Technologies

Dallas, TX • On-site

$56.50 - $75/hr

Contractor

Posted 4 days ago


Job description

Role : Site Reliability Engineer II (SRE II) – Data & Intelligence

Location: Dallas, TX / Overland Park, KS / Atlanta, GA / Bellevue, WA (Onsite)

Job Summary

The Site Reliability Engineer II (SRE II) is responsible for ensuring the reliability, scalability, performance, security, and operational excellence of data platforms, analytics systems, AI/ML services, and business intelligence applications. This role combines software engineering, systems engineering, automation, and operational expertise to build resilient and highly available data services while driving continuous improvement through observability, automation, and reliability engineering practices.

The ideal candidate is passionate about large-scale distributed systems, cloud-native technologies, data platforms, and operational excellence. They partner closely with Data Engineering, Data Science, Analytics, Platform Engineering, and Product teams to maintain and improve critical business services.

Key Responsibilities

Reliability & Operations

•       Ensure the availability, performance, scalability, and reliability of data and intelligence platforms.

•       Manage production environments supporting data ingestion, processing, transformation, storage, analytics, and AI/ML workloads.

•       Participate in on-call rotations and incident response activities.

•       Lead troubleshooting efforts for complex production issues and drive root cause analysis (RCA).

•       Develop and implement service level indicators (SLIs), service level objectives (SLOs), and error budgets.

Automation & Engineering

•       Design and develop automation to improve operational efficiency and system reliability.

•       Build self-healing solutions and automate routine operational tasks.

•       Create tools and scripts to monitor, deploy, and manage large-scale distributed systems.

•       Improve deployment processes through CI/CD pipelines and Infrastructure as Code (IaC).

Observability & Monitoring

•       Design and maintain monitoring, logging, tracing, and alerting solutions.

•       Create dashboards and actionable alerts to proactively identify service degradation.

•       Analyze system performance metrics and recommend optimization opportunities.

•       Drive observability standards across data services and platforms.

Platform & Infrastructure Management

•       Support cloud-based infrastructure and platform services across Azure, AWS, or GCP environments.

•       Optimize compute, storage, networking, and data platform resources.

•       Work with containerized and Kubernetes-based workloads.

•       Ensure high availability and disaster recovery capabilities are implemented and tested.

Data Platform Reliability

•       Support modern data ecosystems including data lakes, warehouses, streaming platforms, and analytics environments.

•       Monitor ETL/ELT pipelines, batch processing, real-time streaming, and data orchestration services.

•       Partner with Data Engineers to improve pipeline reliability and data quality monitoring.

•       Ensure platform scalability for growing data volumes and user demands.

Security & Compliance

•       Implement security best practices and operational controls.

•       Support compliance requirements related to data governance and privacy.

•       Collaborate with security teams to remediate vulnerabilities and improve platform security posture.

Continuous Improvement

•       Conduct post-incident reviews and drive corrective and preventive actions.

•       Identify reliability risks and implement long-term improvements.

•       Promote a culture of operational excellence, resilience, and automation.

•       Contribute to engineering standards, runbooks, knowledge sharing, and best practices.

Required Qualifications

Education

•       Bachelor’s degree in computer science, Information Technology, Engineering, or related field, or equivalent practical experience.

Experience

•       3+ years of experience in Site Reliability Engineering, DevOps, Systems Engineering, Cloud Engineering, or a related role.

•       Experience supporting production environments with high availability requirements.

•       Experience managing cloud infrastructure and distributed systems.

Technical Skills

•       Strong knowledge of Linux systems administration.

•       Proficiency in one or more programming languages such as Python, Java, Go, C#, or JavaScript.

•       Experience with CI/CD tools and deployment automation.

•       Experience with Infrastructure as Code tools such as Terraform, ARM, or Bicep.

•       Experience with Kubernetes and container technologies.

•       Knowledge of monitoring and observability technologies such as Prometheus, Grafana, Datadog, Azure Monitor, Splunk, or OpenTelemetry.

•       Understanding of networking, DNS, load balancing, and distributed systems concepts.

Experience supporting data platforms such as:

o   Azure Data Lake

o   Azure Synapse Analytics

o   Databricks

o   Snowflake

o   Kafka

o   SQL/NoSQL databases

o   Data orchestration platforms

Preferred Qualifications

•       Experience supporting AI/ML platforms and MLOps environments.

•       Experience with Azure cloud-native services.

•       Familiarity with data governance and data quality frameworks.

•       Knowledge of reliability engineering best practices and SRE methodologies.

•       Experience implementing SLOs, SLIs, and error budgets.

•       Experience supporting large-scale analytics and business intelligence environments.

•       Azure, AWS, Kubernetes, Terraform, or DevOps certifications.

Key Competencies

•       Problem-solving and analytical thinking

•       Incident management and troubleshooting

•       Automation-first mindset

•       Collaboration and stakeholder management

•       Strong communication skills

•       Continuous learning and innovation

•       Customer-focused approach

•       Operational excellence and accountability Success Measures

•       Maintaining high availability and reliability targets for critical services.

•       Reducing operational toil through automation.

•       Improving platform observability and incident response effectiveness.

•       Meeting service-level objectives and performance goals.

•       Enhancing deployment reliability and operational efficiency.

•       Driving measurable improvements in system resilience, scalability, and customer experience.