1

Network Reliability Engineer Jobs in Oakbrook Terrace, IL

Site Reliability Engineer III

Chicago, IL

$58.75 - $78/hr

As a Site Reliability Engineer III at JPMorgan Chase within the Commercial and Investment Banking ... Implements infrastructure, configuration, and network as code for the applications and platforms in ...

Site Reliability Engineer - Pharmacy

Deerfield, IL · On-site

$58 - $77/hr

Required : • 7+ years of experience in SRE, platform engineering, or cloud infrastructure ... network policies, pod security standards, cluster autoscaler, and Workload Identity. • Strong ...

Staff Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards ... Advanced networking knowledge - including VPCs, subnets, DNS, load balancing, firewall rules, VPNs ...

Staff Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards ... Advanced networking knowledge - including VPCs, subnets, DNS, load balancing, firewall rules, VPNs ...

Staff Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards ... Advanced networking knowledge - including VPCs, subnets, DNS, load balancing, firewall rules, VPNs ...

Staff Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards ... Advanced networking knowledge - including VPCs, subnets, DNS, load balancing, firewall rules, VPNs ...

Staff Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

The Site Reliability Engineering team drives reliability strategy, elevates engineering standards ... Advanced networking knowledge -- including VPCs, subnets, DNS, load balancing, firewall rules, VPNs ...

Showing results 21-40

Network Reliability Engineer information

See Oakbrook Terrace, IL salary details

$61.3K

$118.6K

$141.8K

How much do network reliability engineer jobs pay per year?

As of Aug 7, 2026, the average yearly pay for network reliability engineer in Oakbrook Terrace, IL is $118,639.00, according to ZipRecruiter salary data. Most workers in this role earn between $103,100.00 and $129,700.00 per year, depending on experience, location, and employer.

What is a network reliability engineer?

A Network Reliability Engineer (NRE) is an IT professional responsible for ensuring the reliability, performance, and scalability of network systems. They combine skills in networking, software engineering, and automation to proactively detect and resolve potential network issues before they affect users. NREs often design and implement monitoring tools, automate network management tasks, and work to improve the overall stability of network infrastructure. Their goal is to minimize downtime and ensure seamless connectivity across an organization’s network.

What is the difference between Network Reliability Engineer vs Network Operations Center (NOC) Technician?

AspectNetwork Reliability EngineerNetwork Operations Center (NOC) Technician
CertificationsCCNA, CCNP, Network+CCNA, Network+
Work EnvironmentDesign, analyze, and improve network infrastructureMonitor, troubleshoot, and maintain networks in real-time
Employer & Industry UsageTelecom, large enterprises, cloud providersISPs, data centers, enterprise networks
Common Search & ComparisonFocus on network reliability and designFocus on network monitoring and incident response

The main difference is that Network Reliability Engineers focus on designing and improving network systems to ensure long-term reliability, while NOC Technicians monitor and troubleshoot networks in real-time to resolve issues quickly. Both roles require relevant certifications and are essential in maintaining network performance, but they serve different functions within network management.

What are the key skills and qualifications needed to thrive as a network reliability engineer, and why are they important?

To thrive as a Network Reliability Engineer, you need a strong background in computer networking, network protocols, troubleshooting, and often a degree in computer science or a related field. Familiarity with tools like Wireshark, Nagios, Cisco IOS, and certifications such as CCNA or CCNP are commonly required. Analytical thinking, proactive problem-solving, and effective communication are standout soft skills in this role. These skills are crucial to maintaining reliable network operations, minimizing downtime, and ensuring seamless communication across organizational systems.

What are some common challenges faced by network reliability engineers, and how are they typically addressed?

Network Reliability Engineers often encounter challenges such as diagnosing intermittent connectivity issues, managing network upgrades with minimal downtime, and maintaining high availability during peak traffic. These are typically addressed by leveraging robust monitoring tools, implementing automation for routine tasks, and collaborating closely with software engineers, network administrators, and incident response teams. Staying current with evolving network technologies and best practices is essential for effectively identifying and resolving problems before they impact users.
What cities near Oakbrook Terrace, IL are hiring for Network Reliability Engineer jobs? Cities near Oakbrook Terrace, IL with the most Network Reliability Engineer job openings:
Infographic showing various Network Reliability Engineer job openings in Oakbrook Terrace, IL as of July 2026, with employment types broken down into 1% As Needed, 77% Full Time, 14% Part Time, 7% Contract, and 1% Nights. Highlights an 92% Physical, 2% Hybrid, and 6% Remote job distribution, with an average salary of $118,639 per year, or $57 per hour.

Sr. Site Reliability Engineer (SRE) (Chicago)

Moonlite

Chicago, IL • On-site

$58.75 - $78/hr

Full-time

Medical, Retirement

Re-posted 5 days ago


Job description

Moonlite delivers high-performance AI infrastructure for organizations running intensive computational research, large-scale model training, and demanding data processing workloads. We provide infrastructure deployed in our facilities or co-located in yours, delivering flexible on‑demand or reserved compute that feels like an extension of your existing data center. Our team of AI infrastructure specialists blends bare‑metal performance with cloud‑native operational simplicity, enabling research teams and enterprises to deploy demanding AI workloads with enterprise‑grade reliability and compliance.

Your Role

You will be instrumental in building and operating production‑grade AI infrastructure with deep Kubernetes expertise at its core. Working closely with our systems engineers, network engineers, and platform engineering team, you’ll architect and operate the Kubernetes infrastructure that powers our control plane and orchestrates compute, storage, and networking at scale. This role requires deep understanding of Kubernetes internals, custom resource definitions (CRDs), storage and network integrations, and building production‑grade clusters from the ground up (not just deploying in managed environments). You'll ensure enterprise‑grade reliability while establishing the automation, observability, and operational practices.

Job Responsibilities
  • Kubernetes Infrastructure Engineering: Design, build, and operate production Kubernetes clusters on bare‑metal infrastructure – including cluster bootstrapping, control plane architecture, etcd management, and scaling strategies for high‑performance compute workloads.
  • Kubernetes Networking & CNIs: Implement and operate custom Kubernetes networking solutions with SR‑IOV for high‑performance GPU interconnects, multi‑tenancy isolation and advanced networking policies. Configure CNI plugins and network segmentation for research workloads.
  • Custom Operators & Controllers: Develop and maintain custom Kubernetes operators and controllers for bare‑metal provisioning, infrastructure lifecycle management, and resource orchestration across compute, storage, and networking domains.
  • GPU Infrastructure Integration: Deploy and optimize NVIDIA GPU operators, device plugins, and other custom scheduling logic for GPU workload placement and utilization optimization.
  • Platform Integration & Storage: Build deep integrations between Kubernetes and underlying infrastructure including CSI drivers for storage, custom admission controllers for policy enforcement, and scheduling extensions for specialized hardware placement.
  • Infrastructure Automation: Design and implement automation using Terraform, Ansible, Helm, and custom operators to orchestrate infrastructure workflows and enable deployments across multiple regions.
  • Production Operations & Reliability: Manage production bare‑metal infrastructure across multiple regions. Build systems ensuring high availability, fault tolerance, and graceful degradation – establishing SLIs, SLOs, and monitoring to meet enterprise reliability commitments.
  • Observability & Incident Response: Build comprehensive monitoring, logging, and alerting using Prometheus, Grafana, and ELK stack. Lead incident response, conduct postmortems, and implement preventative measures to improve reliability and reduce MTTR.
  • Performance & Capacity Planning: Identify and resolve performance bottlenecks across infrastructure domains. Monitor utilization trends, forecast capacity needs, and optimize resource allocation for various workloads.
Requirements
  • Experience: 5+ years in SRE, DevOps, or infrastructure engineering roles with proven experience operating production infrastructure at scale.
  • Kubernetes Infrastructure Expertise: Deep hands‑on experience building and operating production Kubernetes clusters on bare‑metal infrastructure – not just deploying workloads in managed clusters. Must understand cluster bootstrapping, control plane architecture, etcd operations, and scaling strategies.
  • Kubernetes Internals & Integration: Strong understanding of Kubernetes internals including custom resource definitions (CRDs), operators, controllers, admission webhooks, and scheduling. Experience integrating storage (CSI drivers), networking (CNI, SR‑IOV), and specialized hardware (GPU device plugins) with Kubernetes.
  • Linux Systems Experience: Strong fundamentals in Linux systems administration, performance tuning, troubleshooting, and automation in production environments.
  • Infrastructure Automation: Proficiency with infrastructure‑as‑code tools (Terraform, Ansible, Helm) and building automation to reduce operational overhead.
  • Networking Fundamentals: Solid understanding of networking concepts including IPAM, DNS, DHCP, VLAN/VXLAN, routing, load balancing, and experience troubleshooting network issues in production.
  • Observability & Monitoring: Experience building and maintaining comprehensive monitoring solutions using tools such as Prometheus, Grafana, and centralized logging systems.
  • Reliability Practices: Understanding of SRE principles including SLIs/SLOs/SLAs, error budgets, incident management, and blameless postmortems.
  • Scripting & Automation: Strong scripting skills in Go, Python, or Bash for automation, tooling development, and operational efficiency.
  • Problem‑Solving Under Pressure: Demonstrated ability to troubleshoot complex issues under pressure, manage incidents effectively, and communicate clearly during outages.
  • Collaboration & Communication: Excellent communication skills and ability to work across teams including systems engineers, network engineers, and software developers.
Preferred Qualifications
  • Experience building custom Kubernetes operators or controllers for infrastructure orchestration.
  • Deep familiarity with Kubernetes networking (Calico, Cilium, Multus), service mesh technologies, and network policy management.
  • Experience with GPU workload orchestration including NVIDIA GPU Operator, MIG, time‑slicing, and device plugins.
  • Background with advanced Kubernetes features including custom schedulers, admission controllers, and API server extensions.
  • Experience with Kubernetes cluster federation or multi‑cluster management.
  • Knowledge of high‑performance networking technologies (InfiniBand, RDMA, RoCE) and their integration with Kubernetes.
  • Experience with enterprise storage systems (VAST, Lightbits, Ceph, or similar).
  • Familiarity with configuration management at scale and GitOps practices.
  • Understanding of security best practices for Kubernetes and bare‑metal infrastructure.
  • Experience operating infrastructure in regulated industries or co‑located data center environments.
  • Background supporting research institutions, technical computing environments, or enterprise AI infrastructure.
Why Moonlite
  • Build Critical Research Infrastructure: Your work will directly enable quantitative research teams and AI practitioners to push the boundaries of what’s possible in financial modeling and AI research.
  • Enterprise Impact: Build and operate infrastructure that supports mission‑critical research and AI workloads for leading financial institutions and research organizations.
  • Technical Excellence: Join an infrastructure team focused on delivering enterprise‑grade reliability while pushing the boundaries of high‑performance computing capabilities.
  • Hands‑On Ownership: As part of our growing infrastructure team, you’ll have significant ownership over critical systems and the autonomy to influence our operational practices and technology choices.
  • Industry Leadership: Work alongside experienced infrastructure professionals who have built and operated systems for the most demanding computing environments.

We offer a competitive total compensation package combining a base salary, startup equity, and industry‑leading benefits. The total compensation range for this role is $165,000 – $225,000, which includes both base salary and equity. Actual compensation will be determined based on experience, skills, and market alignment. We provide generous benefits, including a 6% 401(k) match, fully covered health insurance premiums, and other comprehensive offerings to support your well‑being and success as we grow together.

Equal Employment Opportunity

As set forth in Moonlite’s Equal Employment Opportunity policy, we do not discriminate on the basis of any protected group status under any applicable law. For government reporting purposes, we ask candidates to respond to the voluntary self‑identification survey. Completion of the form is entirely voluntary. Whatever your decision, it will not be considered in the hiring process or thereafter. Any information that you do provide will be recorded and maintained in a confidential file.

Voluntary Self‑Identification

We are a federal contractor or subcontractor. The law requires us to provide equal employment opportunity to qualified people with disabilities. We have a goal of having at least 7% of our workers as people with disabilities. The law says we must measure our progress towards this goal. To do this, we must ask applicants and employees if they have a disability or have ever had one. People can become disabled, so we need to ask this question at least every five years. Completing this form is voluntary, and we hope that you will choose to do so. Your answer is confidential. No one who makes hiring decisions will see it. Your d