1

Network Operations Manager Jobs in Castro Valley, CA

Network DevOps Engineer

Palo Alto, CA ยท On-site

$62 - $85/hr

Job #105378 Title Network DevOps Engineer - Multi-Cloud Infrastructure Automation Client Leading ... Support deployment and management of virtual networks, routing, load balancing, and traffic ...

... operational improvements on systems such as network management, alarm management, auto-provisioning systems preferred ยท Experience handling customer issues, troubleshooting and resolution follow-up ...

Our WW Operations network delivers millions of packages and smiles to Amazon customers every day ... To achieve this, managers are expected to provide their team with the tools needed for success ...

Our WW Operations network delivers millions of packages and smiles to Amazon customers every day ... To achieve this, managers are expected to provide their team with the tools needed for success ...

External Our WW Operations network delivers millions of packages and smiles to Amazon customers ... To achieve this, managers are expected to provide their team with the tools needed for success ...

Our WW Operations network delivers millions of packages and smiles to Amazon customers every day ... To achieve this, managers are expected to provide their team with the tools needed for success ...

External Our WW Operations network delivers millions of packages and smiles to Amazon customers ... To achieve this, managers are expected to provide their team with the tools needed for success ...

Operations Manager

Stockton, CA ยท On-site

$75K - $80K/yr

Successful completion of the WM Operations Manager Trainee program IV. Physical Requirements Listed ... WM has the largest disposal network and collection fleet in North America, is the largest recycler ...

Operations Manager

Stockton, CA ยท On-site

$75K - $80K/yr

Successful completion of the WM Operations Manager Trainee program IV. Physical Requirements Listed ... WM has the largest disposal network and collection fleet in North America, is the largest recycler ...

Our WW Operations network delivers millions of packages and smiles to Amazon customers every day ... To achieve this, managers are expected to provide their team with the tools needed for success ...

Operations Manager

Stockton, CA ยท On-site

$75K - $80K/yr

Successful completion of the WM Operations Manager Trainee program IV. Physical Requirements Listed ... WM has the largest disposal network and collection fleet in North America, is the largest recycler ...

Router & Core Network Operations: Protocol Maintenance: Manage routing protocols, failover capabilities, and traffic paths. Network Stability: Maintain dynamic routing environments, including OSPF ...

Router & Core Network Operations: Protocol Maintenance: Manage routing protocols, failover capabilities, and traffic paths. Network Stability: Maintain dynamic routing environments, including OSPF ...

Showing results 21-40

Network Operations Manager information

See Castro Valley, CA salary details

$25.4K

$117.3K

$169.1K

How much do network operations manager jobs pay per year?

As of Sep 6, 2026, the average yearly pay for network operations manager in Castro Valley, CA is $117,324.00, according to ZipRecruiter salary data. Most workers in this role earn between $98,100.00 and $131,500.00 per year, depending on experience, location, and employer.

What is a network operations manager?

Network Operations Managers are professionals responsible for overseeing the day-to-day operations, maintenance, and performance of an organization's computer networks. They manage teams of network engineers and technicians, ensure network reliability and security, and handle incidents or outages. Their role includes planning for future network needs, implementing upgrades, and ensuring compliance with industry standards. They also coordinate with other departments to align network services with business goals.

What are the key skills and qualifications needed to thrive as a network operations manager, and why are they important?

To excel as a Network Operations Manager, you need a solid background in network administration, troubleshooting, and IT management, typically backed by a relevant degree and experience in network operations. Familiarity with network monitoring tools (such as SolarWinds or Nagios), ITIL frameworks, and certifications like CCNP or CompTIA Network+ are commonly required. Strong leadership, problem-solving abilities, and effective communication are crucial soft skills for managing teams and responding to network issues. These competencies are essential to ensure network reliability, minimize downtime, and drive operational efficiency in business-critical environments.

What are some common challenges faced by network operations managers, and how can they be effectively addressed?

Network Operations Managers often encounter challenges such as minimizing network downtime, managing multiple priorities, and ensuring rapid response to incidents. Balancing proactive maintenance with reactive troubleshooting requires strong organizational skills and close collaboration with IT teams, vendors, and sometimes clients. Effective communication, implementing robust monitoring tools, and fostering a culture of continuous improvement can help address these challenges and maintain network reliability.

What is the difference between Network Operations Manager vs Network Engineer?

AspectNetwork Operations ManagerNetwork Engineer
CredentialsTypically requires a bachelor's degree in IT or related field, with certifications like Cisco CCNA, CCNP, or CompTIA Network+Requires similar certifications such as Cisco CCNA, CCNP, or CompTIA Network+; often more technical focus
Work EnvironmentOversees network teams, manages operations, and ensures network reliability in enterprise settingsDesigns, implements, and troubleshoots network infrastructure, often working hands-on with hardware and software
Employer & Industry UsageCommonly employed by large corporations, service providers, and government agenciesFound in various industries, including IT firms, telecom, and corporate IT departments

The main difference is that Network Operations Managers focus on overseeing network performance, managing teams, and strategic planning, while Network Engineers are more involved in the technical design, implementation, and troubleshooting of network systems. Both roles require similar certifications and work environments but differ in scope and responsibilities.

What are popular job titles related to Network Operations Manager jobs in Castro Valley, CA?

For Network Operations Manager jobs in Castro Valley, CA, the most frequently searched job titles are:

What job categories do people searching Network Operations Manager jobs in Castro Valley, CA look for?

The top searched job categories for Network Operations Manager jobs in Castro Valley, CA are:

What cities near Castro Valley, CA are hiring for Network Operations Manager jobs?

Cities near Castro Valley, CA with the most Network Operations Manager job openings:

Infographic showing various Network Operations Manager job openings in Castro Valley, CA as of July 2026, with employment types broken down into 100% Full Time. Highlights an 86% In-person, and 14% Remote job distribution, with an average salary of $117,324 per year, or $56.4 per hour.

Network Operations Engineer, AI Networking

OpenAI

San Francisco, CA โ€ข On-site

$140 - $210/hr

Other

Re-posted yesterday


Job description

About the Team

OpenAIโ€™s Infrastructure Operations team is responsible for the availability, reliability, and operational excellence of one of the worldโ€™s largest AI infrastructure networks. The team owns day-to-day operations of production AI networks across Industrial Compute's data centers, working with colocation providers, deployment teams, and hardware vendors to deliver highly available GPU infrastructure for AI training and inference workloads.

About the Role

We are seeking an Infrastructure Operations Engineer to operate and improve the large-scale Ethernet fabrics that support GPU clusters, storage systems, and management infrastructure. This role combines handsโ€‘on production operations with automation, observability, and incident response across a global AI network.

The ideal candidate has experience operating high-availability data center, cloud, AI, or HPC networks and can move comfortably from physicalโ€‘layer troubleshooting to routing and fabric behavior, change execution, and rootโ€‘cause analysis. You will partner closely with network architecture, systems engineering, GPU engineering, storage engineering, security, deployment, site operations, service providers, colocation partners, and hardware vendors to raise reliability and reduce operational toil.

Key Responsibilities
  • Own the operational health, availability, and reliability of production AI network infrastructure across Industrial Compute's data centers.

  • Monitor, troubleshoot, and resolve network incidents while meeting serviceโ€‘level objectives (SLOs), reducing Mean Time to Detect (MTTD), and minimizing Mean Time to Recovery (MTTR).

  • Operate and maintain large-scale Ethernet fabrics supporting GPU compute, storage, and management networks.

  • Execute production network changes, maintenance windows, and capacity expansions with minimal customer impact.

  • Manage the hardware lifecycle, including switch and optics replacements, RMA coordination, software upgrades, and preventive maintenance.

  • Support new AI cluster deployments, data center expansions, and infrastructure migrations in partnership with deployment and engineering teams.

  • Partner with cloud service providers (CSPs), colocation providers, Smart Hands teams, and hardware vendors to maintain production infrastructure.

  • Perform rootโ€‘cause analysis (RCA) for production incidents and drive permanent corrective actions that eliminate recurring issues.

  • Build and maintain monitoring, telemetry, dashboards, and alerting to improve network observability and proactive issue detection.

  • Develop and improve operational runbooks, playbooks, troubleshooting documentation, and standard operating procedures.

  • Automate repetitive operational tasks using Python and infrastructure automation frameworks to reduce toil and improve efficiency.

  • Continuously identify opportunities to improve service reliability, scalability, operational maturity, and engineering efficiency.

Qualifications
  • Bachelorโ€™s degree in Computer Science, Network Engineering, or a related discipline, or equivalent practical experience.

  • 5+ years of experience operating largeโ€‘scale data center, cloud, AI, or HPC network infrastructure.

  • Experience supporting production network environments with highโ€‘availability requirements.

  • Handsโ€‘on experience with one or more of the following platforms: Cisco NXโ€‘OS, Arista EOS, NVIDIA Spectrum / Cumulus Linux, or Juniper JunOS.

  • Strong knowledge of Layer 2 and Layer 3 networking, BGP, OSPF, ECMP, MLAG, LACP, VRFs, and VLANs.

  • Experience troubleshooting physical infrastructure, including fiber optics, transceivers, DAC/AOC cables, and highโ€‘speed Ethernet links.

  • Experience performing software upgrades, hardware maintenance, and production change management.

  • Excellent analytical and troubleshooting skills, with the ability to communicate technical risk clearly across teams.

Preferred Skills
  • Experience operating AI or Highโ€‘Performance Computing (HPC) network environments.

  • Experience with NVIDIA AI networking technologies and GPU infrastructure.

  • Experience supporting RoCEโ€ฏv2 or RDMAโ€‘based Ethernet fabrics, with a strong understanding of Priority Flow Control (PFC), Explicit Congestion Notification (ECN), Data Center Quantized Congestion Notification (DCQCN), Quality of Service (QoS), and lossless Ethernet networking.

  • Experience supporting 100G, 200G, 400G, and 800G Ethernet networks.

  • Experience with GPU platforms including NVIDIA HGX, DGX, GB200, or equivalent AI infrastructure.

  • Experience supporting distributed storage environments such as VAST, DDN, or similar technologies.

  • Experience working with cloud service providers such as AWS, Azure, or Google Cloud, and with thirdโ€‘party colocation providers.

  • Experience with network monitoring and telemetry technologies, including Prometheus, Grafana, gNMI, streaming telemetry, SNMP, or similar tools.

  • Experience developing automation using Python, Git, REST APIs, Terraform, or similar automation frameworks.

Work Environment and Onโ€‘Call
  • Participate in a 24x7 onโ€‘call rotation supporting missionโ€‘critical AI infrastructure.

  • Support timeโ€‘sensitive production incidents, maintenance windows, capacity expansions, and network changes with a focus on service availability and minimal customer impact.

  • This role requires up to 30% travel to data center locations for new turnups and acceptance activities, as needed.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that generalโ€‘purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

For additional information, please see OpenAIโ€™s affirmative action and equal employment opportunity policy statement.

Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for USโ€‘based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and nonโ€‘public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.

To notify OpenAI that you believe this job posting is nonโ€‘compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.

We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.

OpenAI Global Applicant Privacy Policy

At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

#J-18808-Ljbffr