1

Director Infrastructure Operations Jobs in Baltimore, MD

Summary The Director, Engineering is responsible for leading the design, development, delivery ... Partner with infrastructure, operations, SRE, and end-user support teams to ensure stable ...

Head of Network Services

Baltimore, MD · On-site

$201 - $342/hr

... Infrastructure Operations * Lead an organization of managers, senior engineers, and technical ... Azure including Direct Connect and Express Route; demonstrated experience leading ...

Showing results 21-40

Director Infrastructure Operations information

See Baltimore, MD salary details

$33.8K

$107K

$178.4K

How much do director infrastructure operations jobs pay per year?

As of Aug 22, 2026, the average yearly pay for director infrastructure operations in Baltimore, MD is $106,995.00, according to ZipRecruiter salary data. Most workers in this role earn between $75,000.00 and $134,600.00 per year, depending on experience, location, and employer.

What does a director of infrastructure operations do?

A Director of Infrastructure Operations is responsible for overseeing and managing an organization's IT infrastructure, including networks, servers, data centers, and cloud services. They ensure that all systems are running efficiently, securely, and reliably to support business operations. This role often involves strategic planning, budgeting, leading technical teams, and implementing best practices for disaster recovery and security. Directors also collaborate with other departments to align IT infrastructure with organizational goals and may be involved in vendor management and technology upgrades.

What are the key skills and qualifications needed to thrive as a director of infrastructure operations?

To thrive as a Director of Infrastructure Operations, you need deep expertise in IT infrastructure management, network architecture, and systems administration, often supported by a bachelor’s or master’s degree in computer science or a related field. Familiarity with enterprise tools like VMware, cloud platforms (AWS, Azure), ITIL frameworks, and certifications such as PMP or CISSP is typically required. Strong leadership, strategic thinking, and effective communication are essential soft skills for managing teams and aligning IT initiatives with business goals. These skills and qualities ensure the reliable, secure, and scalable operation of an organization’s critical IT systems.

How does a director of infrastructure operations typically collaborate with other departments to ensure seamless IT service delivery?

As a Director of Infrastructure Operations, collaboration with other departments such as security, application development, and business units is essential to align IT infrastructure with organizational goals. This role often involves regular meetings with department heads to understand their requirements, prioritizing projects, and coordinating resource allocation. Directors also work closely with teams to create incident response protocols, ensure compliance, and optimize system uptime. Effective communication and cross-functional teamwork are key to addressing challenges quickly and maintaining high service levels.

What is the difference between Director Infrastructure Operations vs Network Manager?

AspectDirector Infrastructure OperationsNetwork Manager
Primary FocusOversees overall infrastructure strategy, including data centers, servers, and cloud servicesManages network infrastructure, including LAN, WAN, and network security
CertificationsITIL, PMP, Cisco or Microsoft certifications often preferredCCNA, CCNP, or Network+ certifications common
Work EnvironmentExecutive-level, strategic planning, cross-departmental collaborationOperational, technical hands-on management of network systems
Industry UsageUsed in large enterprises, data centers, cloud providersCommon in organizations with complex network needs

The main difference is that the Director Infrastructure Operations oversees the entire infrastructure strategy and management, including data centers and cloud services, while the Network Manager focuses specifically on managing and maintaining network systems. Both roles require technical certifications, but their scope and responsibilities differ significantly.

What are the most commonly searched types of Infrastructure Operations jobs in Baltimore, MD?

The most popular types of Infrastructure Operations jobs in Baltimore, MD are:

What are popular job titles related to Director Infrastructure Operations jobs in Baltimore, MD?

For Director Infrastructure Operations jobs in Baltimore, MD, the most frequently searched job titles are:

What job categories do people searching Director Infrastructure Operations jobs in Baltimore, MD look for?

The top searched job categories for Director Infrastructure Operations jobs in Baltimore, MD are:

What cities near Baltimore, MD are hiring for Director Infrastructure Operations jobs?

Cities near Baltimore, MD with the most Director Infrastructure Operations job openings:

AI Operations & Infrastructure Engineer

Invictus International Consulting, LLC

Fort George G Meade, MD • On-site

$175K - $200K/yr

Full-time

Re-posted yesterday


Job description

Title: AI Operations & Infrastructure Engineer
Location: Fort Meade, MD
Clearance: TS/SCI with a CI Polygraph
Job Details:
  • Manage and maintain AI computing platforms, including GPUs and other specialized hardware
  • Install and configure GPU drivers and software
  • Oversee the AI software stack and tools
  • Implement and manage containerization technologies like Docker and Kubernetes
  • Configure and optimize networking infrastructure for AI workloads, including InfiniBand and Ethernet
  • Manage storage solutions for AI data, considering performance and capacity requirements
  • Deploy and manage data processing units (DPUs) to accelerate data center workloads
  • Monitor and manage AI cluster health and resource utilization
  • Implement workload management and scheduling tools like Slurm and Kubernetes
  • Ensure efficient power and cooling for AI infrastructure to maintain optimal operating conditions
  • Configure high-performance networking solutions for AI and machine learning workloads
  • Optimize network performance to ensure maximum throughput and minimal latency for AI computations
  • Implement and fine-tune network protocols to enhance data transfer speeds and efficiency
  • Integrate NVIDIA networking products with existing AI infrastructure, including servers, GPUs, and storage systems
  • Deploy networking solutions in data centers to ensure seamless connectivity between AI components
  • Diagnose and resolve networking issues impacting AI workloads to maintain optimal system performance
  • Provide technical support and guidance to teams managing AI infrastructure
  • Collaborate with data scientists, researchers, and IT professionals to understand networking requirements and challenges
  • Lead deployment and validation of servers and systems for AI enabled platforms
  • Configure and manage network topologies, BMC, OOB, TPM, power, and cooling
  • Install, upgrade, and validate GPU-based servers, BlueField DPUs, cables, and transceivers
  • Perform firmware upgrades, hardware validation, and storage setup
  • Configure and administer physical and logical resources, including M IG partitioning and BlueField platforms
  • Install and configure operating systems, cluster software, drivers, containers (Docker), and NGC CLI
  • Manage and orchestrate clusters using NVIDIA Base Command Manager, Slurm, Pyxis, Enroot, and Run: Ai
  • Perform stress, benchmarking, and burn-in tests using HPL, NCCL, NVIDIA Nemo, and ClusterKit
  • Verify cabling, firmware/software versions, and network signal quality
  • Troubleshoot and resolve hardware, software, storage, and performance faults
  • Replace faulty components and optimize systems for AMD/Intel platforms
  • Monitor, document, and report on cluster health, resource usage, and job performance
  • Ensure secure, efficient, and scalable operation of NVIDIA AI infrastructure, including user access and workload management

Requirements:
  • Qualified candidates must hold an active NVIDIA Professional Certification in either AI Networking, AI Infrastructure, or AI Operations
  • Prior direct, hands-on professional experience administering NVIDIA GPU and data processing unit (DPU) technologies, AI software stacks, and data center environments for high-performance AI workloads
  • Comprehensive expertise in deploying and maintaining AI compute platforms, requiring proficiency in containerization and workload orchestration using Docker, Kubernetes, Slurm, NVIDIA Base Command Manager, and Run:Ai
  • Must be capable of configuring physical and logical resources, including Multi-Instance GPU (MIG) partitioning and BlueField platforms, while overseeing critical facility elements such as power, cooling, and storage solutions
  • The ability to demonstrate advanced skills in AI networking, specifically configuring and optimizing high-performance InfiniBand and Ethernet fabrics to ensure maximum throughput and minimal latency
  • Current active TS/SCI clearance with a CI Polygraph

Equal Opportunity Employer/Veterans/Disabled