1

Infiniband Jobs in Virginia (NOW HIRING)

Troubleshoot hardware, OS, scheduler, networking, and high-performance interconnect issues (e.g., InfiniBand) * Integrate compute nodes and hardware into clusters * Develop automation and operational ...

HPC Systems Engineer

Charlottesville, VA · On-site

$150K - $200K/yr

Troubleshoot hardware, OS, scheduler, networking, and high-performance interconnect issues (e.g., InfiniBand) * Integrate compute nodes and hardware into clusters * Develop automation and operational ...

HPC Systems Engineer

Charlottesville, VA · On-site

$150K - $200K/yr

Troubleshoot hardware, OS, scheduler, networking, and high-performance interconnect issues (e.g., InfiniBand) * Integrate compute nodes and hardware into clusters * Develop automation and operational ...

HPC Software Engineer

Reston, VA · On-site

$120 - $180/hr

Experience with high-speed interconnects (e.g., InfiniBand) * Understanding of job schedulers and resource managers (e.g., Slurm) * Strong technical documentation and reporting skills * Ability to ...

HPC Software Engineer

Chantilly, VA · On-site

$120 - $150/hr

Experience with high-speed interconnects (e.g., InfiniBand) * Understanding of job schedulers and resource managers (e.g., Slurm) * Strong technical documentation and reporting skills * Ability to ...

Experience with high-speed interconnects (e.g., InfiniBand) * Understanding of job schedulers and resource managers (e.g., Slurm) * Strong technical documentation and reporting skills * Ability to ...

Experience with high-speed interconnects (e.g., InfiniBand) * Understanding of job schedulers and resource managers (e.g., Slurm) * Strong technical documentation and reporting skills * Ability to ...

Experience with high-speed interconnects (e.g., InfiniBand) * Understanding of job schedulers and resource managers (e.g., Slurm) * Strong technical documentation and reporting skills * Ability to ...

Showing results 21-38

Infiniband information

What is InfiniBand?

Infiniband is a high-speed, low-latency networking technology commonly used in data centers and high-performance computing environments. It is designed to connect servers, storage systems, and network devices, providing much faster data transfer rates than traditional Ethernet. Infiniband supports scalable bandwidth and efficient communication, which makes it ideal for applications requiring rapid data movement, such as scientific simulations and large-scale database transactions. Its architecture also supports remote direct memory access (RDMA), which further reduces latency and CPU overhead.

What are the typical responsibilities of an InfiniBand network engineer in a data center environment?

InfiniBand network engineers are primarily responsible for designing, deploying, and maintaining high-performance InfiniBand fabrics that connect servers and storage systems in data centers, especially in HPC (High-Performance Computing) environments. Their daily tasks include monitoring network performance, troubleshooting connectivity or latency issues, and performing firmware and driver updates on InfiniBand switches and host adapters. They also collaborate closely with system administrators and application teams to optimize throughput and ensure reliable, low-latency communication. Additionally, InfiniBand engineers often participate in capacity planning and help scale the network infrastructure to meet growing computational demands.

What are the key skills and qualifications needed to thrive as an InfiniBand network engineer, and why are they important?

To thrive as an InfiniBand Network Engineer, you need a strong background in computer networking, Linux system administration, and high-performance computing (HPC) environments, often supported by a degree in computer science or related field. Familiarity with InfiniBand architecture, experience with tools like OpenFabrics Enterprise Distribution (OFED), and certifications such as CompTIA Network+ are valuable. Strong problem-solving skills, attention to detail, and effective communication are crucial soft skills for this role. These abilities are essential for ensuring efficient, reliable InfiniBand network performance in complex HPC or data center environments.

What is the difference between Infiniband vs Ethernet Network Engineer?

AspectInfinibandEthernet Network Engineer
Required CredentialsNetworking certifications, Cisco, Cisco CCNA, CCNPNetworking certifications, Cisco, CCNA, CCNP
Work EnvironmentData centers, high-performance computing environmentsCorporate networks, data centers, enterprise environments
Industry UsageHigh-performance computing, research institutionsBusiness, telecommunications, enterprise IT
Common Search/ComparisonYesYes

Infiniband and Ethernet Network Engineers both work with network infrastructure, but Infiniband specializes in high-speed, low-latency connections used in data centers and HPC environments. Ethernet Network Engineers focus on standard Ethernet networks used across various industries. While their certifications and skills overlap, their work environments and applications differ significantly.

What job categories do people searching Infiniband jobs in Virginia look for?

The top searched job categories for Infiniband jobs in Virginia are:

What cities in Virginia are hiring for Infiniband jobs?

Cities in Virginia with the most Infiniband job openings:

Infographic showing various Infiniband job openings in Virginia as of August 2026, with employment types broken down into 98% Full Time, and 2% Contract. Highlights an 81% Physical, 4% Hybrid, and 15% Remote job distribution.

HPC Systems Engineer

teKnoluxion

Charlottesville, VA

$150K - $200K/yr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Re-posted 10 days ago


Job description

OverviewHPC Systems EngineerLocation: Charlottesville, VA

Clearance Required:  Active TS (SCI eligibility) 

At Bcore, our strength comes from how we deliver impact to the mission. Whether it's architecting critical IT solutions, producing actionable intelligence, or developing cutting edge technology, we succeed because of the expertise, collaboration, and agility of our teams. Our Mission Services division combines enterprise IT, cloud solutions, DevSecOps, systems engineering, software development, and operational support. Bcore accelerates decisive advantage for warfighters and intelligence professionals by fusing human insight, rapid-fire engineering, precision-measured outcomes, and relentless grit into mission-ready solutions. 

Do you want to join a team that is building tailored technical solutions to modernize our government's mission and our client's business?  Do you have a desire to change how people work?  Are you interested in helping to protect our nation's cyber interests? Join our growing team supporting the Army customer's mission as an HPC Systems Engineer.

ResponsibilitiesWhat you get to do every day:
  • Build, configure, and maintain secure HPC clusters for simulations, scientific computing, and GPU workloads
  • Collaborate with infrastructure teams on cluster platforms, including schedulers, provisioning systems, high-speed interconnects, and distributed nodes
  • Configure and manage job schedulers (Slurm, PBS) with queue setup, resource policies, and job optimization
  • Support containerized workloads (Docker, Podman, Singularity/Apptainer)
  • Assist with cluster provisioning, node management, and initial build-out, including scheduler configuration and validation
  • Troubleshoot hardware, OS, scheduler, networking, and high-performance interconnect issues (e.g., InfiniBand)
  • Integrate compute nodes and hardware into clusters
  • Develop automation and operational tools using Bash, Python, or similar scripting
  • Support authentication and access control via LDAP or Kerberos
  • Analyze performance and identify bottlenecks across compute, storage, and network layers for distributed workloads (MPI/OpenMP)
  • Support GPU-enabled environments and CUDA-based workloads
  • Coordinate with engineering teams to improve cluster performance, stability, and scalability
  • Maintain documentation for configurations, procedures, and troubleshooting
  • Provide technical guidance on HPC best practices for mission workloads
Qualifications

Clearance Required: Active TS clearance (with SCI Eligibility) and eligibility to obtain CI Poly **We are not able to upgrade or sponsor clearances**

Certification Required: Ability to obtain DoD 8140 (8570) IAT Level II certification

Education/Experience:

  • Requires Bachelor's degree in Engineering, Computer Science, or related STEM field (experience in lieu of degree)
  • 6+ years of experience administering Linux based systems in enterprise, research computing, or distributed compute environments, including configuration and troubleshooting of multi-node systems.
Required Skills:
  • Experience supporting distributed compute environments with workload schedulers (e.g., Slurm, PBS, Torque, Grid Engine)
  • Experience supporting multi-node compute environments or HPC clusters
  • Professional experience administering Linux systems via CLI (RHEL derivatives preferred)
  • Experience with scripting and automation (Bash, Python, or similar)
  • Experience troubleshooting server hardware, OS, and distributed computing systems
  • Familiarity with cluster networking and high-speed interconnects
  • Experience diagnosing performance issues across compute, networking, and storage layers
  • Strong troubleshooting and documentation skills

What is ideal?

  • Experience administering multi-node HPC clusters and supporting distributed workloads
  • Knowledge of parallel file systems (e.g., Lustre, BeeGFS, GPFS)
  • Experience with parallel computing frameworks (MPI, OpenMP)
  • Experience with configuration management tools (Ansible, Puppet)
  • Experience supporting GPU-enabled environments and CUDA workloads
  • Familiarity with hybrid HPC architectures (on-prem + cloud, e.g., AWS)
  • Experience supporting HPC systems in research, lab, or mission environments
  • Experience working in DoD or IC environments preferred
What you can expect from us
  • Recognizing great achievements do not go unnoticed by Bcore through service anniversaries, spot awards, and employee referral bonuses
  • You'll join a growing organization of passionate, top-shelf, IT engineering professionals with extensive experience in actively developing the technology revolution in the Intelligence community
  • The expected salary range within the Washington, DC metropolitan area is: $150,000-$200,200.00. Final compensation is unique to each individual and will be determined based on factors such as experience, education, geographic location, and contractual requirements. This is not a guarantee.
  • Benefits include Health/Dental/Vision, 401(k) match, Paid Time Off, STD/LTD/Life Insurance/Voluntary Life Insurance, Stipends, Referral Bonuses, and more.
BCore is proud to be an equal opportunity workplace. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, disability status, protected veteran status, sexual orientation or any other characteristic protected by law.Employment Type: FULL_TIME