1

Hpc Infrastructure Engineer Jobs (NOW HIRING)

Staff HPC Infrastructure Engineer

$110K - $144K/yr

To support Guardant Health's fast growth over the next few years, we need a strong technical engineer who can help maintain and grow the HPC infrastructure through this expansion while working ...

next page

Showing results 1-20

Hpc Infrastructure Engineer information

See salary details

$46.5K

$127.1K

$182K

How much do hpc infrastructure engineer jobs pay per year?

As of Sep 9, 2026, the average yearly pay for hpc infrastructure engineer in the United States is $127,066.00, according to ZipRecruiter salary data. Most workers in this role earn between $107,500.00 and $141,000.00 per year, depending on experience, location, and employer.

What is an HPC Infrastructure Engineer?

HPC Infrastructure Engineers are professionals who design, implement, and manage high-performance computing (HPC) systems and environments. They ensure that computational clusters, storage solutions, and networks are optimized for maximum performance and reliability. These engineers often work closely with researchers, scientists, and IT staff to support complex computing workloads, troubleshoot issues, and maintain system security. Their role is crucial in environments where large-scale simulations, data analysis, or scientific research require significant computational power.

What are some common challenges faced by an HPC Infrastructure Engineer when maintaining high-performance computing clusters?

HPC Infrastructure Engineers often encounter challenges related to scalability, hardware failures, and network bottlenecks when maintaining large computing clusters. Keeping systems up-to-date without causing downtime, troubleshooting complex performance issues, and ensuring efficient resource allocation are daily concerns. Collaborating closely with system administrators, researchers, and software engineers is essential to balance user demands with cluster stability and security. Proactive monitoring and clear documentation can help manage these challenges effectively.

What are the key skills and qualifications needed to thrive as an HPC Infrastructure Engineer?

To thrive as an HPC Infrastructure Engineer, you need expertise in computer science, Linux system administration, networking, and parallel computing, often supported by a relevant degree or industry certifications. Familiarity with HPC cluster management tools (like Slurm or PBS), scripting languages (such as Python or Bash), and hardware management is essential. Strong problem-solving abilities, communication skills, and teamwork are valuable soft skills in this role. These skills ensure the efficient deployment, maintenance, and optimization of high-performance computing systems crucial for research and enterprise applications.

What is the difference between Hpc Infrastructure Engineer vs Network Engineer?

AspectHpc Infrastructure EngineerNetwork Engineer
Required CredentialsBachelor's in Computer Science, Engineering, or related field; certifications like Cisco CCNA or CompTIA Network+Bachelor's in Computer Science, Engineering, or related field; certifications like Cisco CCNA or CompTIA Network+
Work EnvironmentData centers, research labs, high-performance computing clustersCorporate offices, data centers, network operation centers
Employer & Industry UsageResearch institutions, tech companies, scientific organizationsTelecommunications, IT service providers, large enterprises

While both roles require networking knowledge and certifications, the Hpc Infrastructure Engineer specializes in managing high-performance computing systems and clusters, whereas the Network Engineer focuses on designing and maintaining general network infrastructure. The Hpc role is more research and data-intensive, often in scientific or academic settings, while Network Engineers work across various industries to ensure connectivity and network security.

Are HPC Infrastructure engineers in demand?

HPC Infrastructure engineers are in high demand due to the growing need for powerful computing systems in research, scientific, and enterprise environments. They require expertise in cluster management, networking, and tools like Linux and job schedulers, making their skills highly sought after in industries relying on high-performance computing. Employment opportunities are expected to grow as data processing and computational needs increase across sectors.

How much do HPC Infrastructure engineers make in the US?

HPC Infrastructure engineers in the US typically earn between $80,000 and $130,000 annually, depending on experience, location, and certifications. Senior roles or those with specialized skills in high-performance computing environments can earn higher salaries, often exceeding $150,000.

What are popular job titles related to Hpc Infrastructure Engineer jobs?

For Hpc Infrastructure Engineer jobs, the most frequently searched job titles are:

Infographic showing various Hpc Infrastructure Engineer job openings in the United States as of September 2026, with employment types broken down into 1% Internship, 92% Full Time, 3% Part Time, and 4% Contract. Highlights an 85% Physical, 4% Hybrid, and 11% Remote job distribution, with an average salary of $127,066 per year, or $61.1 per hour.

AI & HPC Infrastructure Engineer

San Jose, CA • On-site

$126K - $165K/yr

Other

Posted 10 days ago


Key responsibilities

  • Build, configure, and operate GPU and HPC clusters across compute, storage, and networking environments.

  • Support capacity planning, performance tuning, and resource optimization for AI workloads.

  • Deploy, automate, and manage compute environments across on-premises and cloud platforms.


Job description

About the Company

Prodapt is the largest specialized player in the Connectedness industry. As an AI-first strategic technology partner, Prodapt provides consulting, business reengineering, and managed services for the largest telecom and tech enterprises building networks and digital experiences of tomorrow. A ServiceNow-invested company, Prodapt has been recognized by Gartner as a Large, Telecom-Native, Regional IT Service Provider. Prodapt’s ASIC Services is a leading provider of SoC/ASIC RTL Design, UVM based verification, Emulation, FPGA based validation, DFT, RTL2GDSII, Physical Design using ICC2 and Innovus, Mask Layout, Firmware, Silicon Bringup, and Analog mask layout. Our embedded services include device drivers, RTOS porting, and board bring-up. A “Great Place To Work® Certified™” company, Prodapt employs over 5,000 technology and domain experts in 30+ countries. Prodapt is part of the 130-year-old business conglomerate The Jhaver Group, which employs over 32,000 people across 80+ locations globally.

We are looking for an AI & HPC Infrastructure Engineer to design, build, and operate the compute platforms that power our AI and High-Performance Computing (HPC) workloads.

In this role, you will deploy, automate, and manage GPU-enabled infrastructure across on-premises and cloud environments, enabling scalable, reliable, and cost-effective compute resources for engineering and R&D teams. You will work at the intersection of AI infrastructure, cloud platforms, automation, and operations to support next-generation AI and engineering applications.


Key Responsibilities

  • Build, configure, and operate GPU and HPC clusters across compute, storage, and networking environments.
  • Support capacity planning, performance tuning, and resource optimization for AI training, inference, and compute-intensive workloads.
  • Monitor infrastructure health and ensure high availability and performance.
  • Deploy and manage compute environments across on-premises and public cloud platforms (AWS, Azure, or GCP).
  • Contribute to infrastructure modernization, scalability, and resiliency initiatives.
  • Support cloud adoption and hybrid computing strategies.
  • Implement Infrastructure-as-Code (IaC) and automation frameworks for provisioning and operations.
  • Develop monitoring, logging, and alerting solutions to improve platform reliability.
  • Drive continuous improvements in operational efficiency and resource utilization.
  • Deploy, integrate, and support AI services, including LLM APIs, coding assistants, and AI/agent platforms.
  • Collaborate with engineering teams to enable AI-driven development workflows.
  • Support AI/ML infrastructure requirements and best practices.
  • Troubleshoot and resolve infrastructure, networking, and platform issues.
  • Create and maintain technical documentation, standards, and operational runbooks.
  • Partner with engineering, IT, and platform teams to deliver secure and scalable solutions.


Required Qualifications

  • Bachelor's degree in Computer Science, Engineering, or a related technical discipline.
  • 3+ years of hands-on experience in Infrastructure Engineering, Platform Engineering, Cloud Operations, or HPC environments.
  • Strong experience with Linux-based infrastructure administration.
  • Experience working with public cloud platforms such as AWS, Azure, or GCP.
  • Hands-on experience with GPU/HPC environments and workload orchestration platforms such as Kubernetes or Slurm.
  • Experience with automation, infrastructure provisioning, monitoring, and performance optimization.
  • Solid understanding of compute, storage, networking, virtualization, and container technologies.


Preferred Qualifications

  • Experience supporting AI/ML workloads and GPU-based infrastructure.
  • Knowledge of Kubernetes, Docker, Infrastructure-as-Code tools, and observability platforms.
  • Familiarity with AI platforms, LLM deployment, and modern engineering productivity tools.