1

Hpc Infrastructure Engineer Jobs (NOW HIRING)

Infrastructure Platform Engineer

Oak Ridge, TN · On-site +1

$102K - $134K/yr

The High-Performance Computing Systems Section within a large Research and Development Facility with the Department of Energy is seeking an HPC Infrastructure Platform Engineer to join the HPC ...

Staff Engineer, IT Infrastructure

Menlo Park, CA · On-site

$126K - $166K/yr

Staff Engineer, IT Infrastructure About the Role PacBio is looking for a Staff Engineer, IT ... You will work across Enterprise Linux, HPC and scientific computing, cloud infrastructure ...

Staff Engineer, IT Infrastructure

Menlo Park, CA · On-site

$126K - $166K/yr

Staff Engineer, IT Infrastructure About the Role PacBio is looking for a Staff Engineer, IT ... You will work across Enterprise Linux, HPC and scientific computing, cloud infrastructure ...

$150 - $200/hr

If you prefer, you can work from our offices in London, New York, San Francisco, and Warsaw. #LI-Remote #HPC-Infrastructure-Engineer-GPU-Clusters We are an equal opportunity employer and do not ...

New

HPC Platform Engineer

Houston, TX · Hybrid

$70 - $90/hr

Install, upgrade, and maintain HPC infrastructure * Configure and optimize parallel file systems ... Collaborate with engineering and scientific users * Support HPC environment upgrades and ...

$150 - $200/hr

If you prefer, you can work from our offices in London, New York, San Francisco, and Warsaw. #LI-Remote #HPC-Infrastructure-Engineer-GPU-Clusters We are an equal opportunity employer and do not ...

$150 - $200/hr

If you prefer, you can work from our offices in London, New York, San Francisco, and Warsaw. #LI-Remote #HPC-Infrastructure-Engineer-GPU-Clusters We are an equal opportunity employer and do not ...

New

$150 - $200/hr

Collaborate with cross-functional teams to address AI infrastructure requirements, support AI ... Experience in developingPython basedAI apps and UI * HPC infrastructure engineering for AI/HPC ...

Showing results 41-60

Hpc Infrastructure Engineer information

See salary details

$46.5K

$127.1K

$182K

How much do hpc infrastructure engineer jobs pay per year?

As of Sep 9, 2026, the average yearly pay for hpc infrastructure engineer in the United States is $127,066.00, according to ZipRecruiter salary data. Most workers in this role earn between $107,500.00 and $141,000.00 per year, depending on experience, location, and employer.

What is an HPC Infrastructure Engineer?

HPC Infrastructure Engineers are professionals who design, implement, and manage high-performance computing (HPC) systems and environments. They ensure that computational clusters, storage solutions, and networks are optimized for maximum performance and reliability. These engineers often work closely with researchers, scientists, and IT staff to support complex computing workloads, troubleshoot issues, and maintain system security. Their role is crucial in environments where large-scale simulations, data analysis, or scientific research require significant computational power.

What are some common challenges faced by an HPC Infrastructure Engineer when maintaining high-performance computing clusters?

HPC Infrastructure Engineers often encounter challenges related to scalability, hardware failures, and network bottlenecks when maintaining large computing clusters. Keeping systems up-to-date without causing downtime, troubleshooting complex performance issues, and ensuring efficient resource allocation are daily concerns. Collaborating closely with system administrators, researchers, and software engineers is essential to balance user demands with cluster stability and security. Proactive monitoring and clear documentation can help manage these challenges effectively.

What are the key skills and qualifications needed to thrive as an HPC Infrastructure Engineer?

To thrive as an HPC Infrastructure Engineer, you need expertise in computer science, Linux system administration, networking, and parallel computing, often supported by a relevant degree or industry certifications. Familiarity with HPC cluster management tools (like Slurm or PBS), scripting languages (such as Python or Bash), and hardware management is essential. Strong problem-solving abilities, communication skills, and teamwork are valuable soft skills in this role. These skills ensure the efficient deployment, maintenance, and optimization of high-performance computing systems crucial for research and enterprise applications.

What is the difference between Hpc Infrastructure Engineer vs Network Engineer?

AspectHpc Infrastructure EngineerNetwork Engineer
Required CredentialsBachelor's in Computer Science, Engineering, or related field; certifications like Cisco CCNA or CompTIA Network+Bachelor's in Computer Science, Engineering, or related field; certifications like Cisco CCNA or CompTIA Network+
Work EnvironmentData centers, research labs, high-performance computing clustersCorporate offices, data centers, network operation centers
Employer & Industry UsageResearch institutions, tech companies, scientific organizationsTelecommunications, IT service providers, large enterprises

While both roles require networking knowledge and certifications, the Hpc Infrastructure Engineer specializes in managing high-performance computing systems and clusters, whereas the Network Engineer focuses on designing and maintaining general network infrastructure. The Hpc role is more research and data-intensive, often in scientific or academic settings, while Network Engineers work across various industries to ensure connectivity and network security.

Are HPC Infrastructure engineers in demand?

HPC Infrastructure engineers are in high demand due to the growing need for powerful computing systems in research, scientific, and enterprise environments. They require expertise in cluster management, networking, and tools like Linux and job schedulers, making their skills highly sought after in industries relying on high-performance computing. Employment opportunities are expected to grow as data processing and computational needs increase across sectors.

How much do HPC Infrastructure engineers make in the US?

HPC Infrastructure engineers in the US typically earn between $80,000 and $130,000 annually, depending on experience, location, and certifications. Senior roles or those with specialized skills in high-performance computing environments can earn higher salaries, often exceeding $150,000.

What are popular job titles related to Hpc Infrastructure Engineer jobs?

For Hpc Infrastructure Engineer jobs, the most frequently searched job titles are:

Infographic showing various Hpc Infrastructure Engineer job openings in the United States as of September 2026, with employment types broken down into 1% Internship, 92% Full Time, 3% Part Time, and 4% Contract. Highlights an 85% Physical, 4% Hybrid, and 11% Remote job distribution, with an average salary of $127,066 per year, or $61.1 per hour.

Infrastructure Platform Engineer

Oak Ridge, TN • On-site, Remote

$102K - $134K/yr

Full-time

Posted 27 days ago


Key responsibilities

  • Deploy, configure, and manage HPC-scale services in a Linux environment, primarily Red Hat and Rocky.

  • Build, maintain, and automate internal platforms and tools, including CI/CD pipelines, to enable reliable deployment, monitoring, and scaling of applications.

  • Lead small infrastructure projects, mentor junior staff, and propose improvements to existing systems and processes.


Job description

  • Must be eligible for a federal security clearance (U.S. citizen)
  • On-site or Hybrid preferred (Oak Ridge, TN), but will consider remote 

Overview:  
The High-Performance Computing Systems Section within a large Research and Development Facility with the Department of Energy is seeking an HPC Infrastructure Platform Engineer to join the HPC Infrastructure group.  The preferred candidate will possess commensurate knowledge, skills, and abilities in addition to relevant education, certifications, experience, and demonstrated ability to work as a member of a team.
 
This group provides state-of-the-art computational and data science infrastructure coupled with dedicated technical and scientific professionals tackling large-scale problems across a broad range of scientific domains for accelerating scientific discovery and engineering advances. This group hosts the Oak Ridge Leadership Computing Facility (OLCF), one of the Department of Energy’s (DOE) National User Facilities, which operates Frontier, the nation’s first exascale supercomputer.  
 
Major Duties/Responsibilities:
 
Linux Administration:
  • Deploy, configure, and manage HPC-scale services in a Linux environment, primarily Red Hat and Rocky
  • Perform regular patching, updates, and backups
  • Monitor systems using tools like Nagios and Grafana
  • Respond to and assist in troubleshooting issues
Kubernetes Administration and Automation:
  • Build and maintain foundational internal platforms and tools to enable the HPC Infrastructure team to reliably deploy, monitor, and scale applications
  • Design standardized and automated workflow patterns, build and maintain CI/CD pipelines
  • Offer self-service, excellent documentation, and assistance to HPC Infrastructure group members for efficient consumption of platform services
  • Develop, maintain, and review high-quality code for internal tools using programming languages such as Python, Golang, or Rust
  • Define policies and procedures for automation and configuration management for the team and organization as a whole
Project Management and Leadership:
  • Lead small Infrastructure projects through the project lifecycle
  • Mentor and train junior staff, creating training documentation, holding knowledge sharing sessions, and fostering skill growth throughout the team
  • Propose and implement improvements to existing Infrastructure systems as well as new systems, processes, and procedures
  
Basic Qualifications:
  • Bachelor’s degree in computer science or closely related field and a minimum of 5 years of experience in Linux systems and Kubernetes platform administration, or a master’s degree and a minimum of 4 years of experience in Linux systems and Kubernetes platform administration
  • An equivalent combination of education and experience will be considered
 
Preferred Qualifications:
  • Excellent interpersonal/communication skills and the ability to work within a team
  • Strong experience designing, building, and maintaining Kubernetes platform tools
  • Strong working knowledge of Linux system fundamentals and common network protocols
  • Programming and scripting skills in common languages such as Python and bash
  • Understanding of versioning and code review tools like GitHub and GitLab
  • Experience implementing and supporting highly-available systems and services
  • Experience with configuration management tools such as Puppet or Ansible
  • Experience deploying and maintaining virtual environments using VMware
  • Experience deploying, maintaining, and troubleshooting a variety of infrastructure services such as OpenLDAP, DNS, DHCP, etc.
  • Ability to plan, prioritize, and complete assigned projects with minimal supervision