1

Hpc System Engineer Jobs (NOW HIRING)

AI/HPC System Architect

San Jose, CA · On-site

$155K - $255K/yr

AI/HPC System Architect Office Location: San Jose, CA Job Type: Full-Time Work Model: Onsite About ... D in Electrical and Computer Engineering, or related field with 15+ years of experience in system ...

AI/HPC System Architect Office Location: San Jose, CA Job Type: Full-Time Work Model: Onsite About ... D in Electrical and Computer Engineering, or related field with 15+ years of experience in system ...

$197K - $214K/yr

Analyzes system requirements and leads design and development activities. Guides users in ... Provides system engineering services and supports the HPC System Design & Engineering organization ...

Bachelor's, Master's, or PhD degree in Computer Science, Electrical Engineering, or a related field * Extensive experience (typically 10+ years) in HPC architecture, system design, or a similar role ...

System Software Engineer (HPC)

Milpitas, CA · On-site

$197K - $233K/yr

KLA is a global leader in diversified electronics for the semiconductor manufacturing ecosystem, and they are seeking a System Software Engineer to join their HPC system software engineering team.

HPC Systems Architect

Chicago, IL · On-site

$200K - $225K/yr

Bachelor's, Master's, or PhD degree in Computer Science, Electrical Engineering, or a related field * Extensive experience (typically 10+ years) in HPC architecture, system design, or a similar role ...

HPC Software Engineer III - UPDATED

Green Bank, WV · On-site +1

$46.75 - $63/hr

This includes building modern, scalable data-processing systems in partnership with leading high ... Work with HPC system engineers to tune application performance for specific architectures.

Required : • Expertise in parallel programming models (e.g., MPI, OpenMP) • Strong knowledge of HPC system architectures and workflows • Experience with distributed systems and data-intensive ...

Required : • Expertise in parallel programming models (e.g., MPI, OpenMP) • Strong knowledge of HPC system architectures and workflows • Experience with distributed systems and data-intensive ...

Showing results 21-40

HPC System Engineer information

See salary details

$53.5K

$127.2K

$167K

How much do hpc system engineer jobs pay per year?

As of Aug 16, 2026, the average yearly pay for hpc system engineer in the United States is $127,215.00, according to ZipRecruiter salary data. Most workers in this role earn between $98,000.00 and $157,000.00 per year, depending on experience, location, and employer.

What are some common challenges faced by HPC system engineers?

HPC System Engineers often encounter challenges related to managing large-scale clusters, troubleshooting performance bottlenecks, and ensuring system reliability under demanding workloads. Keeping up with evolving hardware, software updates, and security requirements is also a key part of the job. The role frequently involves responding to urgent issues, supporting a variety of users with different computational needs, and balancing maintenance with ongoing project deadlines. Successfully navigating these challenges requires both strong technical troubleshooting skills and the ability to communicate solutions effectively with researchers and IT peers.

What is an HPC system engineer?

An HPC (High-Performance Computing) System Engineer designs, deploys, and manages supercomputing environments used for complex computations. They optimize hardware and software components, ensuring system performance, scalability, and reliability. Responsibilities include configuring clusters, troubleshooting performance issues, and maintaining parallel file systems. They work with researchers and developers to optimize code for maximum efficiency. Strong knowledge of Linux, networking, and parallel computing is essential for this role.

What are the key skills and qualifications needed to thrive as an HPC system engineer?

Excelling as an HPC System Engineer requires strong expertise in Linux systems administration, parallel computing, and networking, often supported by a degree in computer science or a related field. Familiarity with HPC resource managers (such as Slurm or PBS), file systems like Lustre or GPFS, and certifications like CompTIA Linux+ or RHCE are highly valuable. Effective problem-solving, teamwork, and communication skills help engineers address complex technical issues and interact with diverse research and engineering teams. These competencies are essential to ensure optimized system performance and support for high-demand computational workloads.

More about HPC System Engineer jobs

What are the most commonly searched types of Hpc System Engineer jobs?

The most popular types of Hpc System Engineer jobs are:

Infographic showing various Hpc System Engineer job openings in the United States as of August 2026, with employment types broken down into 1% As Needed, 84% Full Time, 12% Part Time, and 3% Contract. Highlights an 92% Physical, 2% Hybrid, and 6% Remote job distribution, with an average salary of $127,215 per year, or $61.2 per hour.

CAE HPC System Administrator

Toyota Tsusho Systems

Saline, MI • On-site

Contractor

Re-posted 23 days ago


Job description

About TTS-US:

Founded in 2011, Toyota Tsusho Systems US, Inc. (TTS-US) is a Toyota group company, that develops IT solutions wherever global businesses operate. Transforming into a technology and mobility company, TTS-US, with its 8 TTS affiliates worldwide is establishing a secure and resilient Toyota global value chain. The creative capacity to forge such limitless business opportunities is one of the strengths of Toyota Tsusho Systems."                                                                                                                          

Position Summary:                                                                                                               

We are seeking a highly motivated and experienced CAE HPC System Administrator with more than 7 years of experience to join our dynamic digital solution manufacturing team. This position is ideal for a candidate with strong Linux system administration experience and hands-on expertise managing HPC environments for CAE workloads, including job schedulers, system automation, and engineering application support. The successful candidate will be responsible for administering and optimizing HPC clusters, managing job scheduling systems, supporting CAE applications and licensing, automating Linux operations, maintaining infrastructure performance, and ensuring system stability, scalability, and efficient workload execution.                                                                                 

Requirements

Essential Functions:                                                                                                                         

  1. HPC Job Queuing & Workload Management
    • Administer, configure, and optimize HPC job scheduling environments, including IBM Spectrum LSF, Open PBS,  or equivalent schedulers.
    • Design and tune job queues, resource allocation policies, and scheduling strategies to support diverse CAE workloads.
    • Monitor system performance and utilization trends and implement improvements to maximize efficiency and throughput.
  2. CAE Application and Licensing Support
    • Install, upgrade, test, and support CAE applications and simulation tools in production environments.
    • Provide integration support between CAE applications and HPC scheduling systems.
    • Manage CAE software licensing systems (e.g., FlexLM, RLM) and ensure availability.
    • Troubleshoot application-related issues and ensure minimal disruption to engineering activities.
  3. Linux Systems Administration & Automation
    • Administer and maintain Red Hat Enterprise Linux (RHEL) environments across HPC clusters.
    • Perform OS provisioning, deployment, and patch management using automated tools (e.g., PXE, or configuration management solutions).
    • Develop and maintain scripts (Bash, Korn shell, C Shell, Perl, Awk, or equivalent) to automate system monitoring, health checks, and routine administrative tasks.
    • Maintain system logs, monitoring processes, and standard operating procedures.
  4. Hardware & Infrastructure Management
    • Troubleshoot and resolve issues related to servers, storage systems, and high-performance networking (e.g., InfiniBand, high-speed Ethernet).
    • Support hardware lifecycle activities including installation, maintenance, and upgrades.
    • Conduct capacity planning based on system utilization trends and future demand.
  5. Operations, Monitoring & Continuous Improvement
    • Perform system health checks, monitoring, and incident tracking for HPC and CAE environments.
    • Document system configurations, procedures, incidents, and best practices.
    • Track outages, analyze root causes, and implement preventive measures.
    • Follow change management processes for system updates and deployments.
    • Provide accurate reporting (e.g., utilization, incidents, system performance) and support project initiatives.

Minimum qualifications:                                                                                                    

Required Education & Experience:

7+ years of Linux system administration experience (preferably RHEL environments).

Bachelor's degree in mechanical engineering, electrical engineering, computer engineering, computer science, or related field; and/or commensurate work experience

Hands-on experience managing HPC clusters and job schedulers (LSF, Slurm, PBS, or similar).

Proven experience in CAE application support and integration.

Strong scripting skills (Bash, Shell, Perl, or equivalent).

Experience with OS deployment, patching, and system automation.

Solid understanding of enterprise server hardware, storage, and networking fundamentals.

Experience with CAE tools such as Ansys, LS-DYNA, Nastran, or similar.

Familiarity with high-performance networking technologies is plus (e.g., InfiniBand).

Experience developing internal tools or dashboards are plus (e.g., PHP or web-based tooling).

Position Type/Expected Hours of Work:

Hybrid Full-time contract: Standard business hours with flexibility required to support maintenance windows and critical production issues.

Occasional after-hours or weekend work may be required based on business needs.

Benefits