1

Hpc Cluster Administrator Slurm Jobs (NOW HIRING)

Senior HPC Cluster Engineer

Bothell, WA · On-site

$145K - $209K/yr

Experience running scientific applications on a Slurm HPC cluster. * Experience deploying and ... operating HPC clusters using NVIDIA GPUs. * Experience with high-performance networking fabrics ...

New

Senior HPC Cluster Engineer

Santa Clara, CA · On-site

$122K - $168K/yr

... Cluster Engineer. The role involves designing, deploying, and operating GPU Compute Clusters for ... Slurm, LSF, PBS or K8s. Applied experience with AI/HPC workflows that use MPI and NCCL. • ...

HPC Platform Engineer

Houston, TX · Hybrid

$70 - $90/hr

Administer and support HPC clusters in production environments * Configure and tune job schedulers ... Hands-on HPC cluster administration experience * Linux systems administration experience ...

Senior HPC Cluster Engineer

Santa Clara, CA

$122K - $168K/yr

We are seeking a highly skilled and experienced HPC Cluster Engineer to design, deploy, and operate ... Experience with AI/HPC job schedulers and orchestrators, such as Slurm, LSF, PBS or K8s. Applied ...

Senior HPC Cluster Engineer

Redmond, WA

$117K - $160K/yr

We are seeking a highly skilled and experienced HPC Cluster Engineer to design, deploy, and operate ... Experience with AI/HPC job schedulers and orchestrators, such as Slurm, LSF, PBS or K8s. Applied ...

Senior HPC Cluster Engineer

Santa Clara, CA · On-site

$122K - $168K/yr

We are seeking a highly skilled and experienced HPC Cluster Engineer to design, deploy, and operate ... Experience with AI/HPC job schedulers and orchestrators, such as Slurm, LSF, PBS or K8s. Applied ...

Senior HPC Cluster Engineer

Austin, TX

$103K - $142K/yr

We are seeking a highly skilled and experienced HPC Cluster Engineer to design, deploy, and operate ... Experience with AI/HPC job schedulers and orchestrators, such as Slurm, LSF, PBS or K8s. Applied ...

Required : • Proven experience in HPC cluster design, deployment, and management (compute ... Slurm, PBS, or Grid Engine. • Strong skills in profiling, benchmarking, and optimization ...

$83K - $112K/yr

Configure, manage, and optimize job scheduling software (SLURM) * Install and configure free and ... Provide support for SAS researchers using Penn's PARCC central HPC cluster and partner with the ...

HPC Cluster Administration: Deploy, configure, and manage Linux-based HPC clusters. Monitor node ... Job Scheduling: Administer workload managers/job schedulers (e.g., Slurm, PBS, or LSF) to ...

HPC Cluster Administration: Deploy, configure, and manage Linux-based HPC clusters. Monitor node ... Job Scheduling: Administer workload managers/job schedulers (e.g., Slurm, PBS, or LSF) to ...

What We're Looking For * 3+ years in systems engineering or infrastructure with hands-on GPU or HPC ... Familiarity with NVIDIA tooling (DCGM, UFM, NCCL, CUDA) and scheduling via SLURM or Kubernetes

next page

Showing results 1-20

Hpc Cluster Administrator Slurm information

See salary details

$14

$36

$63

How much do hpc cluster administrator slurm jobs pay per hour?

As of Aug 2, 2026, the average hourly pay for hpc cluster administrator slurm in the United States is $36.33, according to ZipRecruiter salary data. Most workers in this role earn between $24.28 and $50.96 per hour, depending on experience, location, and employer.

What is the difference between Hpc Cluster Administrator Slurm vs Hpc Cluster Engineer?

AspectHpc Cluster Administrator SlurmHpc Cluster Engineer
CredentialsTypically requires Linux certifications, HPC-specific training, and Slurm knowledgeRequires similar Linux certifications, scripting skills, and HPC system understanding
Work EnvironmentFocuses on managing and maintaining HPC clusters using Slurm workload managerDesigns, develops, and optimizes HPC systems and workflows
Industry UsageCommonly employed in research institutions, universities, and labs using SlurmFound in research, scientific computing, and high-performance computing sectors

Hpc Cluster Administrator Slurm primarily manages and maintains HPC clusters with a focus on Slurm workload management, ensuring system stability and job scheduling. In contrast, Hpc Cluster Engineer designs and develops HPC systems, often working on performance optimization and infrastructure development. Both roles require Linux expertise and HPC knowledge but differ in their core responsibilities and focus areas.

What does an HPC Cluster Administrator do, specifically with Slurm?

An HPC (High Performance Computing) Cluster Administrator who specializes in Slurm is responsible for managing, configuring, and maintaining large-scale computer clusters used for scientific and engineering computations. They install and optimize the Slurm workload manager, which schedules and allocates computing resources to users and jobs efficiently. Their duties include monitoring cluster health, troubleshooting issues, managing user access, and updating software and hardware components. Additionally, they help users with job submissions, ensure security protocols are followed, and work to maximize system uptime and performance.

What are some common challenges faced by an HPC Cluster Administrator working with Slurm, and how can they be addressed?

One common challenge HPC Cluster Administrators encounter is efficiently managing resource allocation and job scheduling in environments with diverse workloads. Balancing user demands, optimizing Slurm configurations, and troubleshooting job failures require both technical expertise and strong communication skills. Proactively monitoring system health, keeping Slurm up-to-date, and collaborating closely with researchers and IT teams help address these challenges. Regular training and staying engaged with the Slurm user community also enable administrators to implement best practices and quickly resolve issues.

What are the key skills and qualifications needed to thrive as an HPC Cluster Administrator (Slurm), and why are they important?

To thrive as an HPC Cluster Administrator specializing in Slurm, you need expertise in Linux system administration, networking, and parallel computing, often supported by a degree in computer science or a related field. Familiarity with Slurm workload manager, scripting languages (such as Bash or Python), and configuration management tools like Ansible is essential, and certifications like RHCE can be advantageous. Strong problem-solving skills, attention to detail, and effective communication are important soft skills for collaborating with researchers and troubleshooting complex issues. These abilities ensure the efficient operation, reliability, and scalability of high-performance computing environments vital for research and development.
More about Hpc Cluster Administrator Slurm jobs
What cities are hiring for Hpc Cluster Administrator Slurm jobs? Cities with the most Hpc Cluster Administrator Slurm job openings:
What states have the most Hpc Cluster Administrator Slurm jobs? States with the most job openings for Hpc Cluster Administrator Slurm jobs include:
What job categories do people searching Hpc Cluster Administrator Slurm jobs look for? The top searched job categories for Hpc Cluster Administrator Slurm jobs are:
Infographic showing various Hpc Cluster Administrator Slurm job openings in the United States as of July 2026, with employment types broken down into 50% Full Time, and 50% Contract. Highlights an 100% In-person job distribution, with an average salary of $75,575 per year, or $36.3 per hour.

Senior HPC and AI Cluster Administrator

Accenture Federal Services

Tampa, FL • On-site

$81K - $110K/yr

Other

Re-posted 15 days ago


Accenture Federal Services rating

8.4

Company rating: 8.4 out of 10

Based on 19 frontline employees who took The Breakroom Quiz

58th of 481 rated business services


Job description

AFS is looking for a Senior HPC and AI Cluster Administrator to support software and data solutions for our customers. We are integrating supercomputers and AI clusters based on existing technologies. We are looking for a system administrator to be a key player to enable artificial intelligence and GPU computing solutions. 

You will work with many scientific researchers, developers, and customers to create improved workflows and develop unique solutions. You will interact with HPC, OS, GPU compute, and systems specialist to architect, develop and bring up large scale performance platforms.  

Key Responsibilities: 

  • Design, Deploy, and maintain HPC/AI clusters 
  • Manage AI jobs workflows using various scheduling technology, such as Kubernetes. 
  • Support and maintain continuous integration and delivery pipelines 
  • Troubleshooting and fixing, bottom up from bare metal, operating system, software stack and application level 
  • Support Research, Development and Operational activities. 

Basic Qualifications: 

  • Bachelor's Degree in Computer Science, Engineering, or a related field; or equivalent experience 
  • 5 years of experience in any of the following:
    • Knowledge of HPC and AI solution technologies to include hardware, hypervisors, CPU's and GPU's. 
    • Experience with job scheduling workloads and orchestration tools such as Slurm & K8s 
    • Excellent knowledge of Linux (i.e. Redhat, Ubuntu) networking (Routing, Switching) and internals, ACLs and OS level security protection and common protocols e.g. TCP, DHCP, DNS, etc. 
    • Experience with multiple storage solutions such as Lustre, GPFS, zfs and xfs. Familiarity with newer and emerging storage technologies. 
    • Automation and configuration management tools such as Python, Bash within a Gitops workflows. 
    • Knowledge of Networking Protocols like InfiniBand, Ethernet 
    • Experience with private cloud platforms (for example VMware, Hyper-V, KVM, Openshift and Nutanix) 
    • Familiarity with public cloud computing platforms (e.g. AWS, Azure) 
  • Must possess and maintain required DoD 8140 certifications. 

Ways to stand out from the crowd: 

  • Knowledge of GPU architectures, time-slicing, Multi-instance GPU (MIG) 
  • Experience with container orchestration technologies i.e. Kubernetes, Docker 
  • Experience designing, deploying AI workflow technologies such as Apache Airflow, Prefect, Dagster. 
  • Background with RDMA (InfiniBand or RoCE) fabrics 
  • Experience working in regulated industries and applying compliance requirements (i.e. DISA STIG, CIS etc.) 
  • NVIDIA Certifications (AI Infrastructure, AI Operations, AI networking) 
  • VMWARE Certifications (Certified Professional / Advanced Professional) 

Clearance

  • An active TS/SCI federal security clearance is required

What Accenture Federal Services employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom