1

Slurm Jobs (NOW HIRING)

HPC Platform Engineer

Houston, TX · Hybrid

$70 - $90/hr

Configure and tune job schedulers (Slurm preferred) * Support distributed multi-node workloads (MPI) * Install, upgrade, and maintain HPC infrastructure * Configure and optimize parallel file systems ...

HPC/ML Infrastructure Engineer

San Francisco, CA · On-site

$126K - $166K/yr

Responsibilities : • Lead bringup, administration, and operations on the largest anime AI training cluster. • Serve as the bridge between researchers and GPU machines. • Ensure SLURM jobs are ...

Experience with NVidia, Cray, managing clusters, Slurm, Linux, CUDA, high-speed networks as well as an understanding of the customer's A&A process. Qualifications: * Active Top Secret/Sensitive ...

Staff AI Infrastructure Engineer

Redwood City, CA · On-site

$131K - $172K/yr

  • Retirement

  • PTO

Our clusters run Slurm on Kubernetes infrastructure and support everything from day-to-day AI researcher workflows to multi-node hero training runs at thousands of GPUs. The team works at the ...

HPC/ML Infrastructure Engineer

San Francisco, CA · On-site

$126K - $166K/yr

You'll serve as the bridge between our researchers and the bare GPU machines, helping to make sure that SLURM jobs are running, parallel filesystems are serving, network is transmitting, and that the ...

Staff AI Infrastructure Engineer

Redwood City, CA · On-site

$241K - $331K/yr

  • Retirement

  • PTO

Our clusters run Slurm on Kubernetes infrastructure and support everything from day-to-day AI researcher workflows to multi-node hero training runs at thousands of GPUs. The team works at the ...

Staff AI Infrastructure Engineer

Redwood City, CA · On-site +1

$131K - $172K/yr

  • Retirement

  • PTO

Our clusters run Slurm on Kubernetes infrastructure and support everything from day-to-day AI researcher workflows to multi-node hero training runs at thousands of GPUs. The team works at the ...

HPC/ML Infrastructure Engineer

San Francisco, CA · On-site

$126K - $166K/yr

You'll serve as the bridge between our researchers and the bare GPU machines, helping to make sure that SLURM jobs are running, parallel filesystems are serving, network is transmitting, and that the ...

next page

Showing results 1-20

Slurm information

See salary details

$11K

$106.4K

$212.5K

How much do slurm jobs pay per year?

As of Aug 17, 2026, the average yearly pay for slurm in the United States is $106,432.00, according to ZipRecruiter salary data. Most workers in this role earn between $41,000.00 and $205,000.00 per year, depending on experience, location, and employer.
More about Slurm jobs

What cities are hiring for Slurm jobs?

Cities with the most Slurm job openings:

What states have the most Slurm jobs?

States with the most job openings for Slurm jobs include:

What job categories do people searching Slurm jobs look for?

The top searched job categories for Slurm jobs are:

Infographic showing various Slurm job openings in the United States as of August 2026, with employment types broken down into 99% Full Time, and 1% Contract. Highlights an 84% Physical, 5% Hybrid, and 11% Remote job distribution, with an average salary of $106,432 per year, or $51.2 per hour.

Senior DevOps Engineer - HPC / EDA / SLURM

CYNET SYSTEMS

Rancho Cordova, CA • On-site

$76.80 - $81.80/hr

Other

Medical, Dental, Vision, Life, Retirement

Posted 7 days ago


Job description

Job Overview: Pay Range: $76.80hr - $81.80hr Requirement/Must Have: 5+ years of experience in a DevOps, Platform Engineering, or Linux Systems Engineering role. Hands-on HPC cluster administration experience, including SLURM or equivalent workload managers. Demonstrated experience supporting EDA or scientific computing environments. Strong Ansible automation skills with production-grade playbook and role development. Experience with bare-metal provisioning tools (RackN, Cobbler, or equivalent). Proven ability to plan and execute datacenter or infrastructure migrations with minimal disruption. Familiarity with enterprise Linux identity and authentication stacks (SSSD, LDAP, AD, NIS, Okta). Experience with NetApp or comparable enterprise storage platforms in HPC contexts. Ability to author formal technical documentation (MOPs, runbooks, architecture diagrams). Strong verbal and written communication skills. Ability to communicate effectively with stakeholders. Responsibilities: Support and administer SLURM-based HPC compute environments, including partition configuration and migration planning. Plan and execute HPC/EDA compute and storage infrastructure migrations across datacenters. Develop migration strategies and evaluate implementation options, risks, dependencies, and operational tradeoffs. Author formal Method of Procedure (MOP) documents and runbooks for infrastructure changes and service cutovers. Coordinate cross-functionally with EDA/SPE teams, storage teams, and IDAM to deliver coordinated platform changes. Verify storage volumes, application access, and service continuity following migrations or infrastructure changes. Define HPC storage service tiers and gather performance and capacity requirements for EDA workloads. Administer SUSE Linux Enterprise Server (SLES) 12 and SLES 15 systems in a production HPC environment. Build and maintain custom Linux OS images and installation media using Kiwi NG and related tooling. Enable and maintain bare-metal provisioning workflows via RackN / Digital Rebar Provision for SLES and ESXi deployments. Provision and configure VMware vSphere virtual machines for HPC service workloads. Troubleshoot production issues in Linux HPC operational scripts, services, and system daemons. Develop, maintain, and extend Ansible playbooks and roles for Linux system setup, authentication, and platform configuration. Ensure multi-version Ansible playbook compatibility across SLES 12 and SLES 15. Manage Git repositories and Artifactory artifact storage; migrate large binaries and configuration artifacts out of source control. Contribute GitHub pull requests, conduct code reviews, and manage inner-source infrastructure repositories. Drive production environment changes through change management workflows using ServiceNow. Integrate and configure enterprise identity systems including Okta, Active Directory, LDAP, NIS, VAS, and SSSD for Linux/HPC environments. Audit and reconcile Linux user and group identity data (UID/GID) across multiple directory and authentication domains. Validate authentication methods and access behavior across HPC compute and storage environments. Extend SSSD-based corporate authentication to new compute environments and author corresponding Ansible automation. Assess and implement log management strategies, including evaluation of Client integration for HPC system logs. Investigate and remediate operational issues in production Linux services (VNC, NIS, AutoFS, Zabbix, etc.). Produce technical documentation, architecture diagrams, implementation guides, and end-user instructions in Confluence. Nice to Have: Experience with SUSE Linux Enterprise Server (SLES) 12 and/or 15 in an enterprise environment. Familiarity with RackN / Digital Rebar Provision for bare-metal OS deployment. Hands-on experience with Kiwi NG or similar tools for custom OS image creation. Knowledge of VMware vSphere for HPC support VM provisioning. Experience migrating configuration artifacts and binaries to Artifactory. Background in semiconductor, storage, or high-tech manufacturing IT environments. Skills: SLURM. HPC compute/storage administration. EDA infrastructure. SLES 12. SLES 15. ESXi 8.0. Kiwi NG. Ansible. RackN / Digital Rebar Provision. VMware vSphere. SSSD. Okta. Active Directory. LDAP. NIS. VAS. NetApp SVM. NFS. AutoFS. Git. GitHub. Artifactory. Client . Zabbix. Python. Perl. Bash. ServiceNow. Confluence. Jira. Benefits Our Benefits Include: Medical, Dental, and Vision Insurance 401(k) Retirement Plan Health Savings Account (HSA) Disability Insurance (Short-Term and Long-Term) Life and AD&D Insurance Paid Sick Leave (where required by applicable state or local law) Supplemental Insurance Plans Identity Theft Protection Pet Insurance Employee Wellness Programs Employee Assistance Program (EAP) Career Growth and Professional Development Opportunities Disclaimer: Benefits eligibility, accrual rates, and usage limits may vary based on employment status, length of service, and work location. Paid Sick Leave is provided in strict accordance with applicable state and municipal mandates. Cynet Systems Inc. reserves the right to modify, amend, or terminate any benefit plans at any time in accordance with applicable laws. About Cynet Systems Founded in 2010 and headquartered in the Washington, DC metro area, Cynet Systems Inc. is a leading technology staffing and workforce solutions company serving Fortune 500 companies, government agencies, and enterprise organizations across the United States and Canada. We deliver agile, scalable talent solutions across IT, engineering, life sciences, clinical, and professional staffing, powered by a high-performing recruitment engine operating across North America and Asia. As a nationally and locally certified Minority Business Enterprise (MBE), Cynet Systems is committed to helping organizations build high-performing teams while empowering professionals to grow rewarding careers. Our organization is certified to ISO 9001, ISO 14001, ISO 27001, and SOC 2 Type II standards, reflecting our commitment to quality, security, operational excellence, and customer success.

Cynet Systems logo

About Cynet Systems

Sourced by ZipRecruiter

Cynet Systems Inc is a staffing and recruiting corporation nestled in Ashburn, VA, USA. Established in 2010, the company operates within the Information Technology and Services sector, specializing in providing effective workforce solutions to different business needs, including IT consulting, direct hire, and contract staffing services. Through the years, Cynet Systems has built an impressive portfolio, going beyond borders and expanding its operations internationally in Canada and India. Rooted in its core values of teamwork, leadership, and commitment, Cynet Systems helps businesses unlock their full potential by providing versatile and competent professionals that perfectly align with their needs. Fueled by their unwavering mission to deliver top-tier talent to businesses worldwide, Cynet Systems garnered various recognitions including SIA's fastest-growing staffing firms and Best Place to Work in Virginia for 2019.

Industry

It services

Company size

501 - 1,000 Employees

Headquarters location

Sterling, VA, US

Year founded

2010

Social media