1

Slurm Jobs in Colorado (NOW HIRING)

... e.g., Slurm, PBS, Grid Engine). • Monitor system health, troubleshoot issues, and resolve performance bottlenecks. • Ensure optimal configuration and high availability of HPC resources. • ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput, stability, and convergence issues * Build evaluation harnesses and benchmark infrastructure, with held ...

next page

Showing results 1-20

Slurm information

What are popular job titles related to Slurm jobs in Colorado?

For Slurm jobs in Colorado, the most frequently searched job titles are:

Infographic showing various Slurm job openings in Colorado as of August 2026, with employment types broken down into 94% Full Time, 3% Part Time, and 3% Contract. Highlights an 84% Physical, 6% Hybrid, and 10% Remote job distribution.

Senior HPC Specialist

MDAEdge

Denver, CO • On-site

Full-time

Re-posted 13 days ago


Job description

Job Summary:
MDAEdge is seeking a highly skilled and experienced Senior HPC Specialist to design, implement, and maintain high-performance computing (HPC) systems and solutions. The ideal candidate will play a critical role in optimizing computational performance and ensuring the reliability of the infrastructure.
Responsibilities:
• Design and deploy HPC clusters, including compute, storage, and networking components.
• Evaluate and integrate new HPC technologies to enhance system performance and scalability.
• Manage Linux-based HPC systems using job schedulers (e.g., Slurm, PBS, Grid Engine).
• Monitor system health, troubleshoot issues, and resolve performance bottlenecks.
• Ensure optimal configuration and high availability of HPC resources.
• Profile and fine-tune applications and workloads for peak performance on HPC systems.
• Analyze job performance and provide recommendations to users for enhancements.
• Administer large-scale parallel file systems (e.g., Lustre, GPFS, BeeGFS).
• Implement data transfer and storage strategies for high-throughput workloads.
• Provide technical support and training to researchers and end users.
• Work with interdisciplinary teams to understand and meet computational requirements.
• Adhere to security best practices and compliance standards for HPC systems.
• Develop and manage backup and disaster recovery solutions.
Qualifications:
Required:
• Proven experience in HPC cluster design, deployment, and management (compute, storage, networking).
• Proficiency in administering Linux systems (RedHat, CentOS, Ubuntu).
• Hands-on experience with job schedulers like Slurm, PBS, or Grid Engine.
• Strong skills in profiling, benchmarking, and optimization techniques.
• Expertise in managing Lustre, GPFS, or BeeGFS.
• Advanced knowledge of scripting languages (Bash, Python, Perl).
• Ability to provide technical documentation, training, and collaboration.
• Knowledge of system hardening, backups, and disaster recovery processes.
Preferred:
• Experience working in interdisciplinary teams to meet diverse computational needs.
• Passion for exploring emerging HPC technologies to enhance capabilities.
Company:
The world doesn't have a talent shortage. It has a talent alignment problem. MDA Edge exists to fix that. Founded in , the company is headquartered in Sheridan, WY, US, , with a team of 51-200 employees. The company is currently Growth Stage.