1

Slurm Jobs in Colorado (NOW HIRING)

... e.g., Slurm, PBS, Grid Engine). • Monitor system health, troubleshoot issues, and resolve performance bottlenecks. • Ensure optimal configuration and high availability of HPC resources. • ...

Experience with container orchestration (e.g., Kubernetes), workload management (e.g., Slurm, Terraform), and monitoring tools (e.g., Grafana). * Public Cloud Knowledge: Familiarity with other public ...

Senior Staff Solutions Engineer (NYC)

Denver, CO · On-site

$56.75 - $73.25/hr

Exposure to Slurm, but with a primary focus on containerized MLOps over traditional HPC * Multi-cloud deployment or migration experience (especially AWS ➝ Crusoe transitions) * Content ...

Cloud Support Engineer

Denver, CO · On-site

$57.50 - $76.75/hr

Experience with container orchestration (e.g., Kubernetes), workload management (e.g., Slurm, Terraform), and monitoring tools (e.g., Grafana). * Public Cloud Knowledge: Familiarity with other public ...

Experience with container orchestration (e.g., Kubernetes), workload management (e.g., Slurm, Terraform), and monitoring tools (e.g., Grafana). * Public Cloud Knowledge: Familiarity with other public ...

Senior Staff Solutions Engineer (NYC)

Denver, CO · On-site

$56.75 - $73.25/hr

Exposure to Slurm, but with a primary focus on containerized MLOps over traditional HPC * Multi-cloud deployment or migration experience (especially AWS ➝ Crusoe transitions) * Content ...

Cloud Support Engineer

Denver, CO · On-site

$57.50 - $76.75/hr

Experience with container orchestration (e.g., Kubernetes), workload management (e.g., Slurm, Terraform), and monitoring tools (e.g., Grafana). * Public Cloud Knowledge: Familiarity with other public ...

Slurm information

How to see Slurm jobs?

To see Slurm jobs, use the command 'squeue' to display queued and running jobs, or 'sacct' for accounting information on completed jobs. These commands help system administrators and users monitor job status in a high-performance computing environment. Proper permissions and environment modules are often required to access job details.

How many jobs can Slurm handle?

Slurm is a workload manager used in high-performance computing environments, capable of handling thousands to hundreds of thousands of jobs simultaneously depending on the system's hardware and configuration. Its scalability allows efficient scheduling and management of large job queues in supercomputing clusters. Proper system tuning and resource allocation are essential for optimal performance when managing large job volumes.

Is Slurm still used?

Slurm is a widely used open-source workload manager for high-performance computing clusters. It is actively maintained and commonly employed in research, scientific, and enterprise environments to schedule and manage jobs efficiently.

What is a Slurm job?

A Slurm job refers to a task or set of tasks submitted to a Slurm workload manager, which schedules and manages compute jobs on high-performance computing clusters. Users submit jobs with specific resource requirements, and Slurm handles job queuing, execution, and monitoring. Knowledge of command-line tools and job scripting is essential for managing Slurm jobs effectively.
What are popular job titles related to Slurm jobs in Colorado? For Slurm jobs in Colorado, the most frequently searched job titles are:
Infographic showing various Slurm job openings in Colorado as of July 2026, with employment types broken down into 94% Full Time, 5% Part Time, and 1% Contract. Highlights an 91% Physical, 5% Hybrid, and 4% Remote job distribution.

Senior HPC Specialist

MDAEdge

Denver, CO • On-site

Full-time

Posted 2 days ago


Job description

Job Summary:
MDAEdge is seeking a highly skilled and experienced Senior HPC Specialist to design, implement, and maintain high-performance computing (HPC) systems and solutions. The ideal candidate will play a critical role in optimizing computational performance and ensuring the reliability of the infrastructure.
Responsibilities:
• Design and deploy HPC clusters, including compute, storage, and networking components.
• Evaluate and integrate new HPC technologies to enhance system performance and scalability.
• Manage Linux-based HPC systems using job schedulers (e.g., Slurm, PBS, Grid Engine).
• Monitor system health, troubleshoot issues, and resolve performance bottlenecks.
• Ensure optimal configuration and high availability of HPC resources.
• Profile and fine-tune applications and workloads for peak performance on HPC systems.
• Analyze job performance and provide recommendations to users for enhancements.
• Administer large-scale parallel file systems (e.g., Lustre, GPFS, BeeGFS).
• Implement data transfer and storage strategies for high-throughput workloads.
• Provide technical support and training to researchers and end users.
• Work with interdisciplinary teams to understand and meet computational requirements.
• Adhere to security best practices and compliance standards for HPC systems.
• Develop and manage backup and disaster recovery solutions.
Qualifications:
Required:
• Proven experience in HPC cluster design, deployment, and management (compute, storage, networking).
• Proficiency in administering Linux systems (RedHat, CentOS, Ubuntu).
• Hands-on experience with job schedulers like Slurm, PBS, or Grid Engine.
• Strong skills in profiling, benchmarking, and optimization techniques.
• Expertise in managing Lustre, GPFS, or BeeGFS.
• Advanced knowledge of scripting languages (Bash, Python, Perl).
• Ability to provide technical documentation, training, and collaboration.
• Knowledge of system hardening, backups, and disaster recovery processes.
Preferred:
• Experience working in interdisciplinary teams to meet diverse computational needs.
• Passion for exploring emerging HPC technologies to enhance capabilities.
Company:
The world doesn't have a talent shortage. It has a talent alignment problem. MDA Edge exists to fix that. Founded in , the company is headquartered in Sheridan, WY, US, , with a team of 51-200 employees. The company is currently Growth Stage.