Search NASA⌕ Search

DOE OSTI · 3019946

Evaluating HPC Scheduling Strategies for Urgent Workloads

Abstract

Scientific computing centers increasingly face workloads with diverse urgency requirements, driven by applications that demand rapid or even immediate execution. Appropriately configured scheduling policies can significantly improve both user satisfaction and overall cluster utilization. In this work, we present a systematic analysis of scheduler configurations under scenarios where a fraction of jobs have urgent computing needs. We evaluate multiple job scheduling simulators, develop a lightweight job-submission emulation framework, and create tools to analyze and visualize the resulting scheduling data. Our study identifies key trade-offs between responsiveness, fairness, and efficiency, and offers a set of practical scheduling configurations (particularly for Slurm) that can be tailored to HPC environments supporting mixed-urgency workloads.

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Maheshwari, Ketan [ORNL] (ORCID:000000033800662X), Borch, Anderson [Colorado State University, Fort Collins], Webb, Jordan [ORNL] (ORCID:0000000338877601), Etz, Brian [ORNL] (ORCID:0000000208554863), Miller, Ross [ORNL] (ORCID:000000022179495X), Suter, Fred [ORNL] (ORCID:0000000319021955), Oral, Sarp [ORNL] (ORCID:0000000187457078), Ferreira Da Silva, Rafael [ORNL] (ORCID:0000000217200928). 2025-11-01. Evaluating HPC Scheduling Strategies for Urgent Workloads. https://doi.org/10.1145/3731599.3767474

Cite the original work for its findings. Save a collection to share your selection of sources.