A 33 percentage point increase in cluster utilization resulted simply from changing the order of submitted jobs. Researchers at Hugging Face found that specific scheduling sequences reduce hardware idling during large-scale training. This discovery proves that software orchestration often bottlenecks performance more than raw compute. Practitioners should optimize queue logic to maximize GPU efficiency.