Engineering PapersSearch

NASA NTRS · 20000069004

Scheduling for Parallel Supercomputing: A Historical Perspective of Achievable Utilization

Abstract

The NAS facility has operated parallel supercomputers for the past 11 years, including the Intel iPSC/860, Intel Paragon, Thinking Machines CM-5, IBM SP-2, and Cray Origin 2000. Across this wide variety of machine architectures, across a span of 10 years, across a large number of different users, and through thousands of minor configuration and policy changes, the utilization of these machines shows three general trends: (1) scheduling using a naive FIFO first-fit policy results in 40-60% utilization, (2) switching to the more sophisticated dynamic backfilling scheduling algorithm improves utilization by about 15 percentage points (yielding about 70% utilization), and (3) reducing the maximum allowable job size further increases utilization. Most surprising is the consistency of these trends. Over the lifetime of the NAS parallel systems, we made hundreds, perhaps thousands, of small changes to hardware, software, and policy, yet, utilization was affected little. In particular these results show that the goal of achieving near 100% utilization while supporting a real parallel supercomputing workload is unrealistic.

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Jones, James Patton, Nitzberg, Bill. 1999-01-01. Scheduling for Parallel Supercomputing: A Historical Perspective of Achievable Utilization. https://ntrs.nasa.gov/citations/20000069004

Cite the original work for its findings. Save a collection to share your selection of sources.