Scalable Node Allocation for Improved Performance in Regular and Anisotropic 3D Torus Supercomputers
Scalable Node Allocation for Improved Performance in Regular and Anisotropic 3D Torus Supercomputers
复制标题
可扩展的节点分配可提高常规和各向异性 3D Torus 超级计算机的性能
DOI:
10.1007/978-3-642-24449-0_9
复制
发表时间:
2011
期刊:
影响因子:
--
通讯作者:
H. Mills
中科院分区:
文献类型:
--
作者:
Carl Albing;N. Troullier;Stephen Whalen;R. Olson;Joe Glenski;H. Pritchard;H. Mills
MPI application performance can vary based on the scheduler’s placing of ranks, whether between nodes or on cores in the same multi-core chip. MPI applications, by default, are at the mercy of the application placement software decision that assigns nodes to a job. We describe herein the general approach of node ordering for allocation in a 3D torus, how it improved MPI application performance, even in the face of an anisotropic interconnect. We demonstrate, quantitatively, that our topologically-based ordering results in improved performance for several MPI applications running on a Top10 supercomputer.