Performance Analysis of Wavefront Algorithms on Very-Large Scale Distributed Systems

Performance Analysis of Wavefront Algorithms on Very-Large Scale Distributed Systems
复制标题

超大规模分布式系统上波前算法的性能分析

DOI:
--
复制
发表时间:
1998
期刊:
Wide Area Networks and High Performance Computing
影响因子:
--
通讯作者:
H. Wasserman
H. Wasserman
中科院分区:
--
文献类型:
--
作者:
A. Hoisie;O. Lubeck;H. Wasserman

文献摘要

被引文献

相似文献

我们提出了一个算法的并行性能模型,该模型由在消息传递环境中实现的并发二维波前组成。该模型结合了计算和通信波前的单独贡献。我们在三个重要的超级计算机系统(最多 500 个处理器)上验证了该模型。我们使用来自 ASCI 工作负载的确定性粒子传输应用程序的数据,尽管该模型对于在 2-D 处理器域上实现的任何波前算法都是通用的。我们还使用经过验证的模型来估计 100-TFLOPS 计算机系统上的波前算法的性能和可扩展性,这些计算机系统预计将在未来十年内作为 ASCI 计划和其他项目的一部分存在。在此类机器上,我们的分析表明,与传统观点相反,处理器间通信性能并不是瓶颈。单节点效率是主导因素。
We present a model for the parallel performance of algorithms that consist of concurrent, two-dimensional wavefronts implemented in a message passing environment. The model combines the separate contributions of computation and communication wavefronts. We validate the model on three important supercomputer systems, on up to 500 processors. We use data from a deterministic particle transport application taken from the ASCI workload, although the model is general to any wavefront algorithm implemented on a 2-D processor domain. We also use the validated model to make estimates of performance and scalability of wavefront algorithms on 100-TFLOPS computer systems expected to be in existence within the next decade as part of the ASCI program and elsewhere. On such machines our analysis shows that, contrary to conventional wisdom, inter-processor communication performance is not the bottleneck. Single-node efficiency is the dominant factor.