Performance evaluation of SDN-enhanced MPI allreduce on a cluster system with fat-tree interconnect

Performance evaluation of SDN-enhanced MPI allreduce on a cluster system with fat-tree interconnect
复制标题

DOI:
10.1109/hpcsim.2014.6903768
复制
发表时间:
2014-07
期刊:
2014 International Conference on High Performance Computing & Simulation (HPCS)
影响因子:
--
通讯作者:
Keichi Takahashi;Khureltulga Dashdavaa;Yasuhiro Watashiba;Y. Kido;S. Date;S. Shimojo
Keichi Takahashi;Khureltulga Dashdavaa;Yasuhiro Watashiba;Y. Kido;S. Date;S. Shimojo
中科院分区:
其他
文献类型:
--
作者:
Keichi Takahashi;Khureltulga Dashdavaa;Yasuhiro Watashiba;Y. Kido;S. Date;S. Shimojo

文献摘要

相似文献

如今,超级计算机在高性能计算中发挥着至关重要的作用。一般来说,现代超级计算机被构建为集群系统,这是一个由多台计算机在网络上互联的系统。在这样的集群系统上编写并行程序时,使用了MPI(消息传递接口)。在本文中,我们的目标是减少MPI ALLREDUE的执行时间,这是一种在许多仿真代码中经常使用的MPI集合通信。为此,我们将软件定义网络的网络可编程性集成到MPI AllReduce中,从而有效地利用了集群系统互连的带宽。在胖树互连的集群系统上进行的实验表明,在OpenMPI实现上,我们提出的MPI ALLREDUE优于MPI ALLREDUE。
Nowadays, supercomputers play an essential role in high-performance computing. In general, modern supercomuputers are built as a cluster system, which is a system of multiple computers interconnected on a network. In coding a parallel program on such a cluster system, MPI (Message Passing Interface) is utilized. In this paper, we aim to reduce the execution time of MPI Allreduce, a frequently used MPI collective communication in many simulation codes. To this end, we have integrated network programmability by Software Defined Networking into MPI Allreduce so that it effectively uses the bandwidth of the interconnect of the cluster system. An experiment conducted on a cluster system with fat-tree interconnect indicates that our proposed MPI Allreduce is superior to MPI Allreduce in OpenMPI implementations.