Practical performance portability in the Parallel Ocean Program (POP)

Practical performance portability in the Parallel Ocean Program (POP)
复制标题

DOI:
10.1002/cpe.894
复制
发表时间:
2005-08
期刊:
Concurrency and Computation: Practice and Experience
影响因子:
--
通讯作者:
Philip W. Jones;P. Worley;Y. Yoshida;James B. White;J. Levesque
Philip W. Jones;P. Worley;Y. Yoshida;James B. White;J. Levesque
中科院分区:
其他
文献类型:
--
作者:
Philip W. Jones;P. Worley;Y. Yoshida;James B. White;J. Levesque

文献摘要

被引文献

相似文献

本文介绍了并行海洋程序POP的设计,着重讨论了POP的可移植性。POP的性能在各种计算架构上呈现,包括向量架构和商品集群。跨机器的POP性能分析用于表征性能并确定改进,同时保持可移植性。POP模型的一个新的设计,包括缓存阻塞和陆点消除计划,描述了一些初步的性能结果。出版社:John Wiley & Sons,Ltd.
The design of the Parallel Ocean Program (POP) is described with an emphasis on portability. Performance of POP is presented on a wide variety of computational architectures, including vector architectures and commodity clusters. Analysis of POP performance across machines is used to characterize performance and identify improvements while maintaining portability. A new design of the POP model, including a cache blocking and land point elimination scheme, is described with some preliminary performance results. Published in 2005 by John Wiley & Sons, Ltd.