The OPS Domain Specific Abstraction for Multi-block Structured Grid Computations

The OPS Domain Specific Abstraction for Multi-block Structured Grid Computations
复制标题

多块结构化网格计算的 OPS 域特定抽象

DOI:
10.1109/wolfhpc.2014.7
复制
发表时间:
2014
期刊:
--
影响因子:
--
通讯作者:
Reguly I
Reguly I
中科院分区:
--
文献类型:
--
作者:
Reguly I

文献摘要

参考文献

被引文献

相似文献

代码可维护性、性能可移植性和面向未来是高性能计算快速变化时代的一些关键挑战。领域特定语言和活动库通过关注单一应用领域并提供高级编程方法,然后使用领域知识在各种硬件上提供高性能来解决这些挑战。在本文中,我们介绍了针对多块结构化网格计算的OPS高级抽象和活动库,并讨论了其一些关键设计点;我们演示了如何将 OPS 嵌入到 C/C++ 中,并使 API 看起来像传统库,以及如何通过简单的文本操作和后端逻辑的组合,我们可以使用不同的并行编程方法在各种硬件上执行。依靠 OPS 抽象的访问执行描述,我们引入了许多自动化执行技术,这些技术可以实现分布式内存并行化、通信优化 模式、检查点和缓存阻塞。使用 Mantevo 基准测试套件中 CloverLeaf 的性能结果,我们展示了 OPS 的实用性。
Code maintainability, performance portability and future proofing are some of the key challenges in this era of rapid change in High Performance Computing. Domain Specific Languages and Active Libraries address these challenges by focusing on a single application domain and providing a high-level programming approach, and then subsequently using domain knowledge to deliver high performance on various hardware.In this paper, we introduce the OPS high-level abstraction and active library aimed at multi-block structured grid computations, and discuss some of its key design points; we demonstrate how OPS can be embedded in C/C++ and the API made to look like a traditional library, and how through a combination of simple text manipulation and back-end logic we can enable execution on a diverse range of hardware using different parallel programming approaches.Relying on the access-execute description of the OPS abstraction, we introduce a number of automated execution techniques that enable distributed memory parallelization, optimization of communication patterns, checkpointing and cache-blocking. Using performance results from CloverLeaf from the Mantevo suite of benchmarks, we demonstrate the utility of OPS.
DOI: --
发表时间: 2014
期刊: ACM SIGPLAN Symposium on Principles & Practice of Parallel Programming
影响因子: --
作者:
P. Balaji;M. Guo;Zhiyi Huang
通讯作者: Zhiyi Huang
使用 OP2 加速全面的工业 CFD 应用
DOI: 10.1109/tpds.2015.2453972
发表时间: 2014
影响因子: 5.3
作者:
I. Reguly;G. Mudalige;C. Bertolli;M. Giles;A. Betts;P. Kelly;David Radford
通讯作者: David Radford
具有运行时 SIMD 并行化的 xeon phi 编程系统
DOI: --
发表时间: 2014
期刊: International Conference on Supercomputing
影响因子: --
作者:
Xin Huo;Bin Ren;G. Agrawal
通讯作者: G. Agrawal
DOI: 10.1093/comjnl/bxr062
发表时间: 2011
期刊: The Computer Journal
影响因子: --
作者:
Giles M
通讯作者: Giles M
平铺作为并行性和数据局部性的持久抽象
DOI: --
发表时间: 2013
期刊:
影响因子: --
作者:
D. Unat;Cy Chan;W. Zhang;J. Bell;J. Shalf
通讯作者: J. Shalf