CUDASA: Compute Unified Device and Systems Architecture

CUDASA: Compute Unified Device and Systems Architecture
复制标题

CUDASA:计算统一设备和系统架构

DOI:
--
复制
发表时间:
2008
期刊:
EGPGV@Eurographics
影响因子:
--
通讯作者:
T. Ertl
T. Ertl
中科院分区:
--
文献类型:
--
作者:
M. Strengert;C. Müller;Carsten Dachsbacher;T. Ertl

文献摘要

被引文献

相似文献

我们提出了CUDA编程语言的扩展名,该语言将并行性扩展到多GPU系统和GPU群集环境。遵循现有的模型,该模型揭示了GPU的内部并行性,我们的扩展编程语言为从总线和网络互连的额外,更高级别的并行抽象提供了一致的开发接口。新引入的图层提供了当前图形硬件的体系结构和可编程性的特定关键功能,而基础通信和调度机制完全隐藏在用户中。原始编程语言的所有扩展都由一个独立的编译器处理,该编译器很容易嵌入CUDA编译过程中。我们使用两个不同的示例应用程序评估我们的系统,并讨论不同系统体系结构上的缩放行为和性能。
We present an extension to the CUDA programming language which extends parallelism to multi-GPU systems and GPU-cluster environments. Following the existing model, which exposes the internal parallelism of GPUs, our extended programming language provides a consistent development interface for additional, higher levels of parallel abstraction from the bus and network interconnects. The newly introduced layers provide the key features specific to the architecture and programmability of current graphics hardware while the underlying communica- tion and scheduling mechanisms are completely hidden from the user. All extensions to the original programming language are handled by a self-contained compiler which is easily embedded into the CUDA compile process. We evaluate our system using two different sample applications and discuss scaling behavior and performance on different system architectures.