GPU Fast Convolution via the Overlap-and-Save Method in Shared Memory

GPU Fast Convolution via the Overlap-and-Save Method in Shared Memory
复制标题

通过共享内存中的重叠保存方法进行 GPU 快速卷积

DOI:
10.1145/3394116
复制
发表时间:
2020
影响因子:
1.6
通讯作者:
Adámek K
Adámek K
中科院分区:
计算机科学3区
文献类型:
--
作者:
Adámek K

文献摘要

参考文献

被引文献

相似文献

我们提出了一种实现的和保存的方法,非常长的信号与短的响应函数,这是专门为GPU的卷积的方法。我们已经实现了几种FFT算法(使用CUDA编程语言),它们利用GPU共享内存,允许GPU加速卷积。我们比较我们的实现与实现的的快速和保存算法,利用NVIDIA FFT库(cuFFT)。我们证明,通过使用基于共享内存的FFT,我们可以实现显着的速度为某些问题的大小,并降低GPU上的存储和保存方法的内存要求。
We present an implementation of the overlap-and-save method, a method for the convolution of very long signals with short response functions, which is tailored to GPUs. We have implemented several FFT algorithms (using the CUDA programming language), which exploit GPU shared memory, allowing for GPU accelerated convolution. We compare our implementation with an implementation of the overlap-and-save algorithm utilizing the NVIDIA FFT library (cuFFT). We demonstrate that by using a shared-memory-based FFT, we can achieved significant speed-ups for certain problem sizes and lower the memory requirements of the overlap-and-save method on GPUs.
使用共享内存复用的高效 FFT
DOI: --
发表时间: 2014
期刊: Numerical Computations with GPUs
影响因子: --
作者:
Yi Yang;Huiyang Zhou
通讯作者: Huiyang Zhou
DOI: --
发表时间: 2015
期刊:
影响因子: --
作者:
A. Richards
通讯作者: A. Richards
DOI: --
发表时间: 2018
影响因子: 8.7
作者:
S. Dimoudi;K. Adámek;P. Thiagaraj;S. Ransom;A. Karastergiou;W. Armour
通讯作者: W. Armour
使用X射线相位成像法观察轮岛涂
DOI: --
发表时间: 2019
期刊:
影响因子: --
作者:
宮部さやか;藤井規史;藤本慎司;岡本博之,藤森茜,森川公彦,水野薫
通讯作者: 岡本博之,藤森茜,森川公彦,水野薫
DOI: --
发表时间: 2017
期刊:
影响因子: --
作者:
Karel Ad'amek;S. Dimoudi;M. Giles;W. Armour
通讯作者: W. Armour