A parallel sparse tensor benchmark suite on CPUs and GPUs
A parallel sparse tensor benchmark suite on CPUs and GPUs
复制标题
DOI:
10.1145/3332466.3374513
复制
发表时间:
2020-01
期刊:
影响因子:
--
通讯作者:
Jiajia Li;M. Lakshminarasimhan;Xiaolong Wu;Ang Li;C. Olschanowsky;K. Barker
中科院分区:
文献类型:
--
作者:
Jiajia Li;M. Lakshminarasimhan;Xiaolong Wu;Ang Li;C. Olschanowsky;K. Barker
Tensor computations present significant performance challenges that impact a wide spectrum of applications. Efforts on improving the performance of tensor computations include exploring data layout, execution scheduling, and parallelism in common tensor kernels. This work presents a benchmark suite for arbitrary-order sparse tensor kernels using state-of-the-art tensor formats: coordinate (COO) and hierarchical coordinate (HiCOO). It demonstrates a set of reference tensor kernel implementations and some observations on Intel CPUs and NVIDIA GPUs. The full paper can be referred to at http://arxiv.org/abs/2001.00660.