Improving GPU Throughput through Parallel Execution Using Tensor Cores and CUDA Cores
Improving GPU Throughput through Parallel Execution Using Tensor Cores and CUDA Cores
复制标题
使用 Tensor 核心和 CUDA 核心通过并行执行提高 GPU 吞吐量
DOI:
10.1109/isvlsi54635.2022.00051
复制
发表时间:
2022
期刊:
影响因子:
--
通讯作者:
Mohanty, Saraju
中科院分区:
文献类型:
--
作者:
Ho, Khoa;Zhao, Hui;Jog, Adwait;Mohanty, Saraju
DOI:
--
发表时间:
2019
期刊:
International Conference on Supercomputing
影响因子:
--
作者:
Xianwei Cheng;Hui Zhao;M. Kandemir;Beilei Jiang;Gayatri Mehta
通讯作者:
Gayatri Mehta
DOI:
10.1145/3466752.3480063
发表时间:
2021-10
期刊:
MICRO-54: 54th Annual IEEE/ACM International Symposium on Microarchitecture
影响因子:
--
作者:
Vijay Kandiah;Scott Peverelle;Mahmoud Khairy;Junrui Pan;Amogh Manjunath;Timothy G. Rogers;Tor M. Aamodt;Nikolaos Hardavellas
通讯作者:
Vijay Kandiah;Scott Peverelle;Mahmoud Khairy;Junrui Pan;Amogh Manjunath;Timothy G. Rogers;Tor M. Aamodt;Nikolaos Hardavellas