Scalable Multi-Queue Data Transfer Scheme for FPGA-Based Multi-Accelerators

Scalable Multi-Queue Data Transfer Scheme for FPGA-Based Multi-Accelerators
复制标题

基于 FPGA 的多加速器的可扩展多队列数据传输方案

DOI:
--
复制
发表时间:
2018
期刊:
ICCD
影响因子:
--
通讯作者:
E. Bozorgzadeh
E. Bozorgzadeh
中科院分区:
--
文献类型:
--
作者:
Siavash Rezaei;Kanghee Kim;E. Bozorgzadeh

文献摘要

被引文献

相似文献

在本文中,我们提出了MQMAI,基于多Queue的数据传输机制,该机制可以通过PCIE有效且可扩展对FPGA上的多个加速器的访问。所提出的框架为用户友好的非阻滞API提供了增强的I/O并行性,可访问基于FPGA的加速器。 MQMAI由主机CPU中的I/O软件堆栈和FPGA中的加速器控制器组成。制定了循环策略,以允许通过多个线程/应用程序要求的共享加速器有效地访问运行时。我们的实验结果表明,与当前最新框架RIFFA相比,我们提出的框架可大大降低多加速器访问的总延迟,而每个加速器访问的4个线程。
In this paper, we present MQMAI, multi-queue command-based data transfer mechanism that enables efficient and scalable access to multiple accelerators on FPGA via PCIe. The proposed framework provides a user-friendly non-blocking API with enhanced I/O parallelism to access FPGA-based accelerators. MQMAI is composed of I/O software stack in host CPU and accelerator controller in FPGA. A round-robin policy is developed to allow efficient run-time access to shared accelerators requested through multiple threads/applications. Our experimental results show that our proposed framework reduces the total latency of multi accelerators access significantly compared to the current state-of-the-art framework, RIFFA, by 18x for 4 threads per accelerator access.