ASSASIN: Architecture Support for Stream Computing to Accelerate Computational Storage

ASSASIN: Architecture Support for Stream Computing to Accelerate Computational Storage
复制标题

DOI:
10.1109/micro56248.2022.00035
复制
发表时间:
2022-10
期刊:
2022 55th IEEE/ACM International Symposium on Microarchitecture (MICRO)
影响因子:
--
通讯作者:
Chen Zou;A. Chien
Chen Zou;A. Chien
中科院分区:
其他
文献类型:
--
作者:
Chen Zou;A. Chien

文献摘要

被引文献

相似文献

计算存储在存储设备中添加了计算,从而在卸载,减少数据和较低的能源方面应匹配增长的闪光灯带宽,这又需要高SSD DRAM内存带宽。由SSD的严格的功率和成本约束。对最近的计算SSD研究的调查表明,许多计算存储卸载适合流式计算,以利用这一机会。对流计算的架构支持加速计算存储。即使Flash数据布局不均匀,并且在闪光灯翻译层中保留了页面布局决策的独立性。 SSD Cache-DRAM内存层次结构中的数据,从而解决了内存墙。评估表明,与最先进的计算SSD体系结构相比,Assasin提供了1.5倍-2.4倍的速度方法可以提高2.0倍的功率效率和3.2倍的区域效率,并且在计算SSD的水平上,这些性能效益转化为1.1倍-1.5倍 - 端到端的端到端速度。
Computational storage adds computing to storage devices, providing potential benefits in offload, data-reduction, and lower energy. Successful computational SSD architectures should match growing flash bandwidth, which in turn requires high SSD DRAM memory bandwidth. This creates a memory wall scaling problem, resulting from SSDs’ stringent power and cost constraints.A survey of recent computational SSD research shows that many computational storage offloads are suited to stream computing. To exploit this opportunity, we propose a novel general-purpose computational SSD and core architecture, called ASSASIN (Architecture Support for Stream computing to Accelerate computatIoNal Storage). ASSASIN provides a unified set of compute engines between SSD DRAM and the flash array. This eliminates the SSD DRAM bottleneck by enabling direct computing on flash data streams. ASSASIN further employs a crossbar to achieve performance even when flash data layout is uneven and preserve independence for page layout decisions in the flash translation layer. With stream buffers and scratchpad memories, ASSASIN core’s memory hierarchy and instruction set extensions provide superior low-latency access at low-power and effectively keep streaming flash data out of the in-SSD cache-DRAM memory hierarchy, thereby solving the memory wall.Evaluation shows that ASSASIN delivers 1.5x - 2.4x speedup for offloaded functions compared to state-of-the-art computational SSD architectures. Further, ASSASIN’s streaming approach yields 2.0x power efficiency and 3.2x area efficiency improvement. And these performance benefits at the level of computational SSDs translate to 1.1x - 1.5x end-to-end speedups on data analytics workloads.