Accelerating search and recognition workloads with SSE 4.2 string and text processing instructions
Accelerating search and recognition workloads with SSE 4.2 string and text processing instructions
复制标题
使用 SSE 4.2 字符串和文本处理指令加速搜索和识别工作负载
DOI:
10.1109/ispass.2011.5762731
复制
发表时间:
2011
期刊:
影响因子:
--
通讯作者:
Mikko H. Lipasti
中科院分区:
文献类型:
--
作者:
Guangyu Shi;Min Li;Mikko H. Lipasti
Today's information is increasing rapidly, doubling every three years. Consequently, the search and recognition stages in computer applications will consume a growing portion of the total CPU time. The SSE 4.2 instruction set, first implemented in Intel's Core i7, provides string and text processing instructions (STTNI) that utilize SIMD operations for processing character data. Though originally conceived for accelerating string, text, and XML processing, the powerful new capabilities of these instructions are useful outside of these domains, and it is worth revisiting the search and recognition stages of numerous applications to utilize STTNI to improve performance. In this paper, we explored the feasibility and potential benefit of using STTNI to improve the CPU and memory performance of search-and-recognition applications. We optimized four benchmark applications — cache simulation, B+tree search algorithm, template matching, Basic Local Alignment Search Tool (BLAST) — with STTNI, and the new applications outperform their respective original implementations by a factor of 1.4× to 13×.