Efficient Aggregation Query Processing for Large-Scale Multidimensional Data by Combining RDB and KVS
Efficient Aggregation Query Processing for Large-Scale Multidimensional Data by Combining RDB and KVS
复制标题
DOI:
10.1007/978-3-319-98809-2_9
复制
发表时间:
2018-09
期刊:
影响因子:
--
通讯作者:
Y. Watari;Atsushi Keyaki;Jun Miyazaki;Masahide Nakamura
中科院分区:
文献类型:
--
作者:
Y. Watari;Atsushi Keyaki;Jun Miyazaki;Masahide Nakamura
This paper presents a highly efficient aggregation query processing method for large-scale multidimensional data. Recent developments in network technologies have led to the generation of a large amount of multidimensional data, such as sensor data. Aggregation queries play an important role in analyzing such data. Although relational databases (RDBs) support efficient aggregation queries with indexes that enable faster query processing, increasing data size may lead to bottlenecks. On the other hand, the use of a distributed key-value store (D-KVS) is key to obtaining scale-out performance for data insertion throughput. However, querying multidimensional data sometimes requires a full data scan owing to its insufficient support for indexes. The proposed method combines an RDB and D-KVS to use their advantages complementarily. In addition, a novel technique is presented wherein data are divided into several subsets called grids, and the aggregated values for each grid are precomputed. This technique improves query processing performance by reducing the amount of scanned data. We evaluated the efficiency of the proposed method by comparing its performance with current state-of-the-art methods and showed that the proposed method performs better than the current ones in terms of query and insertion.