Distributed Skyline Computation of Vertically Splitted Databases by Using MapReduce

Distributed Skyline Computation of Vertically Splitted Databases by Using MapReduce
复制标题

DOI:
10.1007/978-3-662-43984-5_3
复制
发表时间:
2014-04
期刊:
--
影响因子:
--
通讯作者:
M. A. Siddique;Hao Tian;Y. Morimoto
M. A. Siddique;Hao Tian;Y. Morimoto
中科院分区:
其他
文献类型:
--
作者:
M. A. Siddique;Hao Tian;Y. Morimoto

文献摘要

被引文献

相似文献

Skyline查询检索不受另一个对象支配的对象。天际线查询的结果相对较小,不包含较不重要的对象,并且对于选择对象是有用的。在MapReduce框架下,提出了一种Skyline查询的计算方法,MapReduce框架是大数据分析中的事实标准。目前,我们必须意识到数据泄露。因此,我们提出了一种分布式计算方法,其中每台计算机只使用一个投影数据库,是垂直分裂从原始数据库,计算天际线查询。由于一台计算机只能看到投影值,除了MapReduce的效率优势外,数据库中的敏感信息可以在所提出的方法中本地化。大量的实验表明,该算法的效率合成数据集。
Skyline query retrieve objects that are not dominated by another object. A result of a skyline query is relatively small, does not contain less important objects, and is useful for selecting an object. In this paper, we consider a method for computing skyline query in MapReduce framework, which is a de facto standard in big data analysis. Currently, we have to be aware of data disclosure. Therefore, we propose a distributed computation method, in which each computer uses only a projected database that is vertically splitted from an original database, for computing skyline query. Since one computer can see only projected values, sensitive information in a database can be localized in the proposed method in addition to the advantage of the efficiency of MapReduce. Extensive experiments demonstrate the efficiency of proposed algorithm for synthetic datasets.