Answering Frequent Probabilistic Inference Queries in Databases
Answering Frequent Probabilistic Inference Queries in Databases
复制标题
DOI:
10.1109/tkde.2010.146
复制
发表时间:
2011-04
影响因子:
8.9
通讯作者:
Shaoxu Song;Lei Chen;J. Yu
中科院分区:
文献类型:
--
作者:
Shaoxu Song;Lei Chen;J. Yu
Existing solutions for probabilistic inference queries mainly focus on answering a single inference query, but seldom address the issues of efficiently returning results for a sequence of frequent queries, which is more popular and practical in many real applications. In this paper, we mainly study the computation caching and sharing among a sequence of inference queries in databases. The clique tree propagation (CTP) algorithm is first introduced in databases for probabilistic inference queries. We use the materialized views to cache the intermediate results of the previous inference queries, which might be shared with the following queries, and consequently reduce the time cost. Moreover, we take the query workload into account to identify the frequently queried variables. To optimize probabilistic inference queries with CTP, we cache these frequent query variables into the materialized views to maximize the reuse. Due to the existence of different query plans, we present heuristics to estimate costs and select the optimal query plan. Finally, we present the experimental evaluation in relational databases to illustrate the validity and superiority of our approaches in answering frequent probabilistic inference queries.