Scatter/Gather as a Tool for the Navigation of Retrieval Results

Scatter/Gather as a Tool for the Navigation of Retrieval Results
复制标题

分散/聚集作为检索结果导航的工具

DOI:
--
复制
发表时间:
1995
期刊:
影响因子:
--
通讯作者:
Jan O. Pedersen
Jan O. Pedersen
中科院分区:
--
文献类型:
--
作者:
Marti A. Hearst;David R Karger;Jan O. Pedersen

文献摘要

被引文献

相似文献

一个重要的信息访问问题出现时,用户面临着非常大量的文件,已检索响应查询。在本文中,我们将探讨使用一种技术,称为分散/聚集,检索到的文档的大集合的导航。Scatter/Gather将文档动态聚类为语义一致的组,并向用户呈现组的描述性摘要。这些组可以以多种方式使用:识别要使用其他工具仔细阅读的有用文档子集,消除内容不相关的子集,或者选择有希望的文档子集重新聚类为更精细的组。本文介绍了分散/聚集算法,并通过两个例子说明其应用于检索结果。
An important information access problem arises when the user is confronted with a very large number of documents that have been retrieved in response to a query. In this paper we explore the use of a technique, called Scatter/Gather, for the navigation of large collections of retrieved documents. Scatter/Gather clusters the documents into semantically coherent groups on-the-fly and presents descriptive summaries of the groups to the user. These groups can be used in several ways: to identify useful subsets of documents to be perused with other tools, to eliminate subsets whose contents appear nonrelevant, or to select promising document subsets for reclustering into more refined groups. This paper describes the Scatter/Gather algorithm and illustrates its application to retrieval results via two examples.