Page Sets as Web Search Answers

Page Sets as Web Search Answers
复制标题

作为 Web 搜索答案的页面集

DOI:
10.1007/11931584_27
复制
发表时间:
2006
期刊:
International Conference on Asian Digital Libraries
影响因子:
--
通讯作者:
Katsumi Tanaka
Katsumi Tanaka
中科院分区:
--
文献类型:
--
作者:
T. Yumoto;Katsumi Tanaka

文献摘要

被引文献

相似文献

在阅读网页或编辑文字处理文档时,我们经常使用页面或文档中的术语作为查询的一部分来搜索Web。因此,搜索的目的与正在阅读或编辑的文档之间存在关联。因此,修改查询以反映该目的可以提高搜索结果的相关性。已经多次尝试从搜索项周围的文本中提取关键字并将它们添加到初始查询中。然而,识别适当的附加关键字是困难的;此外,现有方法依赖于预计算的领域知识。我们已经开发了上下文匹配器:一种查询修改方法,它使用初始搜索结果中搜索项周围的文本以及正在阅读或编辑的文档中搜索项周围的文本,即“源文档”。它使用初始结果中搜索项周围的文本来加权源文档中的候选关键字,以便在查询修改中使用。实验表明,与仅在源文档或搜索结果中使用上下文的基线方法相比,我们的方法经常发现与源文档更相关的文档。
When reading a Web page or editing a word processing document, we often search the Web by using a term on the page or in the document as part of a query. There is thus a correlation between the purpose for the search and the document being read or edited. Modifying the query to reflect this purpose can thus improve the relevance of the search results. There have been several attempts to extract keywords from the text surrounding the search term and add them to the initial query. However, identifying appropriate additional keywords is difficult; moreover, existing methods rely on precomputed domain knowledge. We have developed Context Matcher: a query modification method that uses the text surrounding the search term in the initial search results as well as the text surrounding the term in the document being read or edited, the “source document”. It uses the text surrounding the search term in the initial results to weight candidate keywords in the source document for use in query modification. Experiments showed that our method often found documents more related to the source document than baseline methods that use context either in only the source document or search results.