Homonymy and Polysemy in Information Retrieval

Homonymy and Polysemy in Information Retrieval
复制标题

信息检索中的同音异义和一词多义

DOI:
--
复制
发表时间:
1997
期刊:
Annual Meeting of the Association for Computational Linguistics
影响因子:
--
通讯作者:
Robert Krovetz
Robert Krovetz
中科院分区:
--
文献类型:
--
作者:
Robert Krovetz

文献摘要

被引文献

相似文献

本文讨论了信息检索系统中词义判别的研究。我们进行了实验,有三个证据来源,使这些区别:形态,词性,和短语。我们集中讨论了同音异义和一词多义(无关意义和相关意义)之间的区别。我们的研究结果支持区分同音异义和一词多义的必要性。我们发现:1)对形态变体进行分组显著提高了检索性能,2)词典中词性不同的所有单词中有一半以上在意义上是相关的,以及3)将信用分配给短语的组成词是至关重要的。这些实验提供了更好的理解基于词的方法,并建议在自然语言处理可以提供检索性能的进一步改善。
This paper discusses research on distinguishing word meanings in the context of information retrieval systems. We conducted experiments with three sources of evidence for making these distinctions: morphology, part-of-speech, and phrases. We have focused on the distinction between homonymy and polysemy (unrelated vs. related meanings). Our results support the need to distinguish homonymy and polysemy. We found: 1) grouping morphological variants makes a significant improvement in retrieval performance, 2) that more than half of all words in a dictionary that differ in part-of-speech are related in meaning, and 3) that it is crucial to assign credit to the component words of a phrase. These experiments provide better understanding of word-based methods, and suggest where natural language processing can provide further improvements in retrieval performance.