Fine-Grained Word Sense Disambiguation Based on Parallel Corpora, Word Alignment, Word Clustering and Aligned Wordnets

Fine-Grained Word Sense Disambiguation Based on Parallel Corpora, Word Alignment, Word Clustering and Aligned Wordnets
复制标题

基于并行语料库、词对齐、词聚类和对齐词网的细粒度词义消歧

DOI:
10.3115/1220355.1220547
复制
发表时间:
2004
期刊:
ArXiv
影响因子:
--
通讯作者:
Nancy Ide
Nancy Ide
中科院分区:
--
文献类型:
--
作者:
D. Tufis;Radu Ion;Nancy Ide

文献摘要

被引文献

相似文献

提出一种基于平行语料库的词义消歧方法。该方法利用了词对齐和词聚类方面的最新进展,基于自动提取翻译等价物,并得到语料库中语言的可用对齐词网的支持。根据 EuroWordNet 制定的原则,该词网与普林斯顿词网保持一致。实施本文描述的方法对 WSD 系统的评估显示出非常令人鼓舞的结果。验证模式中使用的相同系统可用于检查和发现多语言对齐词网(如 BalkaNet 和 EuroWordNet)中的对齐错误。
The paper presents a method for word sense disambiguation based on parallel corpora. The method exploits recent advances in word alignment and word clustering based on automatic extraction of translation equivalents and being supported by available aligned wordnets for the languages in the corpus. The wordnets are aligned to the Princeton Wordnet, according to the principles established by EuroWordNet. The evaluation of the WSD system, implementing the method described herein showed very encouraging results. The same system used in a validation mode, can be used to check and spot alignment errors in multilingually aligned wordnets as BalkaNet and EuroWordNet.