N-gram based paraphrase generator from large text document

N-gram based paraphrase generator from large text document
复制标题

来自大型文本文档的基于 N 元语法的释义生成器

DOI:
--
复制
发表时间:
2016
期刊:
2016 International Conference on Computation System and Information Technology for Sustainable Solutions (CSITSS)
影响因子:
--
通讯作者:
B. Sagar
B. Sagar
中科院分区:
--
文献类型:
--
作者:
Ashwini I. Gadag;B. Sagar

文献摘要

被引文献

相似文献

本文描述了基于n元语法方法的释义生成。N-gram是文本文档中可应用于一系列自然语言处理(NLP)应用的相关单词。候选释义是基于三文法生成的。参考释义(关键短语)是相关释义的集合,其作用类似于用于生成候选释义的训练数据集。释义生成的任务类似于机器翻译,因此我们使用机器翻译评估指标。R-精度评估度量用于找出候选释义和参考释义之间的共同词数。
This paper describes the paraphrase generation based on n-gram approach. N-grams are relevant words of text document that can be applied for a range of Natural Language Processing (NLP) applications. The candidate paraphrases are generated based on trigrams approach. The reference paraphrases (keyphrases) are the set of relevant paraphrases, which acts like training data set for generating candidate paraphrases. The task of paraphrase generation is similar to machine translation; hence we used machine translation evaluation metrics. R-precision evaluation metric is used to find the number of common words between candidate and reference paraphrases.