Discourse Constraints for Document Compression

Discourse Constraints for Document Compression
复制标题

DOI:
10.1162/coli_a_00004
复制
发表时间:
2010-09
影响因子:
9.3
通讯作者:
J. Clarke;Mirella Lapata
J. Clarke;Mirella Lapata
中科院分区:
计算机科学3区
文献类型:
--
作者:
J. Clarke;Mirella Lapata

文献摘要

相似文献

句子压缩在从摘要到字幕生成的许多应用中都有希望。该任务通常在孤立的句子上执行,而不考虑周围的上下文,即使大多数应用程序会在整个文档上运行。在这篇文章中,我们提出了一个话语知情的模型,这是能够产生的文件压缩,是连贯的和翔实的。我们的模型的灵感来自于局部一致性理论,并在整数线性规划的框架内制定。实验结果表明,显着的改进,一个国家的最先进的话语不可知的方法。
Sentence compression holds promise for many applications ranging from summarization to subtitle generation. The task is typically performed on isolated sentences without taking the surrounding context into account, even though most applications would operate over entire documents. In this article we present a discourse-informed model which is capable of producing document compressions that are coherent and informative. Our model is inspired by theories of local coherence and formulated within the framework of integer linear programming. Experimental results show significant improvements over a state-of-the-art discourse agnostic approach.