An empirical resource for discovering cognitive principles of discourse organisation: the ANNODIS corpus

An empirical resource for discovering cognitive principles of discourse organisation: the ANNODIS corpus
复制标题

发现话语组织认知原理的经验资源:ANNODIS 语料库

DOI:
--
复制
发表时间:
2012
期刊:
International Conference on Language Resources and Evaluation
影响因子:
--
通讯作者:
L. Vieu
L. Vieu
中科院分区:
--
文献类型:
--
作者:
Stergos D. Afantenos;Nicholas Asher;Farah Benamara;M. Bras;Cécile Fabre;Mai Ho;A. L. Draoulec;Philippe Muller;Marie;Laurent Prévot;Josette Rebeyrolle;Ludovic Tanguy;Marianne Vergez;L. Vieu

文献摘要

被引文献

相似文献

本文介绍了ANNODIS资源,法语语篇级注释语料库。该语料库结合了两种语篇分析方法:自下而上的方法和自上而下的方法。自下而上的观点是从基本的语篇单位逐步建立一个结构,而自上而下的观点则侧重于对多层次语篇结构的选择性注释。该语料库由体裁、长度和话语组织类型多样化的文本组成。这里遵循的方法涉及注释指南的迭代设计,以达到令人满意的注释者间一致性水平。这使我们能够提出几个相关的问题,比较这样的复杂对象的话语结构。语料库也是话语理论的经验证据来源。我们在这里提出了两个第一次分析,利用这个新的注释语料库-一个测试的制约话语结构的假设,另一个研究的组成和信号的多层次话语结构的变化。
This paper describes the ANNODIS resource, a discourse-level annotated corpus for French. The corpus combines two perspectives on discourse: a bottom-up approach and a top-down approach. The bottom-up view incrementally builds a structure from elementary discourse units, while the top-down view focuses on the selective annotation of multi-level discourse structures. The corpus is composed of texts that are diversified with respect to genre, length and type of discursive organisation. The methodology followed here involves an iterative design of annotation guidelines in order to reach satisfactory inter-annotator agreement levels. This allows us to raise a few issues relevant for the comparison of such complex objects as discourse structures. The corpus also serves as a source of empirical evidence for discourse theories. We present here two first analyses taking advantage of this new annotated corpus --one that tested hypotheses on constraints governing discourse structure, and another that studied the variations in composition and signalling of multi-level discourse structures.