An empirical resource for discovering cognitive principles of discourse organisation: the ANNODIS corpus
An empirical resource for discovering cognitive principles of discourse organisation: the ANNODIS corpus
复制标题
发现话语组织认知原理的经验资源:ANNODIS 语料库
DOI:
--
复制
发表时间:
2012
期刊:
影响因子:
--
通讯作者:
L. Vieu
中科院分区:
文献类型:
--
作者:
Stergos D. Afantenos;Nicholas Asher;Farah Benamara;M. Bras;Cécile Fabre;Mai Ho;A. L. Draoulec;Philippe Muller;Marie;Laurent Prévot;Josette Rebeyrolle;Ludovic Tanguy;Marianne Vergez;L. Vieu
This paper describes the ANNODIS resource, a discourse-level annotated corpus for French. The corpus combines two perspectives on discourse: a bottom-up approach and a top-down approach. The bottom-up view incrementally builds a structure from elementary discourse units, while the top-down view focuses on the selective annotation of multi-level discourse structures. The corpus is composed of texts that are diversified with respect to genre, length and type of discursive organisation. The methodology followed here involves an iterative design of annotation guidelines in order to reach satisfactory inter-annotator agreement levels. This allows us to raise a few issues relevant for the comparison of such complex objects as discourse structures. The corpus also serves as a source of empirical evidence for discourse theories. We present here two first analyses taking advantage of this new annotated corpus --one that tested hypotheses on constraints governing discourse structure, and another that studied the variations in composition and signalling of multi-level discourse structures.