Computing Reliability for Coreference Annotation
Computing Reliability for Coreference Annotation
复制标题
共指注释的可靠性计算
DOI:
10.7916/d8w09fct
复制
发表时间:
2004
期刊:
影响因子:
--
通讯作者:
R. Passonneau
中科院分区:
文献类型:
--
作者:
R. Passonneau
Coreference annotation is annotation of language corpora to indicate which expressions have been used to co-specify the same discourse entity. When annotations of the same data are collected from two or more coders, the reliability of the data may need to be quanti(cid:2)ed. Two obstacles have stood in the way of applying reliability metrics: incommensurate units across annotations, and lack of a convenient representation of the coding values. Given N coders and M coding units, reliability is computed from an N-by-M matrix that records the value assigned to unit M j by coder N k . The solution I present accommodates a wide range of coding choices for the annotator, while preserving the same units across codings. As a consequence, it permits a straightforward application of reliability measurement. In addition, in coreference annotation, disagreements can be complete or partial so I incorporate a distance metric to scale disagreements. This method has also been applied to a quite distinct coding task, namely semantic annotation of summaries.