Building a Diverse Document Leads Corpus Annotated with Semantic Relations
Building a Diverse Document Leads Corpus Annotated with Semantic Relations
复制标题
构建多样化的文档导致语料库带有语义关系注释
DOI:
--
复制
发表时间:
2012
期刊:
影响因子:
--
通讯作者:
S. Kurohashi
中科院分区:
文献类型:
--
作者:
Masatsugu Hangyo;Daisuke Kawahara;S. Kurohashi
In these days, semantic analysis has been actively studied in natural language processing. For the study of semantic analysis, corpora with semantic annotations are essential. Although there are such corpora annotated on newspaper articles, there are various genres and styles, including linguistic expressions that are not found in newspaper articles. In this paper, we build a diverse document leads corpus annotated with semantic relations. To reduce the workload of annotators and annotate as many various documents as possible, we restrict the annotation target of each document to only the first three sentences. We have completed building a corpus of 1,000 documents and report the statistics of this corpus.