A Survey on Large Scale Corpora and Emotion Corpora

A Survey on Large Scale Corpora and Emotion Corpora
复制标题

大规模语料库和情感语料库调查

DOI:
10.11185/imt.9.429
复制
发表时间:
2014
期刊:
Information and Media Technologies
影响因子:
--
通讯作者:
Kenji Araki
Kenji Araki
中科院分区:
--
文献类型:
--
作者:
Michal Ptaszynski;Rafal Rzepka;Satoshi Oyama;Masahito Kurihara;Kenji Araki

文献摘要

相似文献

在本文中,我们提出了一个调查自然语言语料库,特别是大规模的语料库和适用于情感分析。自然语言语料库对于训练各种软件工程应用程序至关重要,从词性标记器和依赖关系解析器到对话系统或情感分析软件。我们比较了几个自然语言语料库创建不同的语言,分析其显着的特点和额外的注释提供的开发人员的语料库。
In this paper we present a survey on natural language corpora, with particular focus on corpora of large scale and those applicable to sentiment analysis. Natural language corpora are crucial for training various Software Engineering applications, from part-of-speech taggers and dependency parsers to dialog systems or sentiment analysis software. We compare several natural language corpora created for different languages, analyze their distinctive features and the amount of additional annotations provided by the developers of those corpora.