Error Annotation for Corpus of Japanese Learner English

Error Annotation for Corpus of Japanese Learner English
复制标题

日语学习英语语料库错误注释

DOI:
--
复制
发表时间:
2005
期刊:
--
影响因子:
--
通讯作者:
H. Isahara
H. Isahara
中科院分区:
--
文献类型:
--
作者:
Emi Izumi;Kiyotaka Uchimoto;H. Isahara

文献摘要

被引文献

相似文献

本文通过对学习者语料库中错误标注方法的研究现状的阐述,探讨了如何对学习者语料库进行错误标注。包括我们编制的NICT JLE(Japanese Learner English)语料库在内的几个学习者语料库都标注了错误标记集,这些错误标记集是通过预先对现有规范语法规则或词性系统中隐含的“可能”错误进行分类而设计的。这种错误标注可以帮助成功地评估学习者掌握基本语言系统,特别是语法的程度,但不足以描述学习者的交际能力。为了克服这一局限性,我们重新审视学习者的语言在NICT JLE语料库集中的“可理解性”和“自然性”,并确定目前的错误标记集应如何修改。
In this paper, we discuss how error annotation for learner corpora should be done by explaining the state of the art of error tagging schemes in learner corpus research. Several learner corpora, including the NICT JLE (Japanese Learner English) Corpus that we have compiled are annotated with error tagsets designed by categorizing “likely” errors implied from the existing canonical grammar rules or POS (part-of-speech) system in advance. Such error tagging can help to successfully assess to what extent learners can command the basic language system, especially grammar, but is insufficient for describing learners’ communicative competence. To overcome this limitation, we reexamined learner language in the NICT JLE Corpus by focusing on “intelligibility” and “naturalness”, and determined how the current error tagset should be revised.