Part-of-speech tagger for Ainu language based on higher order Hidden Markov Model

Part-of-speech tagger for Ainu language based on higher order Hidden Markov Model
复制标题

DOI:
10.1016/j.eswa.2012.04.031
复制
发表时间:
2012-10
期刊:
Expert Syst. Appl.
影响因子:
--
通讯作者:
M. Ptaszynski;Yoshio Momouchi
M. Ptaszynski;Yoshio Momouchi
中科院分区:
其他
文献类型:
--
作者:
M. Ptaszynski;Yoshio Momouchi

文献摘要

相似文献

本文介绍了第一个阿伊努语词性标注器POST-AL。该系统使用了一个基于阿伊努人叙事“yukar”的手工制作的字典。该系统提供三种类型的信息:单词/标记、词性和标记的翻译(日语)。对一套培训材料的评价取得了积极成果。该系统可以在大量的任务有关的阿伊努语言的研究,如内容分析或翻译,到目前为止,主要是手工完成。
This paper presents POST-AL, the first part-of-speech tagger for Ainu language. The system uses a hand-crafted dictionary based on Ainu narratives “yukar”. The system provides three types of information: word/token, part of speech, and translation of the token (in Japanese). Evaluation on a training set provided positive results. The system could be useful in a great number of tasks related to the research on Ainu language, such as content analysis or translation, which till now have been done mostly manually.