Exploiting Cross-Linguistic Similarities in Zulu and Xhosa Computational Morphology

Exploiting Cross-Linguistic Similarities in Zulu and Xhosa Computational Morphology
复制标题

利用祖鲁语和科萨语计算形态学中的跨语言相似性

DOI:
--
复制
发表时间:
2009
期刊:
影响因子:
--
通讯作者:
Sonja E. Bosch
Sonja E. Bosch
中科院分区:
--
文献类型:
--
作者:
L. Pretorius;Sonja E. Bosch

文献摘要

被引文献

相似文献

本文探讨的可能性,跨语言的相似性和相关语言之间的差异提供了自举的形态分析。在这种情况下,现有的祖鲁语形态分析仪原型(ZulMorph)作为科萨语分析仪的基础。调查是围绕所涉及的语言的形态战术和形态音位变化。特别注意的是所谓的“开放”类,它代表了专门名词和动词的词根词典。事实证明,这些词汇的获取和覆盖对于开发中的分析器的成功至关重要。将自举形态分析器应用于平行测试语料库,并对结果进行了讨论。通过语料库中的例子说明了各种跨语言效果。据发现,自举形态分析器表现出显着的结构和词汇相似性的语言可能会卓有成效地利用开发分析器资源较少的语言。
This paper investigates the possibilities that cross-linguistic similarities and dissimilarities between related languages offer in terms of bootstrapping a morphological analyser. In this case an existing Zulu morphological analyser prototype (ZulMorph) serves as basis for a Xhosa analyser. The investigation is structured around the morphotactics and the morphophonological alternations of the languages involved. Special attention is given to the so-called "open" class, which represents the word root lexicons for specifically nouns and verbs. The acquisition and coverage of these lexicons prove to be crucial for the success of the analysers under development. The bootstrapped morphological analyser is applied to parallel test corpora and the results are discussed. A variety of cross-linguistic effects is illustrated with examples from the corpora. It is found that bootstrapping morphological analysers for languages that exhibit significant structural and lexical similarities may be fruitfully exploited for developing analysers for lesser-resourced languages.