Language Documentation meets Language Technology

Language Documentation meets Language Technology
复制标题

语言文档遇见语言技术

DOI:
--
复制
发表时间:
2015
期刊:
影响因子:
--
通讯作者:
J. Wilbur
J. Wilbur
中科院分区:
--
文献类型:
--
作者:
Rogier Blokland;Marina Fedina;C. Gerstenberger;N. Partanen;Michael Rießler;J. Wilbur

文献摘要

被引文献

相似文献

本文件介绍了Pite Saami、Kola Saami和Izhva Komi语言文件项目正在进行的工作,所有这些项目都使用类似的数据和技术框架,并在乌普萨拉、特罗姆瑟、瑟克特夫卡尔和弗赖堡合作进行。我们的项目记录和注释口语数据,以提供全面的语音语料库作为未来研究和这些濒危和未被描述的乌拉尔语语音社区的数据库。在语言文档中应用语言技术有助于我们创建更系统地注释的语料库,而不是折衷的数据集合。最终,我们的项目所创建的多模态语料库将有助于在未来对这些语言进行科学上有意义的定量研究。
Thepaper describes work-in-progress by the Pite Saami, Kola Saami and Izhva Komi language documentation projects, all of which use similar data and technical frameworks and are carried out collaboratively in Uppsala, Tromso, Syktyvkar and Freiburg. Our projects record and annotate spoken language data in order to provide comprehensive speech corpora as databases for future research on and for these endangered – and under-described – Uralic speech communities. Applying language technology in language documentation helps us to create more systematically annotated corpora, rather than eclectic data collections. Ultimately, the multimodal corpora created by our projects will be useful for scientifically significant quantitative investigations on these languages in the future.