Management of Metadata in Linguistic Fieldwork: Experience from the ACLA Project

Management of Metadata in Linguistic Fieldwork: Experience from the ACLA Project
复制标题

语言田野调查中的元数据管理:ACLA 项目的经验

DOI:
--
复制
发表时间:
2004
期刊:
International Conference on Language Resources and Evaluation
影响因子:
--
通讯作者:
J. Simpson
J. Simpson
中科院分区:
--
文献类型:
--
作者:
B. Hughes;D. Penton;Steven Bird;C. Bow;Gillian Wigglesworth;P. McConvell;J. Simpson

文献摘要

被引文献

相似文献

许多语言学研究项目以数字格式收集大量多模态数据。尽管有过多的数据收集应用程序,它往往是很难的研究人员,以确定和整合的应用程序,使收集的多模态数据的管理,除了促进实际的收集过程本身。在涉及大量数据分析的研究项目中,数据管理成为一个关键问题。虽然通过EMELD、HRELP和DOBES等项目传播了关于数据格式本身的最佳做法建议,但除了拉加办和IMDI等实体提供的标准之外,关于外地元数据管理最佳做法的相应信息很少。这些普遍的问题进一步加剧了多个研究人员在地理上不同或连通性受到挑战的位置的背景下。我们描述了一组研究人员在澳大利亚土著社区收集儿童语言习得数据的解决方案的设计。我们描述上下文,确定相关问题,概述解决方案的机制,最后报告实现情况。在这样做的时候,我们提供了一个替代模型和一个开源软件应用程序套件,其目的是足够普遍,其他研究小组可能会考虑采用部分或全部的基础设施。
Many linguistic research projects collect large amounts of multimodal data in digital formats. Despite the plethora of data collection applications available, it is often difficult for researchers to identify and integrate applications which enable the management of collections of multimodal data in addition to facilitating the actual collection process itself. In research projects that involve substantial data analysis, data management becomes a critical issue. Whilst best practice recommendations in regard to data formats themselves are propagated through projects such as EMELD, HRELP and DOBES, there is little corresponding information available regarding best practice for field metadata management beyond the provision of standards by entities such as OLAC and IMDI. These general problems are further exacerbated in the context of multiple researchers in geographically-disparate or connectivity-challenged locations. We describe the design of a solution for a group of researchers collecting data on child language acquisition in Australian indigenous communities. We describe the context, identify pertinent issues, outline the mechanics of a solution, and finally report the implementation. In doing so, we provide an alternative model and an open source software application suite which aims to be sufficiently general that other research groups may consider adopting some or all of the infrastructure.