Review of Arbil: Free Tool for Creating, Editing, and Searching Metadata

Review of Arbil: Free Tool for Creating, Editing, and Searching Metadata
复制标题

Arbil 评论:用于创建、编辑和搜索元数据的免费工具

DOI:
--
复制
发表时间:
2014
期刊:
影响因子:
--
通讯作者:
Rebecca Defina
Rebecca Defina
中科院分区:
--
文献类型:
--
作者:
Rebecca Defina

文献摘要

被引文献

相似文献

1. Arbil2 (Withers 2012)是一个免费的基于java的软件程序,用于创建和编辑元数据。它是由马克斯·普朗克心理语言学研究所语言档案馆的彼得·威瑟斯(Peter Withers)创建的,用于DoBeS3(“濒危语言文献”)项目。Arbil现在由Peter Withers和Twan Goosen领导的团队维护,他们正在为更广泛的用户群开发和扩展它。元数据的创建是任何收集数据的项目的重要组成部分。在这篇综述中,我将重点关注语言描述和文档风格项目以及它们倾向于收集的语言和其他数据的种类。元数据是描述其他数据内容的信息。元数据通常描述数据的类型,例如是视频记录、书面文本还是照片,以及数据的格式,例如MPEG1或MPEG2。它还应该描述收集数据的时间和地点。还应该对所有涉及的人员进行良好的描述。例如,如果你在描述录音,你需要知道谁在录音,谁做的录音,谁做的转录。有关语言元数据的更多信息,请参阅Bowern(2008:56-59)和Thieberger & Berez(2012)。元数据使社区成员和其他研究人员能够发现您拥有的数据类型并找到他们正在寻找的材料。元数据对于作为研究人员的自己也是必要的,它可以保存数据细节的持久记录,并允许您在未来许多年里定位记录和信息。我很高兴我有很好的元数据来记录我五年前的录音。如果没有它,我就会完全迷失,寻找关于某个特定话题的启发课程,或者寻找讲述狗的故事的女人的名字。Arbil不仅是初始创建元数据的好工具,以后还可以使用它来搜索元数据并直接打开相关文件。使用中的语言元数据有几种不同的格式。Arbil旨在以ISLE元数据计划(IMDI)4或组件元数据基础设施(CMDI)5格式创建元数据。这两种格式都是基于xml的元数据模板。它们提供了一个字段列表,其中一些字段是必需的,必须填写才能获得完整的元数据文件,而其他字段是可选的。这对所创建的元数据的类型进行排序和限制。这种结构上的约束
1. OVERVIEW.1 Arbil2 (Withers 2012) is a free Java-based software program for creating and editing metadata. It was created by Peter Withers, of The Language Archive at the Max Planck Institute for Psycholinguistics, for use within the DoBeS3 (Dokumentation bedrohter Sprachen, ‘Documentation of Endangered Languages’) program. Arbil is now being maintained by a team led by Peter Withers and Twan Goosen, who are developing and extending it for a wider user group. The creation of metadata is an important part of any project that collects data. In this review, I will focus on language description and documentation style projects and the kinds of linguistic and other data that they tend to collect. Metadata is information that describes the content of other data. Metadata usually describes the type of data, for instance whether it is a video recording, a written text, or a photo, and the format of the data, for instance MPEG1 or MPEG2. It should also describe when and where the data was collected. There should also be a good description of all the people involved. For instance, if you are describing a transcription you need to know who is speaking in the recording, who made the recording, and who did the transcription. For more information about linguistic metadata see Bowern (2008:56-59) and Thieberger & Berez (2012). Metadata makes it possible for community members and other researchers to find out what kind of data you have and to find material they are looking for. Metadata is also necessary for yourself as a researcher, to keep a lasting record of the details of your data and allow you to locate recordings and information for many years to come. I am very glad that I have good metadata for recordings I made five years ago. Without it I would be completely lost, looking for elicitation sessions on a particular topic or the name of the woman who told that story about the dog. Arbil is not only a good tool for the initial creation of metadata, it can be used later to search your metadata and open the associated files directly. There are several different formats for linguistic metadata in use. Arbil is intended for the creation of metadata in either the ISLE MetaData Initiative (IMDI)4 or the Component MetaData Infrastructure (CMDI)5 format. Both formats are XML-based templates for metadata. They provide a list of fields, some of which are obligatory and must be filled in in order to have a completed metadata file, and others which are optional. This both orders and constrains the type of metadata created. This constraint on the structure of the