Representing Linguistic Corpora and Their Annotations

Representing Linguistic Corpora and Their Annotations
复制标题

表示语言语料库及其注释

DOI:
--
复制
发表时间:
2006
期刊:
International Conference on Language Resources and Evaluation
影响因子:
--
通讯作者:
Laurent Romary
Laurent Romary
中科院分区:
--
文献类型:
--
作者:
Nancy Ide;Laurent Romary

文献摘要

被引文献

相似文献

国际标准组织技术委员会 37 语言资源管理小组委员会 (ISO TC37 SC4) 正在开发语言注释框架 (LAF)。 LAF 旨在提供一种标准化的方法来表示语言数据及其注释,该方法的定义足够广泛以适应所有类型的语言注释,同时提供表示精确且可能复杂的语言信息的方法。 LAF 设计的一般原则之前已有报道(Ide 和 Romary,2003;Ide 和 Romary,2004a)。本文介绍了 LAF 设计的一些更具技术性的方面,这些方面已在最终确定标准规范的过程中得到解决。
A Linguistic Annotation Framework (LAF) is being developed within the International Standards Organization Technical Committee 37 Sub-committee on Language Resource Management (ISO TC37 SC4). LAF is intended to provide a standardized means to represent linguistic data and its annotations that is defined broadly enough to accommodate all types of linguistic annotations, and at the same time provide means to represent precise and potentially complex linguistic information. The general principles informing the design of LAF have been previously reported (Ide and Romary, 2003; Ide and Romary, 2004a). This paper describes some of the more technical aspects of the LAF design that have been addressed in the process of finalizing the specifications for the standard.