XArch: archiving scientific and reference data

XArch: archiving scientific and reference data
复制标题

XArch:归档科学和参考数据

DOI:
--
复制
发表时间:
2008
期刊:
SIGMOD Conference
影响因子:
--
通讯作者:
Ioannis Koltsidas
Ioannis Koltsidas
中科院分区:
--
文献类型:
--
作者:
Heiko Müller;P. Buneman;Ioannis Koltsidas

文献摘要

被引文献

相似文献

数据库归档对于旧版本数据库的检索以及数据历史记录的临时查询非常重要。我们演示了 XArch,一个用于维护、填充和查询分层数据档案的管理系统。 XArch 基于嵌套合并方法,该方法可以在紧凑的存档中有效地存储分层数据的多个版本。通过将元素合并到一个数据结构中,可以通过一次数据传递从存档中检索任何特定版本,并且可以有效跟踪对象历史记录。 XArch 实现了这种方法并以两种重要方式对其进行了扩展。首先,为了合并大型分层数据集,需要根据元素的键值对元素进行排序。我们开发了一种有效的算法来对辅助存储中的分层数据进行排序,并相应地修改了嵌套合并算法。其次,我们设计并实现了一种声明性查询语言,使人们能够查看特定版本的数据并跟踪对象的历史记录。我们使用分子生物学和人口统计参考数据作为例子来证明这一点。
Database archiving is important for the retrieval of old versions of a database and for temporal queries over the history of data. We demonstrate XArch, a management system for maintaining, populating, and querying archives of hierarchical data. XArch is based on a nested merge approach that efficiently stores multiple versions of hierarchical data in a compact archive. By merging elements into one data structure, any specific version is retrievable from the archive in a single pass over the data and efficient tracking of object history is possible. XArch implements this approach and extends it in two important ways. First, in order to merge large hierarchical data sets, elements need to be sorted according to their key values. We developed an efficient algorithm for sorting hierarchical data in secondary storage and modified the nested merge algorithm accordingly. Second, we designed and implemented a declarative query language that enables one both to view data from particular versions and to track the history of objects. We demonstrate this using both molecular biology and demographic reference data as examples.