XArch: archiving scientific and reference data
XArch: archiving scientific and reference data
复制标题
XArch:归档科学和参考数据
DOI:
--
复制
发表时间:
2008
期刊:
影响因子:
--
通讯作者:
Ioannis Koltsidas
中科院分区:
文献类型:
--
作者:
Heiko Müller;P. Buneman;Ioannis Koltsidas
Database archiving is important for the retrieval of old versions of a database and for temporal queries over the history of data. We demonstrate XArch, a management system for maintaining, populating, and querying archives of hierarchical data. XArch is based on a nested merge approach that efficiently stores multiple versions of hierarchical data in a compact archive. By merging elements into one data structure, any specific version is retrievable from the archive in a single pass over the data and efficient tracking of object history is possible. XArch implements this approach and extends it in two important ways. First, in order to merge large hierarchical data sets, elements need to be sorted according to their key values. We developed an efficient algorithm for sorting hierarchical data in secondary storage and modified the nested merge algorithm accordingly. Second, we designed and implemented a declarative query language that enables one both to view data from particular versions and to track the history of objects. We demonstrate this using both molecular biology and demographic reference data as examples.