DiscoveryLink: A system for integrated access to life sciences data sources

DiscoveryLink: A system for integrated access to life sciences data sources
复制标题

DOI:
10.1147/sj.402.0489
复制
发表时间:
2001-01-01
影响因子:
--
通讯作者:
Swope, WC
Swope, WC
中科院分区:
其他
文献类型:
--
作者:
Haas, LM;Schwarz, PM;Swope, WC

文献摘要

被引文献

相似文献

目前,大量生命科学数据存在于专门的数据源中,具有专门的查询处理功能。来自一个来源的数据通常必须与来自其他来源的数据相结合,以向用户提供他们想要的信息。存在响应于单个查询从多个源提取数据的数据库中间件系统。IBM的DiscoveryLink就是这样一个系统,目标是生命科学行业的应用程序。DiscoveryLink为用户提供了一个虚拟数据库,他们可以提出任意复杂的查询,即使回答查询所需的实际数据可能来自几个不同的来源,而这些来源本身都不能回答查询。我们描述了DiscoveryLink产品,侧重于两个关键要素,包装器架构和查询优化器,并说明如何可以用来集成访问生命科学数据从异构数据源。
Vast amounts of life sciences data reside today in specialized data sources, with specialized query processing capabilities. Data from one source often must be combined with data from other sources to give users the information they desire. There are database middleware systems that extract data from multiple sources in response to a single query. IBM's DiscoveryLink is one such system, targeted to applications from the life sciences industry. DiscoveryLink provides users with a virtual database to which they can pose arbitrarily complex queries, even though the actual data needed to answer the query may originate from several different sources, and none of those sources, by itself is capable of answering the query. We describe the DiscoveryLink offering, focusing on two key elements, the wrapper architecture and the query optimizer and illustrate how if can be used to integrate the access to life sciences data from heterogeneous data sources.