GeneCards: a novel functional genomics compendium with automated data mining and query reformulation support

GeneCards: a novel functional genomics compendium with automated data mining and query reformulation support
复制标题

DOI:
10.1093/bioinformatics/14.8.656
复制
发表时间:
1998-01-01
期刊:
影响因子:
5.8
通讯作者:
Lancet, D
Lancet, D
中科院分区:
生物学3区
文献类型:
--
作者:
Rebhan, M;Chalifa-Caspi, V;Lancet, D

文献摘要

被引文献

相似文献

动机:现代生物学正在从“一个基因一个博士后”的方法转向包括同时监测数千个基因的基因组分析。因此,在学术和工业研究中,有效获得简明和综合的生物医学信息以支持数据分析和决策的重要性正在迅速增加。然而,在广泛分散的资源相关的生物医学研究的知识发现往往是一个繁琐和不平凡的木桶。一个需要大量的培训和努力。结果:为了开发一种新型的主题特定的概述资源,提供有效的访问分布式信息,我们设计了一个数据库cc-riled 'GeneCards'的模型。这是一个免费访问的网络资源,为目前由HUGO/GDB命名委员会出版的7000多个人类基因中的每一个提供了一个超文本“卡片”。所提供的信息旨在立即深入了解有关相应基因的现有知识,包括关注其在健康和疾病中的功能。它是由Perl脚本编译,自动提取相关信息,从几个数据库,包括SWISS-PROT OMIM,Genatlas和GDB。分析用户的互动与网络界面的基因卡触发开发易于扫描显示优化人类浏览。此外,我们开发的算法,提供“后点击”查询重构支持,以促进信息检索和探索。许多长期用户转向GeneCnrds,以快速获取有关大量基因功能的信息,例如在使用“DNA芯片”技术或二维蛋白质电泳的大规模表达研究领域。
Motivation: Modem biology is shifting from the 'one gene one postdoc' approach to genomic analyses that include the simultaneous monitoring of thousands of genes. The importance of efficient access to concise and integrated biomedical information to support data analysis and decision making is therefore increasing rapidly, in both academic and industrial research. However, knowledge discovery in the widely scattered resources relevant for biomedical research is often a cumbersome and non-trivial cask. one that requires a significant amount of training and effort.Results: To develop a model for a new type of topic-specific overview resource that provides efficient access to distributed information we designed a database cc-riled 'GeneCards'. It is a freely accessible Web resource that offers one hypertext 'card' for each of the more than 7000 human genes that currently have an approved gene symbol published by the HUGO/GDB nomenclature committee. The presented information aims at giving immediate insight into current knowledge about the respective gene, including a focus on its functions in health and disease. It is compiled by Perl scripts that automatically extract relevant information from several databases including SWISS-PROT OMIM, Genatlas and GDB.Analyses of the interactions of users with the Web interface of GeneCards triggered development of easy-to-scan displays optimized for human browsing. Also, we developed algorithms that offer 'rearly-to-click' query reformulation support to facilitate information retrieval and exploration. Many of the long-term users turn To GeneCnrds to quickly access information about the function of very large sets of genes, for example in the realm of large-scale expression studies using 'DNA chip' technology or two-dimensional protein electrophoresis.