The TIGR Gene Indices: analysis of gene transcript sequences in highly sampled eukaryotic species

The TIGR Gene Indices: analysis of gene transcript sequences in highly sampled eukaryotic species
复制标题

DOI:
10.1093/nar/29.1.159
复制
发表时间:
2001-01-01
影响因子:
14.9
通讯作者:
White, J
White, J
中科院分区:
生物学2区
文献类型:
--
作者:
Quackenbush, J;Cho, J;White, J

文献摘要

被引文献

相似文献

虽然基因组测序计划进展迅速,但EST测序和分析仍然是鉴定和分类各种物种基因序列的主要研究工具,也是基因组序列注释的重要资源。TIGR基因索引(http://www.tigr.org/tdb/tgi.shtml)是物种特定数据库的集合,这些数据库使用高度精细的协议来分析EST序列,试图识别该数据代表的基因并提供有关这些基因的更多信息。基因索引是通过首先聚类,然后组装EST和来自GenBank的目标物种的注释基因序列来构建的。这个过程产生了一组独特的、高保真的虚拟转录本,或暂定共识(TC)序列。TC序列可用于为推测的基因提供功能注释,将转录本与作图和基因组序列数据联系起来,提供同源和类同源基因之间的联系,并作为比较序列分析的资源。
While genome sequencing projects are advancing rapidly, EST sequencing and analysis remains a primary research tool for the identification and categorization of gene sequences in a wide variety of species and an important resource for annotation of genomic sequence. The TIGR Gene Indices (http:// www.tigr.org/tdb/tgi.shtml) are a collection of species-specific databases that use a highly refined protocol to analyze EST sequences in an attempt to identify the genes represented by that data and to provide additional information regarding those genes. Gene Indices are constructed by first clustering, then assembling EST and annotated gene sequences from GenBank for the targeted species. This process produces a set of unique, high-fidelity virtual transcripts, or Tentative Consensus (TC) sequences. The TC sequences can be used to provide putative genes with functional annotation, to link the transcripts to mapping and genomic sequence data, to provide links between orthologous and paralogous genes and as a resource for comparative sequence analysis.