The subsystems approach to genome annotation and its use in the project to annotate 1000 genomes.

The subsystems approach to genome annotation and its use in the project to annotate 1000 genomes.
复制标题

DOI:
10.1093/nar/gki866
复制
发表时间:
2005
影响因子:
14.9
通讯作者:
Vonstein V
Vonstein V
中科院分区:
生物学2区
文献类型:
--
作者:
Overbeek R;Begley T;Butler RM;Choudhuri JV;Chuang HY;Cohoon M;de Crécy-Lagard V;Diaz N;Disz T;Edwards R;Fonstein M;Frank ED;Gerdes S;Glass EM;Goesmann A;Hanson A;Iwata-Reuyl D;Jensen R;Jamshidi N;Krause L;Kubal M;Larsen N;Linke B;McHardy AC;Meyer F;Neuweger H;Olsen G;Olson R;Osterman A;Portnoy V;Pusch GD;Rodionov DA;Rückert C;Steiner J;Stevens R;Thiele I;Vassieva O;Ye Y;Zagnitko O;Vonstein V

文献摘要

参考文献

被引文献

相似文献

第1000个完整的微生物基因组将在未来两到三年内发布。为了迎接这一里程碑,基因组解释研究会(FIG)启动了注释1000个基因组的项目。该项目是围绕这样一个原则建立的:提高高通量注释技术准确性的关键是让专家注释完整基因组集合中的单个子系统,而不是让注释专家尝试注释单个基因组中的所有基因。使用子系统方法,子系统中的专家分析实现子系统的所有基因。创建了一个注释环境,其中已填充的子系统被策展并投影到新的基因组。定义了一个可移植的填充子系统的概念,并开发了用于交换和管理这些对象的工具。还开发了工具来解决填充子系统之间的冲突。SEED是第一个支持这种注释模型的注释环境。在这里,我们描述了子系统的方法,并提供了我们不断增长的填充子系统库的第一个版本。最初发布的数据包括180 177个不同的蛋白质与2133个不同的功能作用。这些数据来自173个子系统和383种不同的生物。
The release of the 1000th complete microbial genome will occur in the next two to three years. In anticipation of this milestone, the Fellowship for Interpretation of Genomes (FIG) launched the Project to Annotate 1000 Genomes. The project is built around the principle that the key to improved accuracy in high-throughput annotation technology is to have experts annotate single subsystems over the complete collection of genomes, rather than having an annotation expert attempt to annotate all of the genes in a single genome. Using the subsystems approach, all of the genes implementing the subsystem are analyzed by an expert in that subsystem. An annotation environment was created where populated subsystems are curated and projected to new genomes. A portable notion of a populated subsystem was defined, and tools developed for exchanging and curating these objects. Tools were also developed to resolve conflicts between populated subsystems. The SEED is the first annotation environment that supports this model of annotation. Here, we describe the subsystem approach, and offer the first release of our growing library of populated subsystems. The initial release of data includes 180 177 distinct proteins with 2133 distinct functional roles. This data comes from 173 subsystems and 383 different organisms.
DOI: 10.1007/bf02138780
发表时间: 1994-01-01
影响因子: 3.6
作者:
GIBSON, KM;LEE, CF;HOFFMANN, GF
通讯作者: HOFFMANN, GF
DOI: 10.1126/science.7542800
发表时间: 1995-07-28
期刊: SCIENCE
影响因子: 56.9
作者:
FLEISCHMANN, RD;ADAMS, MD;VENTER, JC
通讯作者: VENTER, JC
DOI: 10.1159/000070268
发表时间: 2003-01-01
影响因子: 1.2
作者:
Jordan, IK;Henze, K;Galperin, MY
通讯作者: Galperin, MY
DOI: 10.1093/bioinformatics/bti1052
发表时间: 2005-06-01
期刊: BIOINFORMATICS
影响因子: 5.8
作者:
Ye, YZ;Osterman, A;Godzik, A
通讯作者: Godzik, A
DOI: 10.1074/jbc.c500044200
发表时间: 2005-05-27
影响因子: 4.8
作者:
Brand, LA;Strauss, E
通讯作者: Strauss, E