The COG database: new developments in phylogenetic classification of proteins from complete genomes

The COG database: new developments in phylogenetic classification of proteins from complete genomes
复制标题

DOI:
10.1093/nar/29.1.22
复制
发表时间:
2001-01-01
影响因子:
14.9
通讯作者:
Koonin, EV
Koonin, EV
中科院分区:
生物学2区
文献类型:
--
作者:
Tatusov, RL;Natale, DA;Koonin, EV

文献摘要

被引文献

相似文献

直系同源蛋白(COG)的簇数据库,该数据库代表了对完全基因组编码的蛋白质的系统发育分类的尝试,目前由2791个COG组成,包括45 350个蛋白质,包括来自30个细菌,古细菌和酵母菌糖含量塞雷氏菌的30个基因组的45 350个蛋白质(http://www.ncbi.nlm.nih,gov/cog)。此外,还提供了对齿轮的补充,其中在两个多细胞真核生物的基因组中编码的蛋白质,线虫秀丽隐杆线虫和果蝇果蝇甲状腺癌,并与细菌和/或古细菌共享。添加到COG数据库中的新功能包括有关每个COG和文献参考的结构和功能详细信息的信息页面,用于将新蛋白拟合到COGS中的Cognitor程序的改进以及通过使用主要成分构建的基因组和COG的分类分析。
The database of Clusters of Orthologous Groups of proteins (COGs), which represents an attempt on a phylogenetic classification of the proteins encoded in complete genomes, currently consists of 2791 COGs including 45 350 proteins from 30 genomes of bacteria, archaea and the yeast Saccharomyces cerevisiae (http://www.ncbi.nlm.nih,gov/COG). In addition, a supplement to the COGs is available, in which proteins encoded in the genomes of two multicellular eukaryotes, the nematode Caenorhabditis elegans and the fruit fly Drosophila melanogaster, and shared with bacteria and/or archaea were included. The new features added to the COG database include information pages with structural and functional details on each COG and literature references, improvements of the COGNITOR program that is used to fit new proteins into the COGs, and classification of genomes and COGs constructed by using principal component analysis.