A genomic perspective on protein families

A genomic perspective on protein families
复制标题

DOI:
10.1126/science.278.5338.631
复制
发表时间:
1997-10-24
期刊:
影响因子:
56.9
通讯作者:
Lipman, DJ
Lipman, DJ
中科院分区:
综合性期刊1区
文献类型:
--
作者:
Tatusov, RL;Koonin, EV;Lipman, DJ

文献摘要

被引文献

相似文献

为了从快速积累的基因组序列中提取最大量的信息,需要根据它们的同源关系对所有保守基因进行分类。比较编码的蛋白质在7个完整的基因组从5个主要的系统发育谱系和阐明一致的模式的序列相似性允许划定720簇的orthopathic组(COG)。每个COG由来自至少三个谱系的单独的正向同源蛋白或旁系同源物的正向同源组组成。直系同源物通常具有相同的功能,允许将功能信息从一个成员转移到整个COG。这种关系自动产生了一些功能的预测不佳的特征基因组。COG构成了功能和进化基因组分析的框架。
In order to extract the maximum amount of information from the rapidly accumulating genome sequences, all conserved genes need to be classified according to their homologous relationships. Comparison of proteins encoded in seven complete genomes from five major phylogenetic lineages and elucidation of consistent patterns of sequence similarities allowed the delineation of 720 clusters of orthologous groups (COGs). Each COG consists of individual orthologous proteins or orthologous sets of paralogs from at least three lineages. Orthologs typically have the same function, allowing transfer of functional information from one member to an entire COG. This relation automatically yields a number of functional predictions for poorly characterized genomes. The COGs comprise a framework for functional and evolutionary genome analysis.