The National Center for Biotechnology Information's Protein Clusters Database.

The National Center for Biotechnology Information's Protein Clusters Database.
复制标题

DOI:
10.1093/nar/gkn734
复制
发表时间:
2009-01
影响因子:
14.9
通讯作者:
Tatusova T
Tatusova T
中科院分区:
生物学2区
文献类型:
--
作者:
Klimke W;Agarwala R;Badretdin A;Chetvernin S;Ciufo S;Fedorov B;Kiryutin B;O'Neill K;Resch W;Resenchuk S;Schafer S;Tolstoy I;Tatusova T

文献摘要

参考文献

被引文献

相似文献

DNA测序能力的迅速提高导致原核基因组研究产生的数据大量增加,这对研究微生物进化的科学家和那些希望了解微生物系统生物基础的人来说是一个布恩。NCBI蛋白质簇数据库(ProtClustDB)的创建是为了有效地维护和保持大量数据的最新。ProtClustDB包含按序列相似性分组的有策展和无策展的蛋白质簇。2008年5月发布的版本包含来自RefSeq收集的完整染色体和质粒的3806 nt序列编码的170多万个蛋白质的285386个簇,这些蛋白质来自四个主要群体:原核生物,噬菌体以及线粒体和叶绿体细胞器。共有7180个聚类,包含376513个蛋白质,并进行了基因和蛋白质功能注释。为所有集群收集PubMed标识符和外部交叉引用,并提供额外的信息资源。一套网络工具可用于探索更详细的信息,如多重比对,系统发育树和基因组邻域。ProtClustDB为研究人员提供了一种聚合基因和蛋白质注释的有效方法,可在http://www.ncbi.nlm.nih.gov/sites/entrez?上获得db=蛋白质簇。
Rapid increases in DNA sequencing capabilities have led to a vast increase in the data generated from prokaryotic genomic studies, which has been a boon to scientists studying micro-organism evolution and to those who wish to understand the biological underpinnings of microbial systems. The NCBI Protein Clusters Database (ProtClustDB) has been created to efficiently maintain and keep the deluge of data up to date. ProtClustDB contains both curated and uncurated clusters of proteins grouped by sequence similarity. The May 2008 release contains a total of 285 386 clusters derived from over 1.7 million proteins encoded by 3806 nt sequences from the RefSeq collection of complete chromosomes and plasmids from four major groups: prokaryotes, bacteriophages and the mitochondrial and chloroplast organelles. There are 7180 clusters containing 376 513 proteins with curated gene and protein functional annotation. PubMed identifiers and external cross references are collected for all clusters and provide additional information resources. A suite of web tools is available to explore more detailed information, such as multiple alignments, phylogenetic trees and genomic neighborhoods. ProtClustDB provides an efficient method to aggregate gene and protein annotation for researchers and is available at http://www.ncbi.nlm.nih.gov/sites/entrez?db=proteinclusters.
DOI: 10.1093/nar/gkm882
发表时间: 2008-01
影响因子: 14.9
作者:
Kanehisa M;Araki M;Goto S;Hattori M;Hirakawa M;Itoh M;Katayama T;Kawashima S;Okuda S;Tokimatsu T;Yamanishi Y
通讯作者: Yamanishi Y
DOI: 10.1093/nar/gkl841
发表时间: 2007-01
影响因子: 14.9
作者:
Mulder NJ;Apweiler R;Attwood TK;Bairoch A;Bateman A;Binns D;Bork P;Buillard V;Cerutti L;Copley R;Courcelle E;Das U;Daugherty L;Dibley M;Finn R;Fleischmann W;Gough J;Haft D;Hulo N;Hunter S;Kahn D;Kanapin A;Kejariwal A;Labarga A;Langendijk-Genevaux PS;Lonsdale D;Lopez R;Letunic I;Madera M;Maslen J;McAnulla C;McDowall J;Mistry J;Mitchell A;Nikolskaya AN;Orchard S;Orengo C;Petryszak R;Selengut JD;Sigrist CJ;Thomas PD;Valentin F;Wilson D;Wu CH;Yeats C
通讯作者: Yeats C
DOI: 10.1126/science.7542800
发表时间: 1995-07-28
期刊: SCIENCE
影响因子: 56.9
作者:
FLEISCHMANN, RD;ADAMS, MD;VENTER, JC
通讯作者: VENTER, JC
DOI: 10.1093/nar/gkl1031
发表时间: 2007-01-01
影响因子: 14.9
作者:
Wheeler, David L.;Barrett, Tanya;Yaschenko, Eugene
通讯作者: Yaschenko, Eugene
DOI: 10.1128/jb.186.22.7754-7762.2004
发表时间: 2004-11-01
影响因子: 3.2
作者:
Ettema, TJG;Makarova, KS;van der Oost, J
通讯作者: van der Oost, J