PairsDB atlas of protein sequence space.

PairsDB atlas of protein sequence space.
复制标题

DOI:
10.1093/nar/gkm879
复制
发表时间:
2008-01
影响因子:
14.9
通讯作者:
Holm L
Holm L
中科院分区:
生物学2区
文献类型:
--
作者:
Heger A;Korpelainen E;Hupponen T;Mattila K;Ollikainen V;Holm L

文献摘要

参考文献

被引文献

相似文献

序列相似性/数据库搜索是分子生物学的基石。PairsDB是一个数据库,旨在快速轻松地探索蛋白质序列及其相似性关系。PairsDB背后是蛋白质序列的全面集合,以及它们之间的BLAST和PSI-BLAST比对。不是对每个请求单独运行BLAST或PSI-BLAST,而是从预计算对齐的数据库中即时检索结果。过滤选项使您可以找到一组满足一组标准的序列-例如,所有具有已解决结构且没有跨膜片段的人类蛋白质。PairsDB不断更新,涵盖UniProt中的所有序列。数据存储在MySQL关系数据库中。数据文件将在ftp://nic.funet.fi/pub/sci/molbio.上提供下载还可以通过http://pairsdb.csc.fi.以交互方式访问PairsDBPairsDB数据是构建各种下游自动化分析管道的有价值的平台。例如,All-Against-All相似性关系图是聚类蛋白质家族、描绘结构域、通过一致性测量提高比对准确性以及定义同源基因的起点。此外,查询锚定的堆叠序列比对、轮廓和共有序列在序列保守模式的研究中是有用的,以寻找关于可能的功能位点的线索。
Sequence similarity/database searching is a cornerstone of molecular biology. PairsDB is a database intended to make exploring protein sequences and their similarity relationships quick and easy. Behind PairsDB is a comprehensive collection of protein sequences and BLAST and PSI-BLAST alignments between them. Instead of running BLAST or PSI-BLAST individually on each request, results are retrieved instantaneously from a database of pre-computed alignments. Filtering options allow you to find a set of sequences satisfying a set of criteria—for example, all human proteins with solved structure and without transmembrane segments. PairsDB is continually updated and covers all sequences in Uniprot. The data is stored in a MySQL relational database. Data files will be made available for download at ftp://nic.funet.fi/pub/sci/molbio. PairsDB can also be accessed interactively at http://pairsdb.csc.fi. PairsDB data is a valuable platform to build various downstream automated analysis pipelines. For example, the graph of all-against-all similarity relationships is the starting point for clustering protein families, delineating domains, improving alignment accuracy by consistency measures, and defining orthologous genes. Moreover, query-anchored stacked sequence alignments, profiles and consensus sequences are useful in studies of sequence conservation patterns for clues about possible functional sites.
Pfam:氏族、网络工具和服务。
DOI: 10.1093/nar/gkj149
发表时间: 2006-01-01
影响因子: 14.9
作者:
Finn, Robert D.;Mistry, Jaina;Schuster-Bockler, Benjamin;Griffiths-Jones, Sam;Hollich, Volker;Lassmann, Timo;Moxon, Simon;Marshall, Mhairi;Khanna, Ajay;Durbin, Richard;Eddy, Sean R.;Sonnhammer, Erik L. L.;Bateman, Alex
通讯作者: Bateman, Alex
DOI: 10.1093/nar/gkl841
发表时间: 2007-01
影响因子: 14.9
作者:
Mulder NJ;Apweiler R;Attwood TK;Bairoch A;Bateman A;Binns D;Bork P;Buillard V;Cerutti L;Copley R;Courcelle E;Das U;Daugherty L;Dibley M;Finn R;Fleischmann W;Gough J;Haft D;Hulo N;Hunter S;Kahn D;Kanapin A;Kejariwal A;Labarga A;Langendijk-Genevaux PS;Lonsdale D;Lopez R;Letunic I;Madera M;Maslen J;McAnulla C;McDowall J;Mistry J;Mitchell A;Nikolskaya AN;Orchard S;Orengo C;Petryszak R;Selengut JD;Sigrist CJ;Thomas PD;Valentin F;Wilson D;Wu CH;Yeats C
通讯作者: Yeats C
DOI: 10.1093/bioinformatics/17.3.282
发表时间: 2001-03-01
期刊: BIOINFORMATICS
影响因子: 5.8
作者:
Li, WZ;Jaroszewski, L;Godzik, A
通讯作者: Godzik, A
DOI: 10.1093/bioinformatics/btg213
发表时间: 2003-09-01
期刊: BIOINFORMATICS
影响因子: 5.8
作者:
Wall, DP;Fraser, HB;Hirsh, AE
通讯作者: Hirsh, AE
DOI: 10.1093/nar/gkh039
发表时间: 2004-01-01
影响因子: 14.9
作者:
Andreeva, A;Howorth, D;Murzin, AG
通讯作者: Murzin, AG