Fast pseudolikelihood maximization for direct-coupling analysis of protein structure from many homologous amino-acid sequences
Fast pseudolikelihood maximization for direct-coupling analysis of protein structure from many homologous amino-acid sequences
复制标题
DOI:
10.1016/j.jcp.2014.07.024
复制
发表时间:
2014-11-01
影响因子:
4.1
通讯作者:
Aurell, Erik
中科院分区:
文献类型:
--
作者:
Ekeberg, Magnus;Hartonen, Tuomo;Aurell, Erik
Direct-coupling analysis is a group of methods to harvest information about coevolving residues in a protein family by learning a generative model in an exponential family from data. In protein families of realistic size, this learning can only be done approximately, and there is a trade-off between inference precision and computational speed. We here show that an earlier introduced l(2)-regularized pseudolikelihood maximization method called plmDCA can be modified as to be easily parallelizable, as well as inherently faster on a single processor, at negligible difference in accuracy. We test the new incarnation of the method on 143 protein family/structure-pairs from the Protein Families database (PFAM), one of the larger tests of this class of algorithms to date. (C) 2014 Elsevier Inc. Allrightsreserved.