ANALYSIS OF ACCURACY AND IMPLICATIONS OF SIMPLE METHODS FOR PREDICTING SECONDARY STRUCTURE OF GLOBULAR PROTEINS
ANALYSIS OF ACCURACY AND IMPLICATIONS OF SIMPLE METHODS FOR PREDICTING SECONDARY STRUCTURE OF GLOBULAR PROTEINS
复制标题
DOI:
10.1016/0022-2836(78)90297-8
复制
发表时间:
1978-01-01
影响因子:
5.6
通讯作者:
ROBSON, B
中科院分区:
文献类型:
--
作者:
GARNIER, J;OSGUTHORPE, DJ;ROBSON, B
Cooperation between a laboratory interested in developing the theory for protein secondary structure prediction methods and a laboratory interested in applying and comparing such methods has led to the development of a simple predictive algorithm. Four-state predictions, in which each residue is unambiguously assigned 1 conformational state of .alpha.-helix, extended chain, reverse turn or coil, predict 49% of residue states correctly (in a sample of 26 proteins) when the overall helix and extended-chain content is not taken into account. When the relative abundances of helix, extended chain, reverse turn and coil observed by X-ray crystallography are taken into account, a single constant for each protein and type of conformation can be used to bias the prediction. When predictions are optimized in this way, 63% of all residue states are unambiguously and correctly assigned. By analyzing the nature of the bias required, proteins can be classified into helix-rich types, pleated-sheet-rich types and so on. If the type of protein can be determined even approximately by circular dichroism, 57% of residue states can be correctly predicted without taking into account the X-ray structure. Comparable predictions can be obtained if, instead of circular dichroism, preliminary predictions are made to assess the protein type. The numbers quoted here depend on the method used to assess accuracy, and the algorithm is at least as good as, and usually superior to, the reported prediction methods assessed in the same way. Ways of further enhancing predictions by the use of additional information from hydrophobic triplets and homologous sequences are also explored. Hydrophobic triplet information does not significantly improve predictive power and this information is probably used by proteins in the next stage of folding. The use of homologous sequences appears to be very promising. The implication of these results in protein folding is discussed.