Prediction of protein (domain) structural classes based on amino-acid index

被引:45
作者
Bu, WS
Feng, ZP
Zhang, ZD
Zhang, CT [1 ]
机构
[1] Tianjin Univ, Dept Phys, Tianjin 300072, Peoples R China
[2] Tianjin Univ, Chem Engn Res Ctr, Tianjin, Peoples R China
来源
EUROPEAN JOURNAL OF BIOCHEMISTRY | 1999年 / 266卷 / 03期
关键词
structural classes; prediction of structural class; component-coupled algorithm; amino-acid composition; amino-acid index;
D O I
10.1046/j.1432-1327.1999.00947.x
中图分类号
Q5 [生物化学]; Q7 [分子生物学];
学科分类号
071010 ; 081704 ;
摘要
A protein (domain) is usually classified into one of the following four structural classes: all-alpha, all-beta, alpha/beta and alpha + beta. In this paper, a new formulation is proposed to predict the structural class of a protein (domain) from its primary sequence. Instead of the amino-acid composition used widely in the previous structural class prediction work, the auto-correlation functions based on the profile of amino-acid index along the primary sequence of the query protein (domain) are used for the structural class prediction. Consequently, the overall predictive accuracy is remarkably improved. For the same training database consisting of 359 proteins (domains) and the same component-coupled algorithm [Chou, K.C. & Maggiora, G.M. (1998) Protein Eng. 11, 523-538], the overall predictive accuracy of the new method for the jackknife test is 5-7% higher than the accuracy based only on the amino-acid composition. The overall predictive accuracy finally obtained for the jackknife test is as high as 90.5%, implying that a significant improvement has been achieved by making full use of the information contained in the primary sequence for the class prediction. This improvement depends on the size of the training database, the auto-correlation functions selected and the amino-acid index used. We have found that the aminoacid index proposed by Oobatake and Ooi, i.e. the average nonbonded energy per residue, leads to the optimal predictive result in the case for the database sets studied in this paper. This study may be considered as an alternative step towards making the structural class prediction more practical.
引用
收藏
页码:1043 / 1049
页数:7
相关论文
共 45 条
[41]  
ZHANG CT, 1992, PROTEIN SCI, V1, P401
[42]   Prediction of the secondary structure contents of globular proteins based on three structural classes [J].
Zhang, CT ;
Zhang, ZD ;
He, ZM .
JOURNAL OF PROTEIN CHEMISTRY, 1998, 17 (03) :261-272
[43]   PREDICTING PROTEIN STRUCTURAL CLASSES FROM AMINO-ACID-COMPOSITION - APPLICATION OF FUZZY CLUSTERING [J].
ZHANG, CT ;
CHOU, KC ;
MAGGIORA, GM .
PROTEIN ENGINEERING, 1995, 8 (05) :425-435
[44]   Prediction of the helix/strand content of globular proteins based on their primary sequences [J].
Zhang, CT ;
Lin, ZS ;
Zhang, ZD ;
Yan, M .
PROTEIN ENGINEERING, 1998, 11 (11) :971-979
[45]  
ZHOU GF, 1992, EUR J BIOCHEM, V210, P747