On the accuracy of homology modeling and sequence alignment methods applied to membrane proteins

被引:184
作者
Forrest, Lucy R. [1 ]
Tang, Christopher L. [1 ]
Honig, Barry [1 ]
机构
[1] Columbia Univ, Howard Hughes Med Inst, Dept Biochem & Mol Biophys, Ctr Computat Biol & Bioinformat, New York, NY 10032 USA
基金
美国国家科学基金会;
关键词
D O I
10.1529/biophysj.106.082313
中图分类号
Q6 [生物物理学];
学科分类号
071011 ;
摘要
In this study, we investigate the extent to which techniques for homology modeling that were developed for water-soluble proteins are appropriate for membrane proteins as well. To this end we present an assessment of current strategies for homology modeling of membrane proteins and introduce a benchmark data set of homologous membrane protein structures, called HOMEP. First, we use HOMEP to reveal the relationship between sequence identity and structural similarity in membrane proteins. This analysis indicates that homology modeling is at least as applicable to membrane proteins as it is to water-soluble proteins and that acceptable models ( with C alpha-RMSD values to the native of 2 A or less in the transmembrane regions) may be obtained for template sequence identities of 30% or higher if an accurate alignment of the sequences is used. Second, we show that secondary-structure prediction algorithms that were developed for water-soluble proteins perform approximately as well for membrane proteins. Third, we provide a comparison of a set of commonly used sequence alignment algorithms as applied to membrane proteins. We find that high-accuracy alignments of membrane protein sequences can be obtained using state-of-the-art profile-to-profile methods that were developed for water-soluble proteins. Improvements are observed when weights derived from the secondary structure of the query and the template are used in the scoring of the alignment, a result which relies on the accuracy of the secondary-structure prediction of the query sequence. The most accurate alignments were obtained using template profiles constructed with the aid of structural alignments. In contrast, a simple sequence-to-sequence alignment algorithm, using a membrane protein-specific substitution matrix, shows no improvement in alignment accuracy. We suggest that profile-to-profile alignment methods should be adopted to maximize the accuracy of homology models of membrane proteins.
引用
收藏
页码:508 / 517
页数:10
相关论文
共 71 条
[51]   EVA: Large-scale analysis of secondary structure prediction [J].
Rost, B ;
Eyrich, VA .
PROTEINS-STRUCTURE FUNCTION AND BIOINFORMATICS, 2001, :192-199
[52]  
Rost B, 1996, METHOD ENZYMOL, V266, P525
[53]   Review: Protein secondary structure prediction continues to rise [J].
Rost, B .
JOURNAL OF STRUCTURAL BIOLOGY, 2001, 134 (2-3) :204-218
[54]   COMPARATIVE PROTEIN MODELING BY SATISFACTION OF SPATIAL RESTRAINTS [J].
SALI, A ;
BLUNDELL, TL .
JOURNAL OF MOLECULAR BIOLOGY, 1993, 234 (03) :779-815
[55]   STAM: simple Transmembrane Alignment Method [J].
Shafrir, Y ;
Guy, HR .
BIOINFORMATICS, 2004, 20 (05) :758-U644
[56]   On the role of structural information in remote homology detection and sequence alignment: New methods using hybrid sequence profiles [J].
Tang, CL ;
Xie, L ;
Koh, IYY ;
Posy, S ;
Alexov, E ;
Honig, B .
JOURNAL OF MOLECULAR BIOLOGY, 2003, 334 (05) :1043-1062
[57]   BAliBASE 3.0: Latest developments of the multiple sequence alignment benchmark [J].
Thompson, JD ;
Koehl, P ;
Ripp, R ;
Poch, O .
PROTEINS-STRUCTURE FUNCTION AND BIOINFORMATICS, 2005, 61 (01) :127-136
[58]   CLUSTAL-W - IMPROVING THE SENSITIVITY OF PROGRESSIVE MULTIPLE SEQUENCE ALIGNMENT THROUGH SEQUENCE WEIGHTING, POSITION-SPECIFIC GAP PENALTIES AND WEIGHT MATRIX CHOICE [J].
THOMPSON, JD ;
HIGGINS, DG ;
GIBSON, TJ .
NUCLEIC ACIDS RESEARCH, 1994, 22 (22) :4673-4680
[59]   A comprehensive comparison of multiple sequence alignment programs [J].
Thompson, JD ;
Plewniak, F ;
Poch, O .
NUCLEIC ACIDS RESEARCH, 1999, 27 (13) :2682-2690
[60]   Assessment of predictions submitted for the CASP6 comparative modeling category [J].
Tress, M ;
Ezkurdia, L ;
Graña, O ;
López, G ;
Valencia, A .
PROTEINS-STRUCTURE FUNCTION AND BIOINFORMATICS, 2005, 61 :27-45