Insights from modeling protein evolution with context-dependent mutation and asymmetric amino acid selection

被引:4
作者
Saunders, Christopher T. [1 ]
Green, Phil [1 ,2 ]
机构
[1] Univ Washington, Dept Genome Sci, Seattle, WA 98195 USA
[2] Howard Hughes Med Inst, Seattle, WA USA
关键词
protein evolution; amino acid selection; context-dependent mutation; phylogenetic analysis; protein expression; protein structure;
D O I
10.1093/molbev/msm190
中图分类号
Q5 [生物化学]; Q7 [分子生物学];
学科分类号
071010 ; 081704 ;
摘要
We develop an approximate maximum likelihood method to estimate flanking nucleotide context-dependent mutation rates and amino acid exchange-dependent selection in orthologous protein-coding sequences and use it to analyze genome-wide coding sequence alignments from mammals and yeast. Allowing context-dependent mutation provides a better fit to coding sequence data than simpler (context-independent or CpG "hotspot") models and significantly affects selection parameter estimates. Allowing asymmetric (nonreciprocal) selection on amino acid exchanges gives a better fit than simple dN/dS or symmetric selection models. Relative selection strength estimates from our models show good agreement with independent estimates derived from human disease-causing and engineered mutations. Selection strengths depend on local protein structure, showing expected biophysical trends in helical versus nonhelical regions and increased asymmetry on polar-hydrophobic exchanges with increased burial. The more stringent selection that has previously been observed for highly expressed proteins is primarily concentrated in buried regions, supporting the notion that such proteins are under stronger than average selection for stability. Our analyses indicate that a highly parameterized model of mutation and selection is computationally tractable and is a useful tool for exploring a variety of biological questions concerning protein and coding sequence evolution.
引用
收藏
页码:2632 / 2647
页数:16
相关论文
共 69 条
[1]   Metabolic efficiency and amino acid composition in the proteomes of Escherichia coli and Bacillus subtilis [J].
Akashi, H ;
Gojobori, T .
PROCEEDINGS OF THE NATIONAL ACADEMY OF SCIENCES OF THE UNITED STATES OF AMERICA, 2002, 99 (06) :3695-3700
[2]   Gapped BLAST and PSI-BLAST: a new generation of protein database search programs [J].
Altschul, SF ;
Madden, TL ;
Schaffer, AA ;
Zhang, JH ;
Zhang, Z ;
Miller, W ;
Lipman, DJ .
NUCLEIC ACIDS RESEARCH, 1997, 25 (17) :3389-3402
[3]  
[Anonymous], 1978, Atlas of protein sequence and structure
[4]  
[Anonymous], 1974, SOLVING LEAST SQUARE
[5]   Identification and measurement of neighbor-dependent nucleotide substitution processes [J].
Arndt, PF ;
Hwa, T .
BIOINFORMATICS, 2005, 21 (10) :2322-2328
[6]   DNA sequence evolution with neighbor-dependent mutation [J].
Arndt, PF ;
Burge, CB ;
Hwa, T .
JOURNAL OF COMPUTATIONAL BIOLOGY, 2003, 10 (3-4) :313-322
[7]   The decline of isochores in mammals: An assessment of the GC content variation along the mammalian phylogeny [J].
Belle, EMS ;
Duret, L ;
Galtier, N ;
Eyre-Walker, A .
JOURNAL OF MOLECULAR EVOLUTION, 2004, 58 (06) :653-660
[8]   STRUCTURAL BASIS OF AMINO-ACID ALPHA-HELIX PROPENSITY [J].
BLABER, M ;
ZHANG, XJ ;
MATTHEWS, BW .
SCIENCE, 1993, 260 (5114) :1637-1640
[9]   THE INFLUENCE OF NEAREST NEIGHBORS ON THE RATE AND PATTERN OF SPONTANEOUS POINT MUTATIONS [J].
BLAKE, RD ;
HESS, ST ;
NICHOLSONTUELL, J .
JOURNAL OF MOLECULAR EVOLUTION, 1992, 34 (03) :189-200
[10]   The ASTRAL Compendium in 2004 [J].
Chandonia, JM ;
Hon, G ;
Walker, NS ;
Lo Conte, L ;
Koehl, P ;
Levitt, M ;
Brenner, SE .
NUCLEIC ACIDS RESEARCH, 2004, 32 :D189-D192