Improved coarse-graining of Markov state models via explicit consideration of statistical uncertainty

被引:68
作者
Bowman, Gregory R. [1 ,2 ]
机构
[1] Univ Calif Berkeley, Dept Chem, Berkeley, CA 94720 USA
[2] Univ Calif Berkeley, Dept Mol & Cell Biol, Berkeley, CA 94720 USA
基金
美国国家卫生研究院;
关键词
MOLECULAR-DYNAMICS; SIMULATIONS; EFFICIENT;
D O I
10.1063/1.4755751
中图分类号
O64 [物理化学(理论化学)、化学物理学];
学科分类号
070304 ; 081704 ;
摘要
Markov state models (MSMs)-or discrete-time master equation models-are a powerful way of modeling the structure and function of molecular systems like proteins. Unfortunately, MSMs with sufficiently many states to make a quantitative connection with experiments (often tens of thousands of states even for small systems) are generally too complicated to understand. Here, I present a Bayesian agglomerative clustering engine (BACE) for coarse-graining such Markov models, thereby reducing their complexity and making them more comprehensible. An important feature of this algorithm is its ability to explicitly account for statistical uncertainty in model parameters that arises from finite sampling. This advance builds on a number of recent works highlighting the importance of accounting for uncertainty in the analysis of MSMs and provides significant advantages over existing methods for coarse-graining Markov state models. The closed-form expression I derive here for determining which states to merge is equivalent to the generalized Jensen-Shannon divergence, an important measure from information theory that is related to the relative entropy. Therefore, the method has an appealing information theoretic interpretation in terms of minimizing information loss. The bottom-up nature of the algorithm likely makes it particularly well suited for constructing mesoscale models. I also present an extremely efficient expression for Bayesian model comparison that can be used to identify the most meaningful levels of the hierarchy of models from BACE. (C) 2012 American Institute of Physics. [http://dx.doi.org/10.1063/1.4755751]
引用
收藏
页数:7
相关论文
共 36 条
[1]  
[Anonymous], 1976, Denumerable Markov Chains
[2]   Bayesian comparison of Markov models of molecular dynamics with detailed balance constraint [J].
Bacallado, Sergio ;
Chodera, John D. ;
Pande, Vijay .
JOURNAL OF CHEMICAL PHYSICS, 2009, 131 (04)
[3]   MSMBuilder2: Modeling Conformational Dynamics on the Picosecond to Millisecond Scale [J].
Beauchamp, Kyle A. ;
Bowman, Gregory R. ;
Lane, Thomas J. ;
Maibaum, Lutz ;
Haque, Imran S. ;
Pande, Vijay S. .
JOURNAL OF CHEMICAL THEORY AND COMPUTATION, 2011, 7 (10) :3412-3419
[4]   Mapping the dynamics of multi-dimensional systems onto a nearest-neighbor coupled discrete set of states conserving the mean first-passage times: a projective dynamics approach [J].
Biswas, Katja ;
Novotny, M. A. .
JOURNAL OF PHYSICS A-MATHEMATICAL AND THEORETICAL, 2011, 44 (34)
[5]   Protein folded states are kinetic hubs [J].
Bowman, Gregory R. ;
Pande, Vijay S. .
PROCEEDINGS OF THE NATIONAL ACADEMY OF SCIENCES OF THE UNITED STATES OF AMERICA, 2010, 107 (24) :10890-10895
[6]   Network models for molecular kinetics and their initial applications to human health [J].
Bowman, Gregory R. ;
Huang, Xuhui ;
Pande, Vijay S. .
CELL RESEARCH, 2010, 20 (06) :622-630
[7]   Enhanced Modeling via Network Theory: Adaptive Sampling of Markov State Models [J].
Bowman, Gregory R. ;
Ensign, Daniel L. ;
Pande, Vijay S. .
JOURNAL OF CHEMICAL THEORY AND COMPUTATION, 2010, 6 (03) :787-794
[8]   Using generalized ensemble simulations and Markov state models to identify conformational states [J].
Bowman, Gregory R. ;
Huang, Xuhui ;
Pande, Vijay S. .
METHODS, 2009, 49 (02) :197-201
[9]   Coarse master equations for peptide folding dynamics [J].
Buchete, Nicolae-Viorel ;
Hummer, Gerhard .
JOURNAL OF PHYSICAL CHEMISTRY B, 2008, 112 (19) :6057-6069
[10]   Automatic discovery of metastable states for the construction of Markov models of macromolecular conformational dynamics [J].
Chodera, John D. ;
Singhal, Nina ;
Pande, Vijay S. ;
Dill, Ken A. ;
Swope, William C. .
JOURNAL OF CHEMICAL PHYSICS, 2007, 126 (15)