HOW TO SEARCH FOR RNA STRUCTURES THEORETICAL CONCEPTS IN EVOLUTIONARY BIOTECHNOLOGY

被引:45
作者
SCHUSTER, P
机构
[1] Institut für Molekulare Biotechnologie e.V., D-07708 Jena, Beutenbergstraße 11
关键词
APPLIED MOLECULAR EVOLUTION; DARWINS PRINCIPLE; EVOLUTIONARY BIOTECHNOLOGY; INVERSE FOLDING; RNA SECONDARY STRUCTURE; SEQUENCE SPACE; SHAPE SPACE;
D O I
10.1016/0168-1656(94)00085-Q
中图分类号
Q81 [生物工程学(生物技术)]; Q93 [微生物学];
学科分类号
071005 ; 0836 ; 090102 ; 100705 ;
摘要
The relation between RNA sequences and minimum free energy secondary structures is viewed as a mapping from sequence space into shape space. The properties of such mappings depend strongly on the ratios of the numbers of sequences and structures and, hence, substantial differences are observed between samples of structures derived from AUGC, pure AU or pure GC sequences. Statistical analysis of large samples is used to demonstrate that structures from AUGC sequences are much less sensitive to point mutations than those from sequences containing exclusively AU or GC. The frequency with which a structure is realized in sequence space is inversely proportional to some power c > 1 of the structure's frequency rank, thus following a (generalized) Zipf law. For long sequences the exponent approaches c = 1. An inverse folding algorithm is used to compute samples of sequences folding into the same secondary structure. These sequences are distributed randomly in sequence space. Common structures form extended neutral networks along which populations can migrate through the entire sequence space without changing structure. In this migration, moves of Hamming distance d = 1 and d = 2 are accepted in order to allow for base and base pair exchanges, respectively. Around any arbitrarily chosen sequence a ball that contains sequences folding into all common structures can be drawn. This ball has a diameter that is much smaller than the diameter of sequence space. Hence, only a small fraction of sequence space needs to be searched in order to find a given structure. The results derived from the mapping of sequences into structures are used to suggest a rationale for evolutionary searches on RNA structures: selection cycles with high and low mutation rates applied in alternation. Generalizations of the results to RNA 3-D structures and protein structures are discussed.
引用
收藏
页码:239 / 257
页数:19
相关论文
共 47 条
[1]   ISOLATION OF NEW RIBOZYMES FROM A LARGE POOL OF RANDOM SEQUENCES [J].
BARTEL, DP ;
SZOSTAK, JW .
SCIENCE, 1993, 261 (5127) :1411-1418
[2]  
Bauer GJ, 1989, NACHR CHEM TECH LAB, V37, P484, DOI [10.1002/nadc.19890370508, DOI 10.1002/NADC.19890]
[3]  
BAUER GJ, 1990, THESIS U BRAUNSCHWEI
[4]   DIRECTED EVOLUTION OF AN RNA ENZYME [J].
BEAUDRY, AA ;
JOYCE, GF .
SCIENCE, 1992, 257 (5070) :635-641
[5]  
BONHOEFFER S, 1993, EUR BIOPHYS J BIOPHY, V22, P13, DOI 10.1007/BF00205808
[7]   RNA AS AN ENZYME [J].
CECH, TR .
SCIENTIFIC AMERICAN, 1986, 255 (05) :64-&
[8]  
EIGEN M, 1989, ADV CHEM PHYS, V75, P149
[9]   MOLECULAR QUASI-SPECIES [J].
EIGEN, M ;
MCCASKILL, J ;
SCHUSTER, P .
JOURNAL OF PHYSICAL CHEMISTRY, 1988, 92 (24) :6881-6891
[10]   EVOLUTIONARY MOLECULAR ENGINEERING BASED ON RNA REPLICATION [J].
EIGEN, M ;
GARDINER, W .
PURE AND APPLIED CHEMISTRY, 1984, 56 (08) :967-978