Comparative analysis of amino acid repeats in rodents and humans

被引:138
作者
Albà, MM
Guigó, R
机构
[1] Univ Pompeu Fabra, Inst Municipal Invest Med, Dept Ciencias Expt & Salut, Grp Recerca Informat Biomed, Barcelona 08003, Spain
[2] Ctr Regulacio Genom, Barcelona 08003, Spain
关键词
D O I
10.1101/gr.1925704
中图分类号
Q5 [生物化学]; Q7 [分子生物学];
学科分类号
071010 ; 081704 ;
摘要
Amino acid tandem repeats, also called homopolymeric tracts, are extremely abundant in eukaryotic proteins. To gain insight into the genome-wide evolution of these regions in mammals, we analyzed the repeat content in a large data set of rat-mouse-human orthologs. Our results show that human proteins contain more amino acid repeats than rodent proteins and that trinucleotide repeats are also more abundant in human coding sequences. Using the human species as ail outgroup, we were able to address differences in repeat loss and repeat gain in the rat and mouse lineages. In this data set, mouse proteins contain Substantially more repeats than rat proteins, which call be at least partly attributed to a higher repeat loss in the rat lineage. The data are consistent with a role for trinucleotide slippage in the generation of novel amino acid repeats. We confirm the previously observed functional bias of proteins with repeats, with overrepresentation of transcription factors and DNA-binding proteins. We show that genes encoding amino acid repeats tend to have all unusually high GC content, and that differences in coding CC content among orthologs are directly related to the presence/absence of repeats. We propose that the different GC content isochore structure in rodents and humans may result in an increased amino acid repeat prevalence in the human lineage.
引用
收藏
页码:549 / 554
页数:6
相关论文
共 34 条
  • [31] Initial sequencing and comparative analysis of the mouse genome
    Waterston, RH
    Lindblad-Toh, K
    Birney, E
    Rogers, J
    Abril, JF
    Agarwal, P
    Agarwala, R
    Ainscough, R
    Alexandersson, M
    An, P
    Antonarakis, SE
    Attwood, J
    Baertsch, R
    Bailey, J
    Barlow, K
    Beck, S
    Berry, E
    Birren, B
    Bloom, T
    Bork, P
    Botcherby, M
    Bray, N
    Brent, MR
    Brown, DG
    Brown, SD
    Bult, C
    Burton, J
    Butler, J
    Campbell, RD
    Carninci, P
    Cawley, S
    Chiaromonte, F
    Chinwalla, AT
    Church, DM
    Clamp, M
    Clee, C
    Collins, FS
    Cook, LL
    Copley, RR
    Coulson, A
    Couronne, O
    Cuff, J
    Curwen, V
    Cutts, T
    Daly, M
    David, R
    Davies, J
    Delehaunty, KD
    Deri, J
    Dermitzakis, ET
    [J]. NATURE, 2002, 420 (6915) : 520 - 562
  • [32] AN IMMUNOHISTOCHEMICAL STUDY OF THE INCIDENCE AND SIGNIFICANCE OF C-ERBB-2 ONCOPROTEIN OVEREXPRESSION IN OVARIAN NEOPLASIA
    WILKINSON, N
    TODD, N
    BUCKLEY, CH
    GUSTERSON, BA
    FOX, H
    [J]. INTERNATIONAL JOURNAL OF GYNECOLOGICAL CANCER, 1991, 1 (06) : 285 - 289
  • [33] Glutamine-rich domains activate transcription in yeast Saccharomyces cerevisiae
    Xiao, H
    Jeang, KT
    [J]. JOURNAL OF BIOLOGICAL CHEMISTRY, 1998, 273 (36) : 22873 - 22876
  • [34] Young ET, 2000, GENETICS, V154, P1053