Sampling the arabidopsis transcriptome with massively parallel pyrosequencing

被引:237
作者
Weber, Andreas P. M.
Weber, Katrin L.
Carr, Kevin
Wilkerson, Curtis
Ohlrogge, John B. [1 ]
机构
[1] Michigan State Univ, Dept Plant Biol, E Lansing, MI 48824 USA
[2] Michigan State Univ, Bioinformat Support Core, Res Technol Support Facil, E Lansing, MI 48824 USA
关键词
D O I
10.1104/pp.107.096677
中图分类号
Q94 [植物学];
学科分类号
071001 ;
摘要
Massively parallel sequencing of DNA by pyrosequencing technology offers much higher throughput and lower cost than conventional Sanger sequencing. Although extensively used already for sequencing of genomes, relatively few applications of massively parallel pyrosequencing to transcriptome analysis have been reported. To test the ability of this technology to provide unbiased representation of transcripts, we analyzed mRNA from Arabidopsis ( Arabidopsis thaliana) seedlings. Two sequencing runs yielded 541,852 expressed sequence tags (ESTs) after quality control. Mapping of the ESTs to the Arabidopsis genome and to The Arabidopsis Information Resource 7.0 cDNA models indicated: (1) massively parallel pyrosequencing detected transcription of 17,449 gene loci providing very deep coverage of the transcriptome. Performing a second sequencing run only increased the number of genes identified by 10%, but increased the overall sequence coverage by 50%. (2) Mapping of the ESTs to their predicted full-length transcripts indicated that all regions of the transcript were well represented regardless of transcript length or expression level. Furthermore, short, medium, and long transcripts were equally represented. ( 3) Over 16,000 of the ESTs that mapped to the genome were not represented in the existing dbEST database. In some cases, the ESTs provide the first experimental evidence for transcripts derived from predicted genes, and, for at least 60 locations in the genome, pyrosequencing identified likely protein-coding sequences that are not now annotated as genes. Together, the results indicate massively parallel pyrosequencing provides novel information helpful to improve the annotation of the Arabidopsis genome. Furthermore, the unbiased representation of transcripts will be particularly useful for gene discovery and gene expression analysis of nonmodel plants with less complete genomic information.
引用
收藏
页码:32 / 42
页数:11
相关论文
共 34 条
[1]   Pyrosequencing: History, biochemistry and future [J].
Ahmadian, A ;
Ehn, M ;
Hober, S .
CLINICA CHIMICA ACTA, 2006, 363 (1-2) :83-94
[2]  
[Anonymous], 2003, Proceedings of ClusterWorld 2003
[3]   The significance of digital gene expression profiles [J].
Audic, S ;
Claverie, JM .
GENOME RESEARCH, 1997, 7 (10) :986-995
[4]   Arabidopsis thaliana proteomics:: from proteome to genome [J].
Baginsky, S ;
Gruissem, W .
JOURNAL OF EXPERIMENTAL BOTANY, 2006, 57 (07) :1485-1491
[5]   Analysis of the prostate cancer cell line LNCaP transcriptome using a sequencing-by-synthesis approach [J].
Bainbridge, Matthew N. ;
Warren, Rene L. ;
Hirst, Martin ;
Romanuik, Tammy ;
Zeng, Thomas ;
Go, Anne ;
Delaney, Allen ;
Griffith, Malachi ;
Hickenbotham, Matthew ;
Magrini, Vincent ;
Mardis, Elaine R. ;
Sadar, Marianne D. ;
Siddiqui, Asim S. ;
Marra, Marco A. ;
Jones, Steven J. M. .
BMC GENOMICS, 2006, 7 (1)
[6]   Carbocyclic fatty acids in plants:: Biochemical and molecular genetic characterization of cyclopropane fatty acid synthesis of Sterculia foetida [J].
Bao, XM ;
Katz, S ;
Pollard, M ;
Ohlrogge, J .
PROCEEDINGS OF THE NATIONAL ACADEMY OF SCIENCES OF THE UNITED STATES OF AMERICA, 2002, 99 (10) :7172-7177
[7]   Comparative genomics of two closely related unicellular thermo-acidophilic red algae, Galdieria sulphuraria and Cyanidioschyzon merolae, reveals the molecular basis of the metabolic flexibility of Galdieria sulphuraria and significant differences in carbohydrate metabolism of both algae [J].
Barbier, G ;
Oesterhelt, C ;
Larson, MD ;
Halgren, RG ;
Wilkerson, C ;
Garavito, RM ;
Benning, C ;
Weber, APM .
PLANT PHYSIOLOGY, 2005, 137 (02) :460-474
[8]   Gene expression analysis by massively parallel signature sequencing (MPSS) on microbead arrays [J].
Brenner, S ;
Johnson, M ;
Bridgham, J ;
Golda, G ;
Lloyd, DH ;
Johnson, D ;
Luo, SJ ;
McCurdy, S ;
Foy, M ;
Ewan, M ;
Roth, R ;
George, D ;
Eletr, S ;
Albrecht, G ;
Vermaas, E ;
Williams, SR ;
Moon, K ;
Burcham, T ;
Pallas, M ;
DuBridge, RB ;
Kirchner, J ;
Fearon, K ;
Mao, J ;
Corcoran, K .
NATURE BIOTECHNOLOGY, 2000, 18 (06) :630-634
[9]   d2_cluster: A validated method for clustering EST and full-length cDNA sequences [J].
Burke, J ;
Davison, D ;
Hide, W .
GENOME RESEARCH, 1999, 9 (11) :1135-1142
[10]   Sequencing Medicago truncatula expressed sequenced tags using 454 Life Sciences technology [J].
Cheung, Foo ;
Haas, Brian J. ;
Goldberg, Susanne M. D. ;
May, Gregory D. ;
Xiao, Yongli ;
Town, Christopher D. .
BMC GENOMICS, 2006, 7 (1)