Detecting differentially expressed genes in microarrays using Bayesian model selection

被引:110
作者
Ishwaran, H [1 ]
Rao, JS
机构
[1] Cleveland Clin Fdn, Dept Biostat & Epidemiol Wb4, Cleveland, OH 44195 USA
[2] Case Western Reserve Univ, Dept Epidemiol & Biostat, Cleveland, OH 44106 USA
关键词
Bayesian analysis of variance for microarrays; false discovery rate; false nondiscovery rate; heteroscedasticity; ridge; regression; Shrinkage; variance stabilizing transform; weighted regression;
D O I
10.1198/016214503000224
中图分类号
O21 [概率论与数理统计]; C8 [统计学];
学科分类号
020208 ; 070103 ; 0714 ;
摘要
DNA microarrays open up a broad new horizon for investigators interested in studying the genetic determinants of disease. The high throughput nature of these arrays, where differential expression for thousands of genes can be measured simultaneously, creates an enormous wealth of information, but also poses a challenge for data analysis because of the large multiple testing problem involved. The solution has generally been to focus on optimizing false-discovery rates while sacrificing power. The drawback of this approach is that more subtle expression differences will be missed that might give investigators more insight into the genetic environment necessary for a disease process to take hold. We introduce a new method for detecting differentially expressed genes based on a high-dimensional model selection technique, Bayesian ANOVA for microarrays (BAM), which strikes a balance between false rejections and false nonrejections. The basis of the new approach involves a weighted average of generalized ridge regression estimates that provides the benefits of using shrinkage estimation combined with model averaging. A simple graphical tool based on the amount of shrinkage is developed to visualize the trade-off between low false-discovery rates and finding more genes. Simulations are used to illustrate BAM's performance, and the method is applied to a large database of colon cancer gene expression data. Our working hypothesis in the colon cancer analysis is that large differential expressions may not be the only ones contributing to metastasis-in fact, moderate changes in expression of genes may be involved in modifying the genetic environment to a sufficient extent for metastasis to occur. A functional biological analysis of gene effects found by BAM, but not other false-discovery-based approaches, lends support to this hypothesis.
引用
收藏
页码:438 / 455
页数:18
相关论文
共 31 条