Extending the mutual information measure to rank inferred literature relationships

被引:73
作者
Wren, JD [1 ]
机构
[1] Univ Oklahoma, Dept Bot & Microbiol, Adv Ctr Genome Technol, Norman, OK 73019 USA
关键词
D O I
10.1186/1471-2105-5-145
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Background: Within the peer-reviewed literature, associations between two things are not always recognized until commonalities between them become apparent. These commonalities can provide justification for the inference of a new relationship where none was previously known, and are the basis of most observation-based hypothesis formation. It has been shown that the crux of the problem is not finding inferable associations, which are extraordinarily abundant given the scale-free networks that arise from literature-based associations, but determining which ones are informative. The Mutual Information Measure (MIM) is a well-established method to measure how informative an association is, but is limited to direct (i.e. observable) associations. Results: Herein, we attempt to extend the calculation of mutual information to indirect (i.e. inferable) associations by using the MIM of shared associations. Objects of general research interest (e.g. genes, diseases, phenotypes, drugs, ontology categories) found within MEDLINE are used to create a network of associations for evaluation. Conclusions: Mutual information calculations can be effectively extended into implied relationships and a significance cutoff estimated from analysis of random word networks. Of the models tested, the shared minimum MIM (MMIM) model is found to correlate best with the observed strength and frequency of known associations. Using three test cases, the MMIM method tends to rank more specific relationships higher than counting the number of shared relationships within a network.
引用
收藏
页数:13
相关论文
共 31 条
[1]   Gene Ontology: tool for the unification of biology [J].
Ashburner, M ;
Ball, CA ;
Blake, JA ;
Botstein, D ;
Butler, H ;
Cherry, JM ;
Davis, AP ;
Dolinski, K ;
Dwight, SS ;
Eppig, JT ;
Harris, MA ;
Hill, DP ;
Issel-Tarver, L ;
Kasarskis, A ;
Lewis, S ;
Matese, JC ;
Richardson, JE ;
Ringwald, M ;
Rubin, GM ;
Sherlock, G .
NATURE GENETICS, 2000, 25 (01) :25-29
[2]   Tachykinin receptors are involved in the "local efferent" motor response to capsaicin in the guinea-pig small intestine and oesophagus [J].
Barthó, L ;
Lenard, L ;
Patacchini, R ;
Halmai, V ;
Wilhelm, M ;
Holzer, P ;
Maggi, CA .
NEUROSCIENCE, 1999, 90 (01) :221-228
[3]  
Blaschke C, 1999, Proc Int Conf Intell Syst Mol Biol, P60
[4]   Hit and lead generation:: Beyond high-throughput screening [J].
Bleicher, KH ;
Böhm, HJ ;
Müller, K ;
Alanine, AI .
NATURE REVIEWS DRUG DISCOVERY, 2003, 2 (05) :369-378
[5]  
CANDELA M, 1994, CLIN EXP RHEUMATOL, V12, P509
[6]  
CHURCH KW, 1990, 27TH ANNUAL MEETING OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS, P76
[7]  
Conrad J. G., 1994, SIGIR '94. Proceedings of the Seventeenth Annual International ACM-SIGIR Conference on Research and Development in Information Retrieval, P260
[8]   Microarray expression profiling: capturing a genome-wide portrait of the transcriptome [J].
Conway, T ;
Schoolnik, GK .
MOLECULAR MICROBIOLOGY, 2003, 47 (04) :879-889
[9]   FISH-OIL DIETARY SUPPLEMENTATION IN PATIENTS WITH RAYNAUD PHENOMENON - A DOUBLE-BLIND, CONTROLLED, PROSPECTIVE-STUDY [J].
DIGIACOMO, RA ;
KREMER, JM ;
SHAH, DM .
AMERICAN JOURNAL OF MEDICINE, 1989, 86 (02) :158-164
[10]  
Dunning T., 1993, Computational Linguistics, V19, P61