BACKGROUND: Within the peer-reviewed literature, associations between two things are not always recognized until commonalities between them become apparent. These commonalities can provide justification for the inference of a new relationship where none was previously known, and are the basis of most observation-based hypothesis formation. It has been shown that the crux of the problem is not finding inferable associations, which are extraordinarily abundant given the scale-free networks that arise from literature-based associations, but determining which ones are informative. The Mutual Information Measure (MIM) is a well-established method to measure how informative an association is, but is limited to direct (i.e. observable) associations. RESULTS: Herein, we attempt to extend the calculation of mutual information to indirect (i.e. inferable) associations by using the MIM of shared associations. Objects of general research interest (e.g. genes, diseases, phenotypes, drugs, ontology categories) found within MEDLINE are used to create a network of associations for evaluation. CONCLUSIONS: Mutual information calculations can be effectively extended into implied relationships and a significance cutoff estimated from analysis of random word networks. Of the models tested, the shared minimum MIM (MMIM) model is found to correlate best with the observed strength and frequency of known associations. Using three test cases, the MMIM method tends to rank more specific relationships higher than counting the number of shared relationships within a network

Wren, Jonathan D

English

PubMed

Springer - Publisher Connector

Extending the mutual information measure to rank inferred literature relationships

Jonathan D Wren

Abstract Background Within the peer-reviewed literature, associations between two things are not always recognized until commonalities between them become apparent. These commonalities can provide justification for the inference of a new relationship where none was previously known, and are the basis of most observation-based hypothesis formation. It has been shown that the crux of the problem is not finding inferable associations, which are extraordinarily abundant given the scale-free networks that arise from literature-based associations, but determining which ones are informative. The Mutual Information Measure (MIM) is a well-established method to measure how informative an association is, but is limited to direct (i.e. observable) associations. Results Herein, we attempt to extend the calculation of mutual information to indirect (i.e. inferable) associations by using the MIM of shared associations. Objects of general research interest (e.g. genes, diseases, phenotypes, drugs, ontology categories) found within MEDLINE are used to create a network of associations for evaluation. Conclusions Mutual information calculations can be effectively extended into implied relationships and a significance cutoff estimated from analysis of random word networks. Of the models tested, the shared minimum MIM (MMIM) model is found to correlate best with the observed strength and frequency of known associations. Using three test cases, the MMIM method tends to rank more specific relationships higher than counting the number of shared relationships within a network.</p

Wren Jonathan D

Directory of Open Access Journals

BMC Bioinformatics

Alanine AI: Hit and lead generation: beyond high-throughput screening. Nat Rev Drug Discov

DM: Fish-oil dietary supplementation in patients with Raynaud's phenomenon: a double-blind, controlled, prospective study.

Fish oil, Raynaud's syndrome, and undiscovered public knowledge. Perspect Biol Med

INTERNATIONAL SEQUENCING CONSORTIUM: Initial sequencing and analysis of the human genome. Nature

MEDLINE fact sheet [http://www.nlm.nih.gov/pubs/ factsheets/medline.html].

Schoolnik GK: Microarray expression profiling: capturing a genome-wide portrait of the transcriptome. Mol Microbiol

file:///data/remote/core/dit/data/Springer-OA/pdf/ef5/aHR0cDovL2xpbmsuc3ByaW5nZXIuY29tLzEwLjExODYvMTQ3MS0yMTA1LTUtMTQ1LnBkZg==.pdf

Extending the mutual information measure to rank inferred literature relationships

Abstract

Similar works

Full text

Available Versions

Springer - Publisher Connector

Springer - Publisher Connector

Directory of Open Access Journals