Identification of OBO nonalignments and its implications for OBO enrichment
Michael Bada1, Lawrence Hunter
1Department of Pharmacology, University of Colorado at Denver, MS 8303, RC-1 South, 12801 East 17th Avenue, L18-6400A, PO Box 6511, Aurora, CO 80045, USA. mike.bada@uchsc.edu
Motivation:
Existing projects that focus on the semiautomatic addition of links between existing terms in the Open Biomedical Ontologies can take advantage of reasoners that can make new inferences between terms that are based on the added formal definitions and that reflect nonalignments between the linked terms. However, these projects require that these definitions be necessary and sufficient, a strong requirement that often does not hold. If such definitions cannot be added, the reasoners cannot point to the nonalignments through the suggestion of new inferences.
Results:
We describe a methodology by which we have identified over 1900 instances of nonredundant nonalignments between terms from the Gene Ontology (GO) biological process (BP), cellular component (CC) and molecular function (MF) ontologies, Chemical Entities of Biological Interest (ChEBI) and the Cell Type Ontology (CL). Many of the 39.8% of these nonalignments whose object terms are more atomic than the subject terms are not currently examined in other ontology-enrichment projects due to the fact that the necessary and sufficient conditions required for the inferences are not currently examined. Analysis of the ratios of nonalignments to assertions from which the nonalignments were identified suggests that BP-MF, BP-BP, BP-CL and CC-CC terms are relatively well-aligned, while ChEBI-MF, BP-ChEBI and CC-MF terms are relatively not aligned well. We propose four ways to resolve an identified nonalignment and recommend an analogous implementation of our methodology in ontology-enrichment tools to identify types of nonalignments that are currently not detected.
Availability:
The nonalignments discussed in this article may be viewed at http://compbio.uchsc.edu/Hunter_lab/Bada/nonalignments_2008_03_06.html. Code for the generation of these nonalignments is available upon request.
Contact:
mike.bada@uchsc.edu.
More Related Videos
08:29SUMO-Binding Entities SUBEs as Tools for the Enrichment, Isolation, Identification, and Characterization of the SUMO Proteome in Liver Cancer
Published on: November 1, 2019
09:40Identification of a Murine Erythroblast Subpopulation Enriched in Enucleating Events by Multi-spectral Imaging Flow Cytometry
Published on: June 6, 2014
Related Concept Videos
Methods of Classification and Identification
Peptide Identification Using Tandem Mass Spectrometry
This technique helps gather information regarding the protein from which the peptide was obtained and to study the peptides’ amino acid sequence. Identifying peptides from a complex mixture is an important component of the growing field of...
The Soil Ecosystem
Classification of Skeletal Muscle Fibers
Slow-Twitch Muscle Fibers
Slow oxidative, muscle fibers appear red due to large numbers of capillaries and high levels of...
Additional Subnuclear Structures
The nucleus contains many membrane-less subnuclear organelles or nuclear bodies, such as nucleoli, Cajal bodies, speckles,...
Ethics in Research
