Related Experiment Video
Updated: Jun 16, 2026

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
Data-fusion in clustering microarray data: balancing discovery and interpretability
1Department of Public Health Sciences, University of Toronto, Health Sciences Building, Toronto, Ontario, Canada. r.kustra@utoronto.ca
Abstract:
While clustering genes remains one of the most popular exploratory tools for expression data, it often results in a highly variable and biologically uninformative clusters. This paper explores a data fusion approach to clustering microarray data. Our method, which combined expression data and Gene Ontology (GO)-derived information, is applied on a real data set to perform genome-wide clustering. A set of novel tools is proposed to validate the clustering results and pick a fair value of infusion coefficient. These tools measure stability, biological relevance, and distance from the expression-only clustering solution. Our results indicate that a data-fusion clustering leads to more stable, biologically relevant clusters that are still representative of the experimental data.
Related Concept Videos
DNA Microarrays
Tagging and Fusion Proteins