Related Experiment Video
Updated: May 11, 2025

Large-scale Reconstructions and Independent, Unbiased Clustering Based on Morphological Metrics to Classify Neurons in Selective Populations
Published on: February 15, 2017
Two antagonistic objectives for one multi-scale graph clustering framework
Bruno Gaume1,2, Ixandra Achitouv3, David Chavalarias4,5
1Cognition, Langues, Langage, Ergonomie (CLLE, UMR 5263), CNRS, Paris, France. bruno.gaume@iscpif.fr.
Abstract:
In the current state of knowledge, there is no consensus on an objective criterion for evaluating network communities as cohesive sets of nodes with the following two properties: [Formula: see text] Each community is Densely Connected; [Formula: see text] Communities are Weakly Connected to each other. This makes it difficult to conduct comparative studies between dozens of graph clustering methods proposed over more than 20 years. To fill this gap: We propose a graph clustering framework by faithfully formalizing [Formula: see text] with precision and [Formula: see text] with recall, which are two meaningful metrics, simple, well known and already widely used for many tasks in most sciences. The meaning of these metrics in the context of graph clustering is therefore easily interpretable by most users of real-world graphs. We show that for most graphs, these two metrics are antagonistic, i.e. there is no solution that simultaneously maximizes precision and recall. In other words, to select a clustering among the Pareto optimal solutions (clusterings such that no other clustering exist that both increases the precision and the recall) we must first make a subjective compromise, according to our needs between the two properties [Formula: see text] and [Formula: see text]. We then show how to use this framework to compare, even without 'ground truth', the performances of five hitherto incommensurable state-of-the-art clustering methods, as well as that of a new family of clustering methods inspired by our approach.
Related Concept Videos
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Multiple Bar Graph
Each bar or column in the multiple bar graph represents a data value. These graphs are used primarily in interrelating two or more sets of data. The categories of different kinds of data are listed along the horizontal or x-axis, whereas...
Scaling
Friedman Two-way Analysis of Variance by Ranks
Robbers Cave
Comparing the Survival Analysis of Two or More Groups

