Related Experiment Video
Updated: Oct 17, 2025

Large-scale Reconstructions and Independent, Unbiased Clustering Based on Morphological Metrics to Classify Neurons in Selective Populations
Published on: February 15, 2017
Fuzzy Overclustering: Semi-Supervised Classification of Fuzzy Labels with Overclustering and Inverse Cross-Entropy
Lars Schmarje1, Johannes Brünger1, Monty Santarossa1
1Multimedia Information Processing Group, Kiel University, 24118 Kiel, Germany.
This study introduces a new semi-supervised learning framework to address fuzzy labels in underwater image classification. The novel approach uses overclustering to improve prediction consistency, outperforming existing methods on real-world plankton data.
Area of Science:
- Computer Science
- Machine Learning
- Image Classification
Background:
- Deep learning requires large labeled datasets, a challenge for underwater image classification.
- Existing semi-supervised methods struggle with fuzzy labels common in uncurated real-world data due to ambiguous class boundaries.
- Fuzzy labels arise from limited image information and transitional object stages, leading to expert disagreement.
Purpose of the Study:
- To propose a novel framework for semi-supervised classification that effectively handles fuzzy labels.
- To introduce a new loss function that enhances the overclustering capability for improved fuzzy label classification.
- To demonstrate the superiority of the proposed framework over state-of-the-art methods for real-world fuzzy-labeled datasets.
Main Methods:
- Developed a novel semi-supervised classification framework utilizing overclustering to identify substructures within fuzzy labels.
- Introduced a new loss function designed to enhance the overclustering performance of the framework.
- Evaluated the framework on real-world plankton datasets exhibiting fuzzy labels.
Main Results:
- The proposed framework demonstrated superior performance compared to existing state-of-the-art semi-supervised methods.
- Overclustering proved beneficial for handling fuzzy labels in underwater image classification tasks.
- Achieved 5-10% more consistent predictions of substructures within the fuzzy-labeled data.
Conclusions:
- The novel framework effectively addresses the challenge of semi-supervised classification with fuzzy labels.
- Overclustering is a viable strategy for improving classification accuracy and consistency in ambiguous datasets.
- The method shows significant promise for applications in underwater image analysis and other domains with uncurated data.
More Related Videos
12:08From Voxels to Knowledge: A Practical Guide to the Segmentation of Complex Electron Microscopy 3D-Data
Published on: August 13, 2014
06:01Visualization and Quantification of High-Dimensional Cytometry Data using Cytofast and the Upstream Clustering Methods FlowSOM and Cytosplore
Published on: December 12, 2019
Related Concept Videos
Classification of Systems-I
Homogeneity dictates that if an input x(t) is multiplied by a constant c, the output y(t) is multiplied by the same constant. Mathematically, this is expressed as:
Classification of Systems-II
Aggregates Classification
Petrographic classification groups aggregates based on common mineralogical characteristics. Some of the common mineral groups found in aggregates are...
Classification of Signals
A continuous-time signal holds a value at every instant in time, representing information seamlessly. In contrast, a discrete-time signal holds values only at specific moments, often denoted as x(n), where...
Survival Tree
Building a Survival Tree
Constructing a...
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...