Related Experiment Video
Updated: Jun 25, 2025

Inter-Brain Synchrony in Open-Ended Collaborative Learning: An fNIRS-Hyperscanning Study
Published on: July 21, 2021
AdaDFKD: Exploring adaptive inter-sample relationship in data-free knowledge distillation
Jingru Li1, Sheng Zhou1, Liangcheng Li1
1College of Computer Science and Technology, Zhejiang University, Zheda Rd., Hangzhou, 310027, Zhejiang, China.
Data-free knowledge distillation (DFKD) methods generate pseudo samples for training when data is unavailable. AdaDFKD improves DFKD by adaptively learning relationships among pseudo samples, enhancing model performance and reducing reliance on the teacher model.
Area of Science:
- Artificial Intelligence
- Machine Learning
- Deep Learning
Background:
- Data-free knowledge distillation (DFKD) enables model training without direct data access, crucial for privacy and large-scale transmission.
- Existing DFKD methods struggle with static distributions and teacher model dependency.
- Instance-level distribution learning limits adaptability in prior DFKD approaches.
Purpose of the Study:
- Introduce AdaDFKD, a novel DFKD approach.
- Address limitations of static distributions and teacher model reliance in DFKD.
- Develop an adaptive DFKD method that utilizes relationships among pseudo samples.
Main Methods:
- Generate pseudo samples adaptively from easy-to-hard.
- Employ a relationship refinement module (R2M) for optimizing pseudo-sample generation.
- Learn a progressive conditional distribution of negative samples and maximize inter-sample similarity.
Main Results:
- AdaDFKD demonstrates superiority over state-of-the-art DFKD methods.
- Achieved strong performance across diverse benchmarks and model pairs.
- Exhibited robustness and fast convergence properties.
Conclusions:
- AdaDFKD effectively mitigates risks associated with traditional DFKD.
- The proposed method enhances knowledge distillation by learning adaptive pseudo-sample relationships.
- AdaDFKD offers a more robust and efficient solution for data-free knowledge distillation.
More Related Videos
07:59Author Spotlight: Alignment of Synchronized Time-Series Data Using the Characterizing Loss of Cell Cycle Synchrony Model for Cross-Experiment Comparisons
Published on: June 9, 2023
05:15The Spatial Memory Game: Testing the Relationship Between Spatial Language, Object Knowledge, and Spatial Cognition
Published on: February 19, 2018
Related Concept Videos
One-Way ANOVA: Equal Sample Sizes
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
Sampling Distribution
One-Way ANOVA: Unequal Sample Sizes
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Downsampling
The Fourier transform of the decimated sequence reveals a combination of scaled and shifted versions of the original spectrum. This...
Sampling Theorem