Related Experiment Video
Updated: Mar 4, 2026

Sound Source Localization Testing in Single-sided Deafness Following Bone Conduction Intervention
Published on: December 20, 2024
A conditional diffusion-based model for high-resolution acoustic source mapping
Haobo Jia1,2, Feiran Yang3, Jianfei Tong1
1Laboratory of Noise and Audio Research, Institute of Acoustics, Chinese Academy of Sciences, Beijing 100190, China.
This study introduces a novel diffusion-based framework for acoustic source mapping, improving deconvolution accuracy. The generative model effectively captures source structures, outperforming existing methods in generalization tasks.
Area of Science:
- Acoustics
- Signal Processing
- Machine Learning
Background:
- Inverse imaging problems, particularly acoustic source mapping, are challenging due to sparse and peak-shaped source distributions.
- Traditional supervised regression methods struggle to accurately model these complex source structures, often resulting in blurry artifacts.
- Diffusion models offer powerful generative capabilities applicable to inverse problems.
Purpose of the Study:
- To introduce the first diffusion-based framework for acoustic source mapping that directly addresses the deconvolution inverse problem.
- To develop a generative model that explicitly learns the structural prior of source maps, avoiding blurry artifacts.
- To enhance acoustic source mapping accuracy and generalization capabilities.
Main Methods:
- A diffusion-based generative framework conditioned on delay-and-sum beamforming maps and autoencoder-extracted multi-scale point spread function features.
- Utilizing a smoothed target map to guide the model in capturing structural priors.
- Implementing a time-weighted loss function to improve condition exploitation during training.
- Employing an autoencoder to extract frequency-aware features from point spread functions.
Main Results:
- The proposed diffusion model successfully generates high-resolution acoustic source distribution maps with only 20 sampling steps during inference.
- Experimental results demonstrate superior performance compared to traditional and supervised regression-based deep learning methods.
- The framework shows strong generalization across unseen frequencies, varying numbers of sources, and real-world transfer functions.
Conclusions:
- The diffusion-based framework represents a significant advancement in acoustic source mapping, overcoming limitations of previous methods.
- The model's ability to learn structural priors and avoid blurry artifacts leads to more accurate and detailed source maps.
- This approach offers a robust and generalizable solution for acoustic source mapping in complex scenarios.
Related Concept Videos
Echo
Imagine the sound is reflected back to the ears. Assuming that the source is very close to the human, the difference between hearing the two sounds—the emitted sound and the reflected sound—may be more than the minimum time for perceiving distinct sounds. If this is the case,...
Perceiving Loudness, Pitch, and Location
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by...
Linear Approximation in Frequency Domain
In contrast, nonlinear systems do not inherently possess these properties. However, for small deviations around an operating point, a nonlinear system can often be approximated as linear....
Difference from Background: Limit of Detection
The LOD indicates the presence or absence...
Doppler Effect - II

