Related Experiment Video
Updated: Mar 4, 2026

Sound Source Localization Testing in Single-sided Deafness Following Bone Conduction Intervention
Published on: December 20, 2024
A conditional diffusion-based model for high-resolution acoustic source mapping
Haobo Jia1,2, Feiran Yang3, Jianfei Tong1
1Laboratory of Noise and Audio Research, Institute of Acoustics, Chinese Academy of Sciences, Beijing 100190, China.
Abstract:
Diffusion models have recently shown strong generative capabilities in inverse imaging problems. This paper introduces the first diffusion-based framework for acoustic source mapping that directly solves the deconvolution approach for the mapping of acoustic sources inverse problem. Supervised regression-based learning methods cannot well model the sparse and peak-shaped source distributions, but the proposed generative model explicitly learns the structural prior of source maps and thus avoids blurry artifacts in the output map. During training, the diffusion model is conditioned on both the delay-and-sum beamforming map and multi-scale point spread function features extracted by an autoencoder. The beamforming map provides coarse spatial cues on source positions and strengths, while the point spread function features provide frequency-aware information. The target map is a smoothed form of sparse source labels to help the model capture structural priors, and a time-weighted loss is proposed to help model better exploit the conditions. During inference, the model can generate high-resolution source distribution maps in only 20 sampling steps. Experimental results on three generalization tasks, i.e., unseen frequencies, unseen numbers of sources, and real-world transfer functions, demonstrate that the proposed method outperforms existing traditional and supervised regression-based deep learning-based approaches.
Related Concept Videos
Echo
Imagine the sound is reflected back to the ears. Assuming that the source is very close to the human, the difference between hearing the two sounds—the emitted sound and the reflected sound—may be more than the minimum time for perceiving distinct sounds. If this is the case,...
Perceiving Loudness, Pitch, and Location
Place theory, or place coding, suggests that different pitches are heard because various sound waves activate specific locations along the cochlea's basilar membrane. The brain determines the pitch of a sound by...
Linear Approximation in Frequency Domain
In contrast, nonlinear systems do not inherently possess these properties. However, for small deviations around an operating point, a nonlinear system can often be approximated as linear....
Difference from Background: Limit of Detection
The LOD indicates the presence or absence...
Doppler Effect - II

