MSFDnet: A Multi-Scale Feature Dual-Layer Fusion Model for Sound Event Localization and Detection

Yi Chen1, Zhenyu Huang2, Liang Lei3

  • 1School of Big Data and Information Industry, Chongqing City Management College, No. 151, Daxuecheng South Second Road, Shapingba District, Chongqing 401331, China.

PubMed
Summary

This study introduces MSDFnet, a novel model for Sound Event Localization and Detection (SELD) that improves accuracy in complex audio by enhancing feature extraction and fusion. The new model excels in dynamic scenarios, outperforming existing methods.