Related Experiment Video
Updated: Nov 12, 2025

A Lightweight, Headphones-based System for Manipulating Auditory Feedback in Songbirds
Published on: November 26, 2012
Divide and Conquer: A Deep CASA Approach to Talker-independent Monaural Speaker Separation
1Department of Computer Science and Engineering, The Ohio State University, Columbus, OH 43210-1277 USA.
Abstract:
We address talker-independent monaural speaker separation from the perspectives of deep learning and computational auditory scene analysis (CASA). Specifically, we decompose the multi-speaker separation task into the stages of simultaneous grouping and sequential grouping. Simultaneous grouping is first performed in each time frame by separating the spectra of different speakers with a permutation-invariantly trained neural network. In the second stage, the frame-level separated spectra are sequentially grouped to different speakers by a clustering network. The proposed deep CASA approach optimizes frame-level separation and speaker tracking in turn, and produces excellent results for both objectives. Experimental results on the benchmark WSJ0-2mix database show that the new approach achieves the state-of-the-art results with a modest model size.
More Related Videos
Related Concept Videos
Interference: Path Lengths
Two special sources may be considered when they are in phase. This can be easily achieved by feeding the two sources from the same source. An example would be synchronizing the two speakers by feeding them with the same source, such as the sound waves produced by a tuning fork. This setup ensures that the two sources have the same frequency and are...
Design Example
¹H NMR Signal Multiplicity: Splitting Patterns
Impedance Combination
Sound Waves: Interference
Extraction: Partition and Distribution Coefficients
For extracting a solute from an aqueous phase into an...

