A Multi-channel Speech Separation System for Unknown Number of Multiple Speakers

Chao Peng,Yiwen Wang,Xihong Wu,Tianshu Qu

doi:10.1109/icicsp55539.2022.10050619

Abstract

This paper presents a multi-channel speech separation system for an unknown number of speakers. It can be applied to cases with a different number of speakers using a single model by iterative speech separation based on beam signal. It first determines the spatial directions where speakers are located (Direction of Arrival, DOA), and then the beam signals in each direction are obtained with spectral features, spatial features, and directional features by deep neural networks. Finally, the iterative speech separation is performed on the basis of the beam signals. Experimental evaluations show that the proposed method is better than the multi-channel Permutation Invariant Training (PIT) and Deep Clustering (DPCL) for an unknown number of speakers and the one-and-rest speech separation method. Besides, the system can still keep a relatively good separation performance even though the number of speakers is enlarged to 9.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A Multi-channel Speech Separation System for Unknown Number of Multiple Speakers

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Boosting Spatial Information for Deep Learning Based Multichannel Speaker-Independent Speech Separation In Reverberant Environments
Ziye Yang ... Xiao-Lei Zhang
-
Ziye Yang, et. al.Ziye Yang ... Xiao-Lei Zhang
01 Nov 2019
01 Nov 2019

Single-Channel Speech Separation Using Soft-Minimum Permutation Invariant Training
Midia Yousefi ... John H.L Hansen
SSRN Electronic Journal | VOL. -
Midia Yousefi, et. al.Midia Yousefi ... John H.L Hansen
01 Jan 2021
SSRN Electronic Journal | VOL. -

Single-channel speech separation using soft-minimum permutation invariant training
Midia Yousefi ... John H.L Hansen
Speech Communication | VOL. 151
Midia Yousefi, et. al.Midia Yousefi ... John H.L Hansen
18 May 2023
Speech Communication | VOL. 151

Multi-Microphone Speaker Separation based on Deep DOA Estimation
Shlomo E. Chazan ... Jacob Goldberger
-
Shlomo E. Chazan, et. al.Shlomo E. Chazan ... Jacob Goldberger
01 Sep 2019
01 Sep 2019

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A Multi-channel Speech Separation System for Unknown Number of Multiple Speakers

Abstract

Talk to us

Similar Papers