A Contrast Function Based on Generalized Divergences for Solving the Permutation Problem in Convolved Speech Mixtures

Auxiliadora Sarmiento,Sergio Cruces,Ivan Duran-Diaz,Andrzej Cichocki

doi:10.1109/taslp.2015.2447281

Abstract

In this paper, we propose a method for solving the permutation problem that is inherent in the separation of convolved mixtures of speech signals in the time-frequency domain. The proposed method obtains the solution through maximization of a contrast function that exploits the similarity of the temporal envelope of the speech spectrum. For this purpose, the contrast calculation uses a global measure of similarity based on the recently developed family of generalized Alpha-Beta divergences, which depend on two tuning parameters, alpha and beta. This parameterization is exploited to best measure the similarity of the speech spectrum and to obtain solutions that are robust against noise and outliers. The ability of this contrast function to solve the permutation problem is supported by a theoretical study that shows that for a simple time-frequency speech model, the contrast value reaches its maximum when the estimated components are properly aligned. Several performance studies demonstrate that the proposed method maintains a high level of permutation correction accuracy in a wide variety of acoustic environments. Moreover, it produces better results than other state-of-the-art methods for solving permutations in highly reverberant environments.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A Contrast Function Based on Generalized Divergences for Solving the Permutation Problem in Convolved Speech Mixtures

Abstract

Talk to us

Similar Papers

More From: IEEE/ACM Transactions on Audio, Speech, and Language Processing

Lead the way for us

Journal: IEEE/ACM Transactions on Audio, Speech, and Language Processing	Publication Date: Nov 1, 2015
Citations: 34

Similar Papers

Blind separation of underdetermined Convolutive speech mixtures by time–frequency masking with the reduction of musical noise of separated signals
Mahbanou Zohrevandi ... Azam Rabiee
Multimedia Tools and Applications | VOL. 80
Mahbanou Zohrevandi, et. al.Mahbanou Zohrevandi ... Azam Rabiee
12 Jan 2021
Multimedia Tools and Applications | VOL. 80

New approaches for solving permutation indeterminacy and scaling ambiguity in frequency domain separation of convolved mixtures
Zhitang Chen ... Laiwan Chan
-
Zhitang Chen, et. al.Zhitang Chen ... Laiwan Chan
01 Jul 2011
01 Jul 2011

Multisynchrosqueezing Generalized S-Transform and Its Application in Tight Sandstone Gas Reservoir Identification
Xuping Chen ... Rui Li
IEEE Geoscience and Remote Sensing Letters | VOL. 19
Xuping Chen, et. al.Xuping Chen ... Rui Li
15 Dec 2020
IEEE Geoscience and Remote Sensing Letters | VOL. 19

Synchroextracting Transform
Gang Yu ... Mingjin Yu
IEEE Transactions on Industrial Electronics | VOL. 64
Gang Yu, et. al.Gang Yu ... Mingjin Yu
01 Oct 2017
IEEE Transactions on Industrial Electronics | VOL. 64

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A Contrast Function Based on Generalized Divergences for Solving the Permutation Problem in Convolved Speech Mixtures

Abstract

Talk to us

Similar Papers

More From: IEEE/ACM Transactions on Audio, Speech, and Language Processing