Robust text-independent speaker identification using Gaussian mixture speaker models

D.A Reynolds,R.C Rose

doi:10.1109/89.365379

Abstract

This paper introduces and motivates the use of Gaussian mixture models (GMM) for robust text-independent speaker identification. The individual Gaussian components of a GMM are shown to represent some general speaker-dependent spectral shapes that are effective for modeling speaker identity. The focus of this work is on applications which require high identification rates using short utterance from unconstrained conversational speech and robustness to degradations produced by transmission over a telephone channel. A complete experimental evaluation of the Gaussian mixture speaker model is conducted on a 49 speaker, conversational telephone speech database. The experiments examine algorithmic issues (initialization, variance limiting, model order selection), spectral variability robustness techniques, large population performance, and comparisons to other speaker modeling techniques (uni-modal Gaussian, VQ codebook, tied Gaussian mixture, and radial basis functions). The Gaussian mixture speaker model attains 96.8% identification accuracy using 5 second clean speech utterances and 80.8% accuracy using 15 second telephone speech utterances with a 49 speaker population and is shown to outperform the other speaker modeling techniques on an identical 16 speaker telephone speech task.< <ETX xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">></ETX>

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Robust text-independent speaker identification using Gaussian mixture speaker models

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Speech and Audio Processing

Lead the way for us

Journal: IEEE Transactions on Speech and Audio Processing	Publication Date: Jan 1, 1995
Citations: 2811

Similar Papers

Optimized Gaussian mixture models for upper limb motion classification
Y Huang ... B Hudgins
-
Y Huang, et. al.Y Huang ... B Hudgins
01 Jan 2004
01 Jan 2004

Gaussian Mixture Models for Higher-Order Side Channel Analysis
Kerstin Lemke-Rust ... Christof Paar
-
Kerstin Lemke-Rust, et. al.Kerstin Lemke-Rust ... Christof Paar
10 Sep 2007
10 Sep 2007

A Tutorial on Text-Independent Speaker Verification
Frédéric Bimbot ... Dijana Petrovska-Delacrétaz
EURASIP Journal on Advances in Signal Processing | VOL. 2004
Frédéric Bimbot, et. al.Frédéric Bimbot ... Dijana Petrovska-Delacrétaz
21 Apr 2004
EURASIP Journal on Advances in Signal Processing | VOL. 2004

Robust Speaker Identification using Denoised Wave Atom and GMM
Mohammed A H Lubbad ... Mahmoud Z Alkurdi
International Journal of Computer Applications | VOL. 67
Mohammed A H Lubbad, et. al.Mohammed A H Lubbad ... Mahmoud Z Alkurdi
18 Apr 2013
International Journal of Computer Applications | VOL. 67

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Robust text-independent speaker identification using Gaussian mixture speaker models

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Speech and Audio Processing