Maximum likelihood and minimum classification error factor analysis for automatic speech recognition

L.K Saul,M.G Rahim

doi:10.1109/89.824696

Abstract

Hidden Markov models (HMMs) for automatic speech recognition rely on high dimensional feature vectors to summarize the short-time properties of speech. Correlations between features can arise when the speech signal is nonstationary or corrupted by noise. We investigate how to model these correlations using factor analysis, a statistical method for dimensionality reduction. Factor analysis uses a small number of parameters to model the covariance structure of high dimensional data. These parameters can be chosen in two ways: (1) to maximize the likelihood of observed speech signals, or (2) to minimize the number of classification errors. We derive an expectation-maximization (EM) algorithm for maximum likelihood estimation and a gradient descent algorithm for improved class discrimination. Speech recognizers are evaluated on two tasks, one small-sized vocabulary (connected alpha-digits) and one medium-sized vocabulary (New Jersey town names). We find that modeling feature correlations by factor analysis leads to significantly increased likelihoods and word accuracies. Moreover, the rate of improvement with model size often exceeds that observed in conventional HMM's.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Maximum likelihood and minimum classification error factor analysis for automatic speech recognition

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Speech and Audio Processing

Lead the way for us

Journal: IEEE Transactions on Speech and Audio Processing	Publication Date: Mar 1, 2000
Citations: 102

Similar Papers

Low complexity source localization algorithms in sensor networks
Junichi Shirahama ... Toshinobu Kaneko
-
Junichi Shirahama, et. al.Junichi Shirahama ... Toshinobu Kaneko
10 Oct 2005
10 Oct 2005

Fuzzy hidden Markov models for speech and speaker recognition
Dat Tran ... M Wagner
-
Dat Tran, et. al. Dat Tran ... M Wagner
10 Jun 1999
10 Jun 1999

Efficient backward decoding of high-order hidden Markov models
H.A Engelbrecht ... J.A Du Preez
Pattern Recognition | VOL. 43
H.A Engelbrecht, et. al.H.A Engelbrecht ... J.A Du Preez
21 Jun 2009
Pattern Recognition | VOL. 43

Ensemble acoustic modeling in automatic speech recognition
Xin Chen
-
Xin ChenXin Chen
01 Jan 2010
01 Jan 2010

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Maximum likelihood and minimum classification error factor analysis for automatic speech recognition

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Speech and Audio Processing