Parametric subspace modeling of speech transitions

K Reinhard,M Niranjan

doi:10.1016/s0167-6393(98)00067-3

Abstract

This paper describes an attempt at capturing segmental transition information for speech recognition tasks. The slowly varying dynamics of spectral trajectories carries much discriminant information that is very crudely modelled by traditional approaches such as HMMs. In approaches such as recurrent neural networks there is the hope, but not the convincing demonstration, that such transitional information could be captured. The method presented here starts from the very different position of explicitly capturing the trajectory of short time spectral parameter vectors on a subspace in which the temporal sequence information is preserved. This was approached by introducing a temporal constraint into the well known technique of Principal Component Analysis (PCA). On this subspace, an attempt of parametric modelling the trajectory was made, and a distance metric was computed to perform classification of diphones. Using the Principal Curves method of Hastie and Stuetzle and the Generative Topographic map (GTM) technique of Bishop, Svensen and Williams as description of the temporal evolution in terms of latent variables was performed. On the difficult problem of /bee/, /dee/, /gee/ it was possible to retain discriminatory information with a small number of parameters. Experimental illustrations present results on ISOLET and TIMIT database.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Parametric subspace modeling of speech transitions

Abstract

Talk to us

Similar Papers

More From: Speech Communication

Lead the way for us

Journal: Speech Communication	Publication Date: Feb 1, 1999
Citations: 32

Similar Papers

Parametric subspace modelling of speech transitions
K Reinhard ... M Niranjan
-
K Reinhard, et. al.K Reinhard ... M Niranjan
12 May 1998
12 May 1998

Assessment of input variables determination on the SVM model performance using PCA, Gamma test, and forward selection techniques for monthly stream flow prediction
R Noori ... M Ghafari Gousheh
Journal of Hydrology | VOL. 401
R Noori, et. al.R Noori ... M Ghafari Gousheh
22 Feb 2011
Journal of Hydrology | VOL. 401

Hybrid Techniques for Arabic Letter Recognition
Mohamed Hassine
International Journal of Intelligent Information Systems | VOL. 4
Mohamed HassineMohamed Hassine
01 Jan 2015
International Journal of Intelligent Information Systems | VOL. 4

Assessment of water quality variations under non-rainy and rainy conditions by principal component analysis techniques in Lake Doam watershed, Korea
Bal Dev Bhattrai ... Woomyung Heo
Journal of Ecology and Environment | VOL. 38
Bal Dev Bhattrai, et. al.Bal Dev Bhattrai ... Woomyung Heo
28 May 2015
Journal of Ecology and Environment | VOL. 38

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Parametric subspace modeling of speech transitions

Abstract

Talk to us

Similar Papers

More From: Speech Communication