Improvements in linear transform based speaker adaptation

L.F Uebel,P.C Woodland

doi:10.1109/icassp.2001.940764

Improvements in linear transform based speaker adaptation

L.F Uebel, P.C Woodland

https://doi.org/10.1109/icassp.2001.940764

Copy DOI

Publication Date: May 7, 2001

Citations: 64

Affiliation: University of Cambridge

#Maximum Likelihood Linear Regression #Wall Street Journal Database + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

Presents three forms of linear transform based speaker adaptation that can give better performance than standard maximum likelihood linear regression (MLLR) adaptation. For unsupervised adaptation, a lattice-based technique is introduced which is compared to MLLR using confidence scores. For supervised adaptation, estimation of the adaptation matrices using the maximum mutual information criterion is discussed which leads to the MMILR approach. Recognition experiments show that lattice MLLR can reduce word error rates on a Switchboard task by 1.4% absolute. For recognition of non-native speech from the Wall Street Journal database, a reduction in word error rate of 10-16% relative was obtained using MMILR compared to standard MLLR.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.