Bi-Spectral Acoustic Features for Robust Speech Recognition

K Onoe,S Sato,S Homma,A Kobayashi,T Imai,T Takagi

doi:10.1093/ietisy/e91-d.3.631

Bi-Spectral Acoustic Features for Robust Speech Recognition

K Onoe, S Sato + Show 4 more

Open Access

https://doi.org/10.1093/ietisy/e91-d.3.631

Copy DOI

Journal: IEICE Transactions on Information and Systems	Publication Date: Mar 1, 2008
Citations: 9	License type: free

#Features For Robust Speech Recognition #Conventional Mel Frequency Cepstral Coefficients + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

The extraction of acoustic features for robust speech recognition is very important for improving its performance in realistic environments. The bi-spectrum based on the Fourier transformation of the third-order cumulants expresses the non-Gaussianity and the phase information of the speech signal, showing the dependency between frequency components. In this letter, we propose a method of extracting short-time bispectral acoustic features with averaging features in a single frame. Merged with the conventional Mel frequency cepstral coefficients (MFCC) based on the power spectrum by the principal component analysis (PCA), the proposed features gave a 6.9% relative lower a word error rate in Japanese broadcast news transcription experiments.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: IEICE Transactions on Information and Systems

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.