Bangla phoneme recognition using hybrid features

Mohammed Rokibul Alam Kotwal,Foyzul Hassan,Chowdhury Mofizur Rahman,Md Shahadat Hossain,Ghulam Muhammad,Moahammad Nurul Huda

doi:10.1109/icelce.2010.5700793

Mohammed Rokibul Alam Kotwal, Foyzul Hassan + Show 4 more

https://doi.org/10.1109/icelce.2010.5700793

Copy DOI

Abstract

This paper presents a Bangla phoneme recognition method for Automatic Speech Recognition (ASR). The method consists of three stages: i) a multilayer neural network (MLN), which converts acoustic features, mel frequency cepstral coefficients (MFCCs), into phoneme probabilities, ii) the phoneme probabilities obtained from the first stage and corresponding Δ and ΔΔ are inserted into another MLN to improve the phoneme probabilities by reducing the context effect and (iii) the phoneme probabilities of current frame and corresponding MFCCs are fed into a hidden Markov model (HMM) based classifier to obtain more accurate phoneme strings. From the experiments on Bangla speech corpus prepared by us, it is observed that the proposed method provides higher phoneme recognition performance than the existing method. Moreover, it requires a fewer mixture components in the HMMs.

Full Text