Croatian Large Vocabulary Automatic Speech Recognition

Sandaasst Prof Martinčić-Ipšić,Miran Pobar,Ivoprof Ipšić

doi:10.1080/00051144.2011.11828413

Sandaasst Prof Martinčić-Ipšić, Miran Pobar + Show 1 more

Open Access

https://doi.org/10.1080/00051144.2011.11828413

Copy DOI

Journal: Automatika	Publication Date: Jan 1, 2011
Citations: 5	License type: cc-by-nc

Affiliation: University of Rijeka

Abstract

This paper presents procedures used for development of a Croatian large vocabulary automatic speech recognition system (LVASR). The proposed acoustic model is based on context-dependent triphone hidden Markov models and Croatian phonetic rules. Different acoustic and language models, developed using a large collection of Croatian speech, are discussed and compared. The paper proposes the best feature vectors and acoustic modeling procedures using which lowest word error rates for Croatian speech are achieved. In addition, Croatian language modeling procedures are evaluated and adopted for speaker independent spontaneous speech recognition. Presented experiments and results show that the proposed approach for automatic speech recognition using context-dependent acoustic modeling based on Croatian phonetic rules and a parameter tying procedure can be used for efficient Croatian large vocabulary speech recognition with word error rates below 5%.

Full Text