Quantifying the value of pronunciation lexicons for keyword search in lowresource languages

Guoguo Chen,Jan Trmal,David Yarowsky,Sanjeev Khudanpur,Daniel Povey,Oguz Yilmaz

doi:10.1109/icassp.2013.6639336

Abstract

This paper quantifies the value of pronunciation lexicons in large vocabulary continuous speech recognition (LVCSR) systems that support keyword search (KWS) in low resource languages. State-of-the-art LVCSR and KWS systems are developed for conversational telephone speech in Tagalog, and the baseline lexicon is augmented via three different grapheme-to-phoneme models that yield increasing coverage of a large Tagalog word-list. It is demonstrated that while the increased lexical coverage - or reduced out-of-vocabulary (OOV) rate - leads to only modest (ca 1%-4%) improvements in word error rate, the concomitant improvements in actual term weighted value are as much as 60%. It is also shown that incorporating the augmented lexicons into the LVCSR system before indexing speech is superior to using them post facto, e.g., for approximate phonetic matching of OOV keywords in pre-indexed lattices. These results underscore the disproportionate importance of automatic lexicon augmentation for KWS in morphologically rich languages, and advocate for using them early in the LVCSR stage.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Quantifying the value of pronunciation lexicons for keyword search in lowresource languages

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Exemplar-Based Sparse Representation Features: From TIMIT to LVCSR
Tara N Sainath ... David Nahamoo
IEEE Transactions on Audio, Speech, and Language Processing | VOL. 19
Tara N Sainath, et. al.Tara N Sainath ... David Nahamoo
01 Nov 2011
IEEE Transactions on Audio, Speech, and Language Processing | VOL. 19

Automatic generation of subword units for speech recognition systems
R Singh ... R.M Stern
IEEE Transactions on Speech and Audio Processing | VOL. 10
R Singh, et. al.R Singh ... R.M Stern
01 Jan 2002
IEEE Transactions on Speech and Audio Processing | VOL. 10

Acoustic models of the elderly for large‐vocabulary continuous speech recognition
Akira Baba ... Shinichi Yoshizawa
Electronics and Communications in Japan (Part II: Electronics) | VOL. 87
Akira Baba, et. al.Akira Baba ... Shinichi Yoshizawa
09 Jun 2004
Electronics and Communications in Japan (Part II: Electronics) | VOL. 87

Language model cross adaptation for LVCSR system combination
X Liu ... P.C Woodland
Computer Speech & Language | VOL. 27
X Liu, et. al.X Liu ... P.C Woodland
27 Jul 2012
Computer Speech & Language | VOL. 27

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Quantifying the value of pronunciation lexicons for keyword search in lowresource languages

Abstract

Talk to us

Similar Papers