Improved vocabulary independent search with approximate match based on Conditional Random Fields

Upendra V Chaudhari,Michael Picheny

doi:10.1109/asru.2009.5373323

Abstract

We investigate the use of Conditional Random Fields (CRF) to model confusions and account for errors in the phonetic decoding derived from Automatic Speech Recognition output. The goal is to improve the accuracy of approximate phonetic match, given query terms and an indexed database of documents, in a vocabulary independent audio search system. Audio data is ingested, segmented, decoded to produce a sequence of phones, and subsequently indexed using phone N-grams. Search is performed by expanding queries into phone sequences and matching against the index. The approximate match score is derived from a CRF, trained on parallel transcripts, which provides a general framework for modeling the errors that a recognition system may make taking contextual effects into consideration. Our approach differs from other work in the field in that we focus on using CRFs to model context dependent phone level confusions, rather than on explicitly modeling parameters of an edit distance. While, the results we obtain on both in and out of vocabulary (OOV) search tasks improve on previous work which incorporated high order phone confusions, the gains for OOV are more impressive.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Improved vocabulary independent search with approximate match based on Conditional Random Fields

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Matching Criteria for Vocabulary-Independent Search
Upendra V Chaudhari ... Michael Picheny
IEEE Transactions on Audio, Speech, and Language Processing | VOL. 20
Upendra V Chaudhari, et. al.Upendra V Chaudhari ... Michael Picheny
01 Jul 2012
IEEE Transactions on Audio, Speech, and Language Processing | VOL. 20

An approach for efficient open vocabulary spoken term detection
Atta Norouzian ... Richard Rose
Speech Communication | VOL. 57
Atta Norouzian, et. al.Atta Norouzian ... Richard Rose
24 Sep 2013
Speech Communication | VOL. 57

Open vocabulary spoken document retrieval by subword sequence obtained from speech recognizer
Go Kuriki ... Yoshiaki
-
Go Kuriki, et. al. Go Kuriki ... Yoshiaki
01 Dec 2008
01 Dec 2008

Biomedical Named Entity Recognition with less Supervision
Omid Ghiasvand ... Rohit J Kate
-
Omid Ghiasvand, et. al.Omid Ghiasvand ... Rohit J Kate
01 Oct 2015
01 Oct 2015

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Improved vocabulary independent search with approximate match based on Conditional Random Fields

Abstract

Talk to us

Similar Papers