Speech Recognition Engineering Issues in Speech to Speech Translation System Design for Low Resource Languages and Domains

S Narayanan,M Graciarena,M Bulut,A Sethy,S Sundaram,P.G Georgiou,H Franco,Wen Wang Wen Wang,D Vergyri,Jing Zheng Jing Zheng,V Abrash,C Richey,R.R Gadde,K Precoda,E Ettelaie,Dagen Wang Dagen Wang,S Ananthakrishnan,M Frandsen

doi:10.1109/icassp.2006.1661499

S Narayanan, M Graciarena + Show 16 more

https://doi.org/10.1109/icassp.2006.1661499

Copy DOI

Export

Save

Cite

Publication Date: May 14, 2006

Citations: 12

Affiliation: Viterbo University

Abstract
Full-Text
Similar Papers

Abstract

Listen

Engineering automatic speech recognition (ASR) for speech to speech (S2S) translation systems, especially targeting languages and domains that do not have readily available spoken language resources, is immensely challenging due to a number of reasons. In addition to contending with the conventional data-hungry speech acoustic and language modeling needs, these designs have to accommodate varying requirements imposed by the domain needs and characteristics, target device and usage modality (such as phrase-based, or spontaneous free form interactions, with or without visual feedback) and huge spoken language variability arising due to socio-linguistic and cultural differences of the users. This paper, using case studies of creating speech translation systems between English and languages such as Pashto and Farsi, describes some of the practical issues and the solutions that were developed for multilingual ASR development. These include novel acoustic and language modeling strategies such as language adaptive recognition, active-learning based language modeling, class-based language models that can better exploit resource poor language data, efficient search strategies, including N-best and confidence generation to aid multiple hypotheses translation, use of dialog information and clever interface choices to facilitate ASR, and audio interface design for meeting both usability and robustness requirements.

Full Text