Automatic Speech Recognition of English-isiZulu Code-switched Speech from South African Soap Operas

Ewald Van Der Westhuizen,Thomas Niesler

doi:10.1016/j.procs.2016.04.039

Ewald Van Der Westhuizen, Thomas Niesler

Open Access

https://doi.org/10.1016/j.procs.2016.04.039

Copy DOI

Journal: Procedia Computer Science	Publication Date: Jan 1, 2016
Citations: 17	License type: cc-by-nc-nd

Affiliation: Stellenbosch University

Abstract

We introduce a new English-isiZulu code-switched speech corpus compiled from South African soap opera broadcasts. isiZulu itself is currently under-resourced, and automatic speech recognition is made even more challenging by the high prevalence of code-switching in spontaneous speech. Analysis of the corpus reflects effects common in conversational isiZulu, such as vowel deletion and cross-language prefixes and suffixes. Baseline monolingual and code-switched automatic speech recognition systems are developed, including a new language model configuration that explicitly includes switching transitions. For code-switched speech, a system with language-dependent acoustic models and language-dependent language models linked by switching transitions leads to best performance, although word error rates overall remain very high.

Full Text