Turkish broadcast news transcription revisited

Ebru Arisoy,Murat Saraglar

doi:10.1109/siu.2018.8404597

Abstract

In this study a decade old automatic speech recognition system for Turkish broadcast news transcription is revisited and updated with the latest methods. Recently deep learning using artificial neural networks resulted in significant improvements in speech recognition error rates and became the state-of-the-art. Neural network based acoustic and language models are used as the main components of the speech recognition system built in this paper. For acoustic modeling, deep neural networks are optimized using both cross-entropy and sequence discriminative objective functions. In addition, time-delay neural networks are used for modeling long term dependencies with similar performance to recurrent neural networks. The lowest error rates are obtained using discriminatively trained versions of these models. For the language model a recurrent language model is used. It was observed that the word error rates are approximately halved and fell below 10%.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Turkish broadcast news transcription revisited

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

An Investigation of Multilingual TDNN-BLSTM Acoustic Modeling for Hindi Speech Recognition
Ankit Kumar ... Rajesh Kumar Aggarwal
International Journal of Sensors, Wireless Communications and Control | VOL. 12
Ankit Kumar, et. al.Ankit Kumar ... Rajesh Kumar Aggarwal
01 Jan 2021
International Journal of Sensors, Wireless Communications and Control | VOL. 12

Improving Russian LVCSR Using Deep Neural Networks for Acoustic and Language Modeling
Irina Kipyatkova
-
Irina KipyatkovaIrina Kipyatkova
01 Jan 2018
01 Jan 2018

Building Acoustic and Language Model for Continuous Speech Recognition in Bahasa Indonesia
Andreas Widjaja ... Vincent Elbert Budiman
Jurnal Teknik Informatika dan Sistem Informasi | VOL. 6
Andreas Widjaja, et. al.Andreas Widjaja ... Vincent Elbert Budiman
10 Aug 2020
Jurnal Teknik Informatika dan Sistem Informasi | VOL. 6

Exploring recurrent neural network based acoustic and linguistic modeling for children's speech recognition
Sreeram Ganji ... Rohit Sinha
-
Sreeram Ganji, et. al.Sreeram Ganji ... Rohit Sinha
01 Nov 2017
01 Nov 2017

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Turkish broadcast news transcription revisited

Abstract

Talk to us

Similar Papers