Spoken Language Identification with Deep Convolutional Neural Network and Data Augmentation

Can Korkut,Levent M Arslan,Ali Haznedaroğlu

doi:10.1109/siu49456.2020.9302425

Spoken Language Identification with Deep Convolutional Neural Network and Data Augmentation

Can Korkut, Levent M Arslan + Show 1 more

https://doi.org/10.1109/siu49456.2020.9302425

Copy DOI

Publication Date: Oct 5, 2020

Affiliation: Boğaziçi University

#Convolutional Network #Deep Convolutional Network + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

In this paper, a spoken language detection system based on deep convolutional neural networks is presented. The neural network model is trained and tested on a speech dataset containing five languages. Speech signals are first converted into mel-spectrogram features and these features are fed into the deep convolutional neural network. Flattened outputs of the deep convolutional network are then fed into a recurrent layer, and a dense layer with softmax activation function is used as an output layer to predict the output language probabilities. This network results in 0.89 Fl-score in our test data. We also used a data augmentation method, namely SpecAugment, which increased the Fl-score to 0.94.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.