Automatic detection and classification of marmoset vocalizations using deep and recurrent neural networks.

Ya-Jie Zhang,Jun-Feng Huang,Zhen-Hua Ling,Yu Hu,Neng Gong

doi:10.1121/1.5047743

Abstract

This paper investigates the methods to detect and classify marmoset vocalizations automatically using a large data set of marmoset vocalizations and deep learning techniques. For vocalization detection, neural networks-based methods, including deep neural network (DNN) and recurrent neural network with long short-term memory units, are designed and compared against a conventional rule-based detection method. For vocalization classification, three different classification algorithms are compared, including a support vector machine (SVM), DNN, and long short-term memory recurrent neural networks (LSTM-RNNs). A 1500-min audio data set containing recordings from four pairs of marmoset twins and manual annotations is employed for experiments. Two test sets are built according to whether the test samples are produced by the marmosets in the training set (test set I) or not (test set II). Experimental results show that the LSTM-RNN-based detection method outperformed others and achieved 0.92% and 1.67% frame error rate on these two test sets. Furthermore, the deep learning models obtained higher classification accuracy than the SVM model, which was 95.60% and 91.67% on the two test sets, respectively.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Automatic detection and classification of marmoset vocalizations using deep and recurrent neural networks.

Abstract

Talk to us

Similar Papers

More From: The Journal of the Acoustical Society of America

Lead the way for us

Journal: The Journal of the Acoustical Society of America	Publication Date: Jul 1, 2018
Citations: 17

Similar Papers

Long short-term memory recurrent neural network architectures for large scale acoustic modeling
Haşim Sak ... Françoise Beaufays
-
Haşim Sak, et. al.Haşim Sak ... Françoise Beaufays
14 Sep 2014
14 Sep 2014

Semi-Supervised Training in Deep Learning Acoustic Model
Yan Huang ... Yifan Gong
-
Yan Huang, et. al.Yan Huang ... Yifan Gong
08 Sep 2016
08 Sep 2016

Long short-term memory (LSTM) recurrent neural network for low-flow hydrological time series forecasting
Bibhuti Bhusan Sahoo ... Ramakar Jha
Acta Geophysica | VOL. 67
Bibhuti Bhusan Sahoo, et. al.Bibhuti Bhusan Sahoo ... Ramakar Jha
20 Jul 2019
Acta Geophysica | VOL. 67

Advanced Recurrent Neural Networks for Automatic Speech Recognition
Yu Zhang ... Guoguo Chen
-
Yu Zhang, et. al.Yu Zhang ... Guoguo Chen
01 Jan 2017
01 Jan 2017

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Automatic detection and classification of marmoset vocalizations using deep and recurrent neural networks.

Abstract

Talk to us

Similar Papers

More From: The Journal of the Acoustical Society of America