Classification of animal sounds in a hyperdiverse rainforest using convolutional neural networks with data augmentation

Yuren Sun,Tatiana Midori Maeda,Claudia Solís-Lemus,Daniel Pimentel-Alarcón,Zuzana Buřivalová

doi:10.1016/j.ecolind.2022.109621

Yuren Sun, Tatiana Midori Maeda + Show 3 more

Open Access

https://doi.org/10.1016/j.ecolind.2022.109621

Copy DOI

Journal: Ecological Indicators	Publication Date: Nov 3, 2022
Citations: 12	License type: cc-by-nc-nd

Affiliation: University of Wisconsin–Madison

Abstract

To protect tropical forest biodiversity, we need to be able to detect it reliably, cheaply, and at scale. Automated detection of sound producing animals from passively recorded soundscapes via machine-learning approaches is a promising technique towards this goal, but it is constrained by the necessity of large training data sets. Using soundscapes from a tropical forest in Borneo and a Convolutional Neural Network model (CNN), we investigate i) the minimum viable training data set size for accurate prediction of call types (‘sonotypes’), and ii) the extent to which data augmentation and transfer learning can overcome the issue of small and imbalanced training data sets. We found that even relatively high sample sizes (>80 per sonotype) lead to mediocre accuracy, which however improved significantly with data augmentation and transfer learning, including at extremely small sample sizes (3 per sonotype), regardless of taxonomic group or call characteristics. Neither transfer learning nor data augmentation alone achieved high accuracy. Our results suggest that transfer learning and data augmentation could make the use of CNNs to classify species’ vocalizations feasible even for small soundscape-based projects with many rare species. Retraining our open-source model requires only basic programming skills which makes it possible for individual conservation initiatives to match their local context, in order to enable more evidence-informed management of biodiversity.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Classification of animal sounds in a hyperdiverse rainforest using convolutional neural networks with data augmentation

Abstract

Talk to us

Similar Papers

More From: Ecological Indicators

Lead the way for us

Similar Papers

Combining weakly and strongly supervised learning improves strong supervision in Gleason pattern classification
Sebastian Otálora ... Henning Müller
BMC Medical Imaging | VOL. 21
Sebastian Otálora, et. al.Sebastian Otálora ... Henning Müller
08 May 2021
BMC Medical Imaging | VOL. 21

Diagnostic performance evaluation of adult Chiari malformation type I based on convolutional neural networks
Wei-Wei Lin ... Yong-Jian Zhu
European Journal of Radiology | VOL. 151
Wei-Wei Lin, et. al.Wei-Wei Lin ... Yong-Jian Zhu
02 Apr 2022
European Journal of Radiology | VOL. 151

Molecular image-convolutional neural network (CNN) assisted QSAR models for predicting contaminant reactivity toward OH radicals: Transfer learning, data augmentation and model interpretation
Shifa Zhong ... Huichun Zhang
Chemical Engineering Journal | VOL. 408
Shifa Zhong, et. al.Shifa Zhong ... Huichun Zhang
07 Dec 2020
Chemical Engineering Journal | VOL. 408

A novel polynomial reconstruction algorithm‐based 1D convolutional neural network used for transfer learning in Raman spectroscopy application
Lin‐Wei Shang ... Jian‐Hua Yin
Journal of Raman Spectroscopy | VOL. 53
Lin‐Wei Shang, et. al.Lin‐Wei Shang ... Jian‐Hua Yin
24 Oct 2021
Journal of Raman Spectroscopy | VOL. 53

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Classification of animal sounds in a hyperdiverse rainforest using convolutional neural networks with data augmentation

Abstract

Talk to us

Similar Papers

More From: Ecological Indicators