Large-Scale Weakly Supervised Audio Classification Using Gated Convolutional Neural Network

Yong Xu,Wenwu Wang,Qiuqiang Kong,Mark D Plumbley

doi:10.1109/icassp.2018.8461975

Abstract

In this paper, we present a gated convolutional neural network and a temporal attention-based localization method for audio classification, which won the 1st place in the large-scale weakly supervised sound event detection task of Detection and Classification of Acoustic Scenes and Events (DCASE) 2017 challenge. The audio clips in this task, which are extracted from YouTube videos, are manually labelled with one or more audio tags, but without time stamps of the audio events, hence referred to as weakly labelled data. Two subtasks are defined in this challenge including audio tagging and sound event detection using this weakly labelled data. We propose a convolutional recurrent neural network (CRNN) with learnable gated linear units (GLUs) non-linearity applied on the log Mel spectrogram. In addition, we propose a temporal attention method along the frames to predict the locations of each audio event in a chunk from the weakly labelled data. The performances of our systems were ranked the 1st and the 2nd as a team in these two sub-tasks of DCASE 2017 challenge with F value 55.6% and Equal error 0.73, respectively.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Large-Scale Weakly Supervised Audio Classification Using Gated Convolutional Neural Network

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Sound Event Detection of Weakly Labelled Data With CNN-Transformer and Automatic Threshold Optimization
Qiuqiang Kong ... Yong Xu
IEEE/ACM Transactions on Audio, Speech, and Language Processing | VOL. 28
Qiuqiang Kong, et. al.Qiuqiang Kong ... Yong Xu
01 Jan 2020
IEEE/ACM Transactions on Audio, Speech, and Language Processing | VOL. 28

A Region Based Attention Method for Weakly Supervised Sound Event Detection and Classification
Jie Yan ... Li-Rong Dai
-
Jie Yan, et. al.Jie Yan ... Li-Rong Dai
01 May 2019
01 May 2019

Adaptive Memory-Controlled Self-Attention for Polyphonic Sound Event Detection
Mei Wang ... Hongbin Qiu
Symmetry | VOL. 14
Mei Wang, et. al.Mei Wang ... Hongbin Qiu
12 Feb 2022
Symmetry | VOL. 14

Research on Semi-Supervised Sound Event Detection Based on Mean Teacher Models Using ML-LoBCoD-NET
Jinjia Wang ... Jing Xia
IEEE Access | VOL. 8
Jinjia Wang, et. al.Jinjia Wang ... Jing Xia
01 Jan 2020
IEEE Access | VOL. 8

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Large-Scale Weakly Supervised Audio Classification Using Gated Convolutional Neural Network

Abstract

Talk to us

Similar Papers