Robust Raw Waveform Speech Recognition Using Relevance Weighted Representations

Purvi Agrawal,Sriram Ganapathy

doi:10.21437/interspeech.2020-2301

Abstract

Speech recognition in noisy and channel distorted scenarios is often challenging as the current acoustic modeling schemes are not adaptive to the changes in the signal distribution in the presence of noise. In this work, we develop a novel acoustic modeling framework for noise robust speech recognition based on relevance weighting mechanism. The relevance weighting is achieved using a sub-network approach that performs feature selection. A relevance sub-network is applied on the output of first layer of a convolutional network model operating on raw speech signals while a second relevance sub-network is applied on the second convolutional layer output. The relevance weights for the first layer correspond to an acoustic filterbank selection while the relevance weights in the second layer perform modulation filter selection. The model is trained for a speech recognition task on noisy and reverberant speech. The speech recognition experiments on multiple datasets (Aurora-4, CHiME-3, VOiCES) reveal that the incorporation of relevance weighting in the neural network architecture improves the speech recognition word error rates significantly (average relative improvements of 10% over the baseline systems)

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Robust Raw Waveform Speech Recognition Using Relevance Weighted Representations

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Interpretable Representation Learning for Speech and Audio Signals Based on Relevance Weighting
Purvi Agrawal ... Sriram Ganapathy
IEEE/ACM Transactions on Audio, Speech, and Language Processing | VOL. 28
Purvi Agrawal, et. al.Purvi Agrawal ... Sriram Ganapathy
01 Jan 2020
IEEE/ACM Transactions on Audio, Speech, and Language Processing | VOL. 28

Representation Learning for Speech Recognition Using Feedback Based Relevance Weighting
Purvi Agrawal ... Sriram Ganapathy
-
Purvi Agrawal, et. al.Purvi Agrawal ... Sriram Ganapathy
06 Jun 2021
06 Jun 2021

A Joint Training Framework for Robust Automatic Speech Recognition
Zhong-Qiu Wang ... Deliang Wang
IEEE/ACM Transactions on Audio, Speech, and Language Processing | VOL. 24
Zhong-Qiu Wang, et. al.Zhong-Qiu Wang ... Deliang Wang
01 Apr 2016
IEEE/ACM Transactions on Audio, Speech, and Language Processing | VOL. 24

A hybrid CTC+Attention model based on end-to-end framework for multilingual speech recognition
Sendong Liang ... Wei Qi Yan
Multimedia Tools and Applications | VOL. 81
Sendong Liang, et. al.Sendong Liang ... Wei Qi Yan
20 May 2022
Multimedia Tools and Applications | VOL. 81

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Robust Raw Waveform Speech Recognition Using Relevance Weighted Representations

Abstract

Talk to us

Similar Papers