Utterance Weighted Multi-Dilation Temporal Convolutional Networks for Monaural Speech Dereverberation

William Ravenscroft,Stefan Goetze,Thomas Hain

doi:10.1109/iwaenc53105.2022.9914752

Abstract

Speech dereverberation is an important stage in many speech technology applications. Recent work in this area has been dominated by deep neural network models. Temporal convolutional networks (TCNs) are deep learning models that have been proposed for sequence modelling in the task of dereverberating speech. In this work a weighted multi-dilation depthwise-separable convolution is proposed to replace standard depthwise-separable convolutions in TCN models. This proposed convolution enables the TCN to dynamically focus on more or less local information in its receptive field at each convolutional block in the network. It is shown that this weighted multi-dilation temporal convolutional network (WD-TCN) consistently outperforms the TCN across various model configurations and using the WD-TCN model is a more parameter-efficient method to improve the performance of the model than increasing the number of convolutional blocks. The best performance improvement over the baseline TCN is 0.55 dB scale-invariant signal-to-distortion ratio (SISDR) and the best performing WD-TCN model attains 12.26 dB SISDR on the WHAMR dataset.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Utterance Weighted Multi-Dilation Temporal Convolutional Networks for Monaural Speech Dereverberation

Abstract

Talk to us

Similar Papers

Lead the way for us

Publication Date: Sep 5, 2022
Citations: 4	License type: cc-by

Similar Papers

Receptive Field Analysis of Temporal Convolutional Networks for Monaural Speech Dereverberation
William Ravenscroft ... Thomas Hain
-
William Ravenscroft, et. al.William Ravenscroft ... Thomas Hain
29 Aug 2022
29 Aug 2022

Ensemble deep learning models for protein secondary structure prediction using bidirectional temporal convolution and bidirectional long short-term memory.
Lu Yuan ... Yihui Liu
Frontiers in Bioengineering and Biotechnology | VOL. 11
Lu Yuan, et. al.Lu Yuan ... Yihui Liu
13 Feb 2023
Frontiers in Bioengineering and Biotechnology | VOL. 11

Deformable Temporal Convolutional Networks for Monaural Noisy Reverberant Speech Separation
William Ravenscroft ... Thomas Hain
-
William Ravenscroft, et. al.William Ravenscroft ... Thomas Hain
04 Jun 2023
04 Jun 2023

Hybrid model for short-term wind power forecasting based on singular spectrum analysis and a temporal convolutional attention network with an adaptive receptive field
Zhen Shao ... Shanlin Yang
Energy Conversion and Management | VOL. 269
Zhen Shao, et. al.Zhen Shao ... Shanlin Yang
01 Sep 2022
Energy Conversion and Management | VOL. 269

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Utterance Weighted Multi-Dilation Temporal Convolutional Networks for Monaural Speech Dereverberation

Abstract

Talk to us

Similar Papers