Monaural Speech Enhancement Using Intra-Spectral Recurrent Layers in the Magnitude and Phase Responses

Khandokar Md Nayem,Donald S Williamson

doi:10.1109/icassp40776.2020.9054298

Abstract

Speech enhancement has greatly benefited from deep learning. Currently, the best performing deep architectures use long short-term memory (LSTM) recurrent neural networks (RNNs) to model short and long temporal dependencies. These approaches, however, underutilize or ignore spectral-level dependencies within the magnitude and phase responses, respectively. In this paper, we propose a deep learning architecture that leverages both temporal and spectral dependencies within the magnitude and phase responses. More specifically, we first train a LSTM network to predict both the spectral-magnitude response and group delay, where this model captures temporal correlations. We then introduce Markovian recurrent connections in the output layers to capture spectral dependencies within the magnitude and phase responses. We compare our approach with traditional enhancement approaches and approaches that consider spectral dependencies within a single time frame. The results show that considering the within-frame spectral dependencies leads to improvements.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Monaural Speech Enhancement Using Intra-Spectral Recurrent Layers in the Magnitude and Phase Responses

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Incorporating Intra-Spectral Dependencies with a Recurrent Output Layer for Improved Speech Enhancement
Khandokar Md. Nayem ... Donald S. Williamson
-
Khandokar Md. Nayem, et. al.Khandokar Md. Nayem ... Donald S. Williamson
01 Oct 2019
01 Oct 2019

Deep Learning with Long Short-Term Memory Recurrent Neural Network for Daily Container Volumes of Storage Yard Predictions in Port
Yinping Gao ... Daofang Chang
-
Yinping Gao, et. al.Yinping Gao ... Daofang Chang
01 Oct 2018
01 Oct 2018

Kazakh and Russian Languages Identification Using Long Short-Term Memory Recurrent Neural Networks
Zhanibek Kozhirbayev ... Zhandos Yessenbayev
-
Zhanibek Kozhirbayev, et. al.Zhanibek Kozhirbayev ... Zhandos Yessenbayev
01 Sep 2017
01 Sep 2017

Interpretable, highly accurate brain decoding of subtly distinct brain states from functional MRI using intrinsic functional networks and long short-term memory recurrent neural networks
Hongming Li ... Yong Fan
NeuroImage | VOL. 202
Hongming Li, et. al.Hongming Li ... Yong Fan
27 Jul 2019
NeuroImage | VOL. 202

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Monaural Speech Enhancement Using Intra-Spectral Recurrent Layers in the Magnitude and Phase Responses

Abstract

Talk to us

Similar Papers