Self-Supervised Speech Enhancement for Arabic Speech Recognition in Real-World Environments

Bilal Dendani,Halima Bahi,Toufik Sari

doi:10.18280/ts.380212

Abstract

Mobile speech recognition attracts much attention in the ubiquitous context, however, background noises, speech coding, and transmission errors are prone to corrupt the incoming speech. Therein, building a robust speech recognizer requires the availability of a large number of real-world speech samples. Arabic language, like many other languages, lacks such resources; to overcome this limitation, we propose a speech enhancement step, before the recognition begins. For the speech enhancement purpose, we suggest the use of a deep autoencoder (DAE) algorithm. A two-step procedure is suggested: in the first step, an overcomplete DAE is trained in an unsupervised way, and in the second one, a denoising DAE is trained in a supervised way leveraging the clean speech produced in the previous step. Experimental results performed on a real-life mobile database confirmed the potentials of the proposed approach and show a reduction of the WER (Word Error Rate) of a ubiquitous Arabic speech recognizer. Further experiments show an improvement of the perceptual evaluation of speech quality (PESQ), and the short-time objective intelligibility (STOI) as well.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Self-Supervised Speech Enhancement for Arabic Speech Recognition in Real-World Environments

Abstract

Talk to us

Similar Papers

More From: Traitement du Signal

Lead the way for us

Journal: Traitement du Signal	Publication Date: Apr 30, 2021
Citations: 5

Similar Papers

New research on monaural speech segregation based on quality assessment
Xiaoping Xie ... Fei Ding
Computer Speech & Language | VOL. 85
Xiaoping Xie, et. al.Xiaoping Xie ... Fei Ding
05 Dec 2023
Computer Speech & Language | VOL. 85

CITISEN: A Deep Learning-Based Speech Signal-Processing Mobile Application
Yu-Wen Chen ... Kai-Chun Liu
IEEE Access | VOL. 10
Yu-Wen Chen, et. al.Yu-Wen Chen ... Kai-Chun Liu
01 Jan 2021
IEEE Access | VOL. 10

Deep causal speech enhancement and recognition using efficient long-short term memory Recurrent Neural Network.
Zhenqing Li ... Amil Daraz
PloS one | VOL. 19
Zhenqing Li, et. al.Zhenqing Li ... Amil Daraz
03 Jan 2024
PloS one | VOL. 19

A Conditional Generative Model for Speech Enhancement
Zeng-Xi Li ... Li-Rong Dai
Circuits, Systems, and Signal Processing | VOL. 37
Zeng-Xi Li, et. al.Zeng-Xi Li ... Li-Rong Dai
13 Mar 2018
Circuits, Systems, and Signal Processing | VOL. 37

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Self-Supervised Speech Enhancement for Arabic Speech Recognition in Real-World Environments

Abstract

Talk to us

Similar Papers

More From: Traitement du Signal