Audio-visual fuzzy fusion for robust speech recognition

M Malcangi,P Patel,K Ouazzane

doi:10.1109/ijcnn.2013.6706789

Abstract

Improvements of robustness of speech recognition is one of the hottest topics in speech signal processing, particularly when applied within a noisy environment. Most of the research efforts focused in combining audio and visual data to implement an audiovisual speech recognition (AVSR) system. Bimodal approach demonstrated that a superior performance can be gained compared to the separate audio or visual approach. This paper proposes a fuzzy logic-based data fusion method that combines the recognition capabilities of two independent working systems namely the automatic speech recognition system (ASR) and the automatic visual recognition system (AVR). The main purpose is to boost the whole system's performance keeping the ASR separate from the AVR. This approach provides a powerful method that enables simpler data fusion at decision level rather than the more complex at data and features level. Such complexity is also lowered due to the fuzzy logic-based implementation of the data fusion engine. Preliminary experimental results confirms the proposed approach.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Audio-visual fuzzy fusion for robust speech recognition

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Measuring the effect of high-speed video data on the audio-visual speech recognition accuracy
D V Ivanko ... D A Ryumin
Information and Control Systems | VOL. -
D V Ivanko, et. al.D V Ivanko ... D A Ryumin
19 Apr 2019
Information and Control Systems | VOL. -

Combined speech enhancement and auditory modelling for robust distributed speech recognition
Ronan Flynn ... Edward Jones
Speech Communication | VOL. 50
Ronan Flynn, et. al.Ronan Flynn ... Edward Jones
20 May 2008
Speech Communication | VOL. 50

Speaker independent audio-visual continuous speech recognition
Luhong Liang ... A.V Nefian
-
Luhong Liang, et. al. Luhong Liang ... A.V Nefian
07 Nov 2002
07 Nov 2002

Audio-visual speech recognition based on machine learning approach
Saswati Debnath ... Pinki Roy
International Journal of Advanced Intelligence Paradigms | VOL. 21
Saswati Debnath, et. al.Saswati Debnath ... Pinki Roy
01 Jan 2021
International Journal of Advanced Intelligence Paradigms | VOL. 21

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Audio-visual fuzzy fusion for robust speech recognition

Abstract

Talk to us

Similar Papers