Emotion Recognition from Audio and Visual Data using F-score based Fusion

Abhishek Gera,Arnab Bhattacharya

doi:10.1145/2567688.2567690

Abstract

Emotion recognition has been one of the cornerstones of human-computer interaction. Although decades of work has attacked the problem of automatic emotion recognition from either audio or video signals, the fusion of the two modalities is more recent. In this paper, we aim to tackle the problem when both audio and video data are available in a synchronized manner. We address the six basic human emotions, namely, anger, disgust, fear, happiness, sadness, and surprise. We employ an automatic face tracker to extract the different facial points of interest from a video. We then compute feature vectors for each video frame using distances and angles between the tracked points. For audio data, we use the pitch, energy and MFCC to derive feature vectors for each window as well as the entire audio signal. We use two standard techniques, GMM-based HMM and SVM, as the base classifiers. We then design a novel fusion method using the F-score of the base classifiers. We first demonstrate that our fusion approach can increase the accuracy of the base classifiers by as much as 5%. Finally, we show that our fusion-based bi-modal emotion recognition method achieves an overall accuracy of 54% on a publicly available database, which is an improvement upon the current state-of-the-art by 9%.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Emotion Recognition from Audio and Visual Data using F-score based Fusion

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Cross-Modal Harmony: Low-Rank Fusion for Enhanced Artificial Emotion Recognition
...
INTERANTIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING AND MANAGEMENT | VOL. 08
, et. al. ...
16 Oct 2024
INTERANTIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING AND MANAGEMENT | VOL. 08

Audio-visual emotion fusion (AVEF): A deep efficient weighted approach
Yaxiong Ma ... Andrej Košir
Information Fusion | VOL. 46
Yaxiong Ma, et. al.Yaxiong Ma ... Andrej Košir
15 Jun 2018
Information Fusion | VOL. 46

Theory of mind, emotion recognition, delusions and the quality of the therapeutic relationship in patients with psychosis \u2013 a secondary analysis of a randomized-controlled therapy trial
Stephanie Mehl ... Georg Wiedemann
BMC psychiatry | VOL. 20
Stephanie Mehl, et. al.Stephanie Mehl ... Georg Wiedemann
10 Feb 2020
BMC psychiatry | VOL. 20

The Audio-Visual Arabic Dataset for Natural Emotions
Ftoon Abu Shaqra ... Mahmoud Al-Ayyoub
-
Ftoon Abu Shaqra, et. al.Ftoon Abu Shaqra ... Mahmoud Al-Ayyoub
01 Aug 2019
01 Aug 2019

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Emotion Recognition from Audio and Visual Data using F-score based Fusion

Abstract

Talk to us

Similar Papers