Mode Variational LSTM Robust to Unseen Modes of Variation: Application to Facial Expression Recognition

Wissam J Baddar,Yong Man Ro

doi:10.1609/aaai.v33i01.33013215

Abstract

Spatio-temporal feature encoding is essential for encoding the dynamics in video sequences. Recurrent neural networks, particularly long short-term memory (LSTM) units, have been popular as an efficient tool for encoding spatio-temporal features in sequences. In this work, we investigate the effect of mode variations on the encoded spatio-temporal features using LSTMs. We show that the LSTM retains information related to the mode variation in the sequence, which is irrelevant to the task at hand (e.g. classification facial expressions). Actually, the LSTM forget mechanism is not robust enough to mode variations and preserves information that could negatively affect the encoded spatio-temporal features. We propose the mode variational LSTM to encode spatio-temporal features robust to unseen modes of variation. The mode variational LSTM modifies the original LSTM structure by adding an additional cell state that focuses on encoding the mode variation in the input sequence. To efficiently regulate what features should be stored in the additional cell state, additional gating functionality is also introduced. The effectiveness of the proposed mode variational LSTM is verified using the facial expression recognition task. Comparative experiments on publicly available datasets verified that the proposed mode variational LSTM outperforms existing methods. Moreover, a new dynamic facial expression dataset with different modes of variation, including various modes like pose and illumination variations, was collected to comprehensively evaluate the proposed mode variational LSTM. Experimental results verified that the proposed mode variational LSTM encodes spatio-temporal features robust to unseen modes of variation.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Mode Variational LSTM Robust to Unseen Modes of Variation: Application to Facial Expression Recognition

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence

Lead the way for us

Journal: Proceedings of the AAAI Conference on Artificial Intelligence	Publication Date: Jul 17, 2019
Citations: 23

Similar Papers

Encoding features robust to unseen modes of variation with attentive long short-term memory
Wissam J Baddar ... Yong Man Ro
Pattern Recognition | VOL. 100
Wissam J Baddar, et. al.Wissam J Baddar ... Yong Man Ro
18 Dec 2019
Pattern Recognition | VOL. 100

Facial Expression Detection and Recognition through VIOLA-JONES Algorithm and HCNN using LSTM Method
Dinesh Kumar P ... Dr B Rosiline Jeetha
International Journal of Scientific Research in Computer Science, Engineering and Information Technology | VOL. -
Dinesh Kumar P, et. al.Dinesh Kumar P ... Dr B Rosiline Jeetha
12 Jun 2021
International Journal of Scientific Research in Computer Science, Engineering and Information Technology | VOL. -

A MULTIMODAL APPROACH FOR AUTOMATIC DETECTION OF INFANT PAIN USING FACIAL EXPRESSION AND CRYING
...
Journal of critical reviews | VOL. 7
, et. al. ...
01 Apr 2020
Journal of critical reviews | VOL. 7

Deepfake Face Detection Using Machine Learning with LSTM
Pasam Samkeerthana
INTERANTIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING AND MANAGEMENT | VOL. 08
Pasam SamkeerthanaPasam Samkeerthana
28 Apr 2024
INTERANTIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING AND MANAGEMENT | VOL. 08

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Mode Variational LSTM Robust to Unseen Modes of Variation: Application to Facial Expression Recognition

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence