Simple Effective Methods for Decision-Level Fusion in Two-Stream Convolutional Neural Networks for Video Classification

Rukiye Savran Kızıltepe,John Q Gan

doi:10.1007/978-3-030-62362-3_8

Abstract

AbstractConvolutional Neural Networks (CNNs) have recently been applied for video classification applications where various methods for combining the appearance (spatial) and motion (temporal) information from video clips are considered. The most common method for combining the spatial and temporal information for video classification is averaging prediction scores at softmax layer. Inspired by the Mycin uncertainty system for combining production rules in expert systems, this paper proposes using the Mycin formula for decision fusion in two-stream convolutional neural networks. Based on the intuition that spatial information is more useful than temporal information for video classification, this paper also proposes multiplication and asymmetrical multiplication for decision fusion, aiming to better combine the spatial and temporal information for video classification using two-stream convolutional neural networks. The experimental results show that (i) both spatial and temporal information are important, but the decision from the spatial stream should be dominating with the decision from temporal stream as complementary and (ii) the proposed asymmetrical multiplication method for decision fusion significantly outperforms the Mycin method and average method as well.KeywordsDeep learningVideo classificationAction recognitionConvolutional neural networksDecision fusion

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Simple Effective Methods for Decision-Level Fusion in Two-Stream Convolutional Neural Networks for Video Classification

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Spatial and temporal saliency based four-stream network with multi-task learning for action recognition
Ming Zong ... Wanting Ji
Applied Soft Computing | VOL. 132
Ming Zong, et. al.Ming Zong ... Wanting Ji
05 Dec 2022
Applied Soft Computing | VOL. 132

A Refined Non-Driving Activity Classification Using a Two-Stream Convolutional Neural Network
Lichao Yang ... Lee Skrypchuk
IEEE Sensors Journal | VOL. 21
Lichao Yang, et. al.Lichao Yang ... Lee Skrypchuk
15 Jul 2021
IEEE Sensors Journal | VOL. 21

Spatial-temporal interaction learning based two-stream network for action recognition
Tianyu Liu ... Ping Jiang
Information Sciences | VOL. 606
Tianyu Liu, et. al.Tianyu Liu ... Ping Jiang
28 May 2022
Information Sciences | VOL. 606

The Interaction Between Temporal and Spatial Information in the Updating of Situation Model
Xianyou He ... Huijuan Li
Acta Psychologica Sinica | VOL. 45
Xianyou He, et. al.Xianyou He ... Huijuan Li
27 Nov 2013
Acta Psychologica Sinica | VOL. 45

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Simple Effective Methods for Decision-Level Fusion in Two-Stream Convolutional Neural Networks for Video Classification

Abstract

Talk to us

Similar Papers