G-MS2F: GoogLeNet based multi-stage feature fusion of deep CNN for scene recognition

Pengjie Tang,Hanli Wang,Sam Kwong

doi:10.1016/j.neucom.2016.11.023

Abstract

Scene recognition plays an important role in the task of visual information retrieval, segmentation and image/video understanding. Traditional approaches for scene recognition usually utilize handcrafted features and have the drawbacks of poor representation ability, which can be improved by employing deep convolutional neural network (CNN) features that contain more semantic and structure information and thus possess more discriminative ability via multiple linear and non-linear transformations. However, an amount of detailed information may be lost when only the final output features which have gone through a certain number of transformations are applied to scene recognition. The features which are generated from the intermediate layers are not fully utilized. In this work, the GoogLeNet model is employed and divided into three parts of layers from bottom to top. The output features from each of the three parts are applied for scene recognition, which leads to the proposed GoogLeNet based multi-stage feature fusion (G-MS2F). What's more, the product rule is used to generate the final decision for scene recognition from the three outputs corresponding to the three parts of the proposed model. The experimental results demonstrate that the proposed model is superior to a number of state-of-the-art CNN models for scene recognition, and obtains the recognition accuracy of 92.90%, 79.63% and 64.06% on the benchmark scene recognition datasets Scene15, MIT67 and SUN397, respectively.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

G-MS2F: GoogLeNet based multi-stage feature fusion of deep CNN for scene recognition

Abstract

Talk to us

Similar Papers

More From: Neurocomputing

Lead the way for us

Journal: Neurocomputing	Publication Date: Nov 19, 2016
Citations: 193

Similar Papers

Enhancement of ELDA Tracker Based on CNN Features and Adaptive Model Update.
Changxin Gao ... Jin-Gang Yu
Sensors | VOL. 16
Changxin Gao, et. al.Changxin Gao ... Jin-Gang Yu
15 Apr 2016
Sensors | VOL. 16

Transferred Deep Convolutional Neural Network Features for Extensive Facial Landmark Localization
Shaohua Zhang ... Zhou-Ping Yin
IEEE Signal Processing Letters | VOL. 23
Shaohua Zhang, et. al.Shaohua Zhang ... Zhou-Ping Yin
01 Apr 2016
IEEE Signal Processing Letters | VOL. 23

Object detection and classification: a joint selection and fusion strategy of deep convolutional neural network and SIFT point features
Muhammad Rashid ... Muhammad Masood Sarfraz
Multimedia Tools and Applications | VOL. 78
Muhammad Rashid, et. al.Muhammad Rashid ... Muhammad Masood Sarfraz
08 Dec 2018
Multimedia Tools and Applications | VOL. 78

Evaluation of Feature Channels for Correlation-Filter-Based Visual Object Tracking in Infrared Spectrum
Erhan Gundogdu ... Berkan Solmaz
-
Erhan Gundogdu, et. al.Erhan Gundogdu ... Berkan Solmaz
01 Jun 2016
01 Jun 2016

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

G-MS2F: GoogLeNet based multi-stage feature fusion of deep CNN for scene recognition

Abstract

Talk to us

Similar Papers

More From: Neurocomputing