Lighter and Faster Two-Pathway CMRNet for Video Saliency Prediction

Sai Phani Kumar Malladi,Mohamed-Chaker Larabi,Jayanta Mukhopadhyay,Santanu Chaudhury

doi:10.1109/icip46576.2022.9897252

Abstract

Existing dynamic saliency prediction models face challenges like inefficient spatio-temporal feature integration, ineffective multi-scale feature extraction, and lacking domain adaptation because of huge pre-trained backbone networks. In this paper, we propose a two pathway architecture with effective feature integration of spatial and temporal domains at multiple scales for video saliency prediction. Frame and optical flow pathways extract features from video frame and optical flow maps, respectively using a series of cross-concatenated multi-scale residual (CMR) blocks. We name this network as two-pathway CMRNet (TP-CMRNet). Every CMR block follows a feature fusion and attention module for merging features from two pathways and guiding the network to weigh salient regions, respectively. A bi-directional LSTM module is used for learning the task by looking at previous and next video frames. We build a simple decoder for feature reconstruction into the final attention map. TP-CMRNet is comprehensively evaluated using three benchmark datasets: DHF1K, Hollywood-2, and UCF sports. We observe that our model performs at par with other deep dynamic models. In particular, we outperform all the other models with a lesser number of model parameters and lower inference time.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Lighter and Faster Two-Pathway CMRNet for Video Saliency Prediction

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Optical flow correction network for repairing video
Fujie Huang ... Bin Luo
Signal, Image and Video Processing | VOL. 17
Fujie Huang, et. al.Fujie Huang ... Bin Luo
20 Feb 2023
Signal, Image and Video Processing | VOL. 17

Cross-Modality Fusion and Progressive Integration Network for Saliency Prediction on Stereoscopic 3D Images
Yudong Mao ... Runmin Cong
IEEE Transactions on Multimedia | VOL. 24
Yudong Mao, et. al.Yudong Mao ... Runmin Cong
21 May 2021
IEEE Transactions on Multimedia | VOL. 24

Weakly Supervised Video Object Segmentation
Yufei Wang ... Junhu Wang
-
Yufei Wang, et. al.Yufei Wang ... Junhu Wang
01 Oct 2018
01 Oct 2018

Saliency Prediction Network for $360^\circ$ Videos
Youqiang Zhang ... Yike Ma
IEEE Journal of Selected Topics in Signal Processing | VOL. 14
Youqiang Zhang, et. al.Youqiang Zhang ... Yike Ma
05 Dec 2019
IEEE Journal of Selected Topics in Signal Processing | VOL. 14

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Lighter and Faster Two-Pathway CMRNet for Video Saliency Prediction

Abstract

Talk to us

Similar Papers