Efficient modal-aware feature learning with application in multimodal hashing

Hanlu Chu,Hanjiang Lai,Haien Zeng,Yong Tang

doi:10.3233/ida-215780

Abstract

Many retrieval applications can benefit from multiple modalities, for which how to represent multimodal data is the critical component. Most deep multimodal learning methods typically involve two steps to construct the joint representations: 1) learning of multiple intermediate features, with each intermediate feature corresponding to a modality, using separate and independent deep models; 2) merging the intermediate features into a joint representation using a fusion strategy. However, in the first step, these intermediate features do not have previous knowledge of each other and cannot fully exploit the information contained in the other modalities. In this paper, we present a modal-aware operation as a generic building block to capture the non-linear dependencies among the heterogeneous intermediate features, which can learn the underlying correlation structures in other multimodal data as soon as possible. The modal-aware operation consists of a kernel network and an attention network. The kernel network is utilized to learn the non-linear relationships with other modalities. The attention network finds the informative regions of these modal-aware features that are favorable for retrieval. We verify the proposed modal-aware feature learning in the multimodal hashing task. The experiments conducted on three public benchmark datasets demonstrate significant improvements in the performance of our method relative to state-of-the-art methods.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Efficient modal-aware feature learning with application in multimodal hashing

Abstract

Talk to us

Similar Papers

More From: Intelligent Data Analysis

Lead the way for us

Similar Papers

Hybrid CF on Modeling Feature Importance with Joint Denoising AutoEncoder and SVD++
Qing Yang ... Ya Zhou
-
Qing Yang, et. al.Qing Yang ... Ya Zhou
01 Jan 2020
01 Jan 2020

Using LFDA to Learn Subset-Haar-Like Intermediate Feature Weights for Pedestrian Detection
Kai Zang ... Wankou Yang
-
Kai Zang, et. al.Kai Zang ... Wankou Yang
01 Jan 2017
01 Jan 2017

Addi-Reg: A Better Generalization-Optimization Tradeoff Regularization Method for Convolutional Neural Networks.
Yao Lu ... Zheng Zhang
IEEE transactions on cybernetics | VOL. 52
Yao Lu, et. al.Yao Lu ... Zheng Zhang
22 Mar 2021
IEEE transactions on cybernetics | VOL. 52

Two-Stage Sketch Colorization With Color Parsing
Hui Ren ... Nan Gao
IEEE Access | VOL. 8
Hui Ren, et. al.Hui Ren ... Nan Gao
01 Jan 2020
IEEE Access | VOL. 8

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Efficient modal-aware feature learning with application in multimodal hashing

Abstract

Talk to us

Similar Papers

More From: Intelligent Data Analysis