An End-to-end Heterogeneous Restraint Network for RGB-D Cross-modal Person Re-identification

Jingjing Wu,Jianguo Jiang,Cuiqun Chen,Meibin Qi,Jingjing Zhang

doi:10.1145/3506708

Abstract

The RGB-D cross-modal person re-identification (re-id) task aims to identify the person of interest across the RGB and depth image modes. The tremendous discrepancy between these two modalities makes this task difficult to tackle. Few researchers pay attention to this task, and the deep networks of existing methods still cannot be trained in an end-to-end manner. Therefore, this article proposes an end-to-end module for RGB-D cross-modal person re-id. This network introduces a cross-modal relational branch to narrow the gaps between two heterogeneous images. It models the abundant correlations between any cross-modal sample pairs, which are constrained by heterogeneous interactive learning. The proposed network also exploits a dual-modal local branch, which aims to capture the common spatial contexts in two modalities. This branch adopts shared attentive pooling and mutual contextual graph networks to extract the spatial attention within each local region and the spatial relations between distinct local parts, respectively. Experimental results on two public benchmark datasets, that is, the BIWI and RobotPKU datasets, demonstrate that our method is superior to the state-of-the-art. In addition, we perform thorough experiments to prove the effectiveness of each component in the proposed method.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

An End-to-end Heterogeneous Restraint Network for RGB-D Cross-modal Person Re-identification

Abstract

Talk to us

Similar Papers

More From: ACM Transactions on Multimedia Computing, Communications, and Applications

Lead the way for us

Journal: ACM Transactions on Multimedia Computing, Communications, and Applications	Publication Date: Mar 4, 2022
Citations: 8

Similar Papers

Person re-identification based on frequency channel attention networks under the surveillance scenario
Shengbo Chen ... Hongchang Zhang
Journal of Physics: Conference Series | VOL. 1966
Shengbo Chen, et. al.Shengbo Chen ... Hongchang Zhang
01 Jul 2021
Journal of Physics: Conference Series | VOL. 1966

IGIE-net: Cross-modality person re-identification via intermediate modality image generation and discriminative information enhancement
Jiangtao Guo ... Minghao Zhang
Image and Vision Computing | VOL. 147
Jiangtao Guo, et. al.Jiangtao Guo ... Minghao Zhang
09 May 2024
Image and Vision Computing | VOL. 147

Towards robust registration of heterogeneous multispectral UAV imagery: A two-stage approach for cotton leaf lesion grading
Xinzhou Li ... Mingzhou Lu
Computers and Electronics in Agriculture | VOL. 212
Xinzhou Li, et. al.Xinzhou Li ... Mingzhou Lu
27 Aug 2023
Computers and Electronics in Agriculture | VOL. 212

Image Analysis on Symmetric Positive Definite Manifolds
Azadeh Alavi
-
Azadeh AlaviAzadeh Alavi
19 Dec 2014
19 Dec 2014

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

An End-to-end Heterogeneous Restraint Network for RGB-D Cross-modal Person Re-identification

Abstract

Talk to us

Similar Papers

More From: ACM Transactions on Multimedia Computing, Communications, and Applications