Negative Pre-aware for Noisy Cross-Modal Matching

Xu Zhang,Mang Ye,Hao Li

doi:10.1609/aaai.v38i7.28564

Abstract

Cross-modal noise-robust learning is a challenging task since noisy correspondence is hard to recognize and rectify. Due to the cumulative and unavoidable negative impact of unresolved noise, existing methods cannot maintain a stable performance when the noise increases. In this paper, we present a novel Negative Pre-aware Cross-modal (NPC) matching solution for large visual-language model fine-tuning on noisy downstream tasks. It is featured in two aspects: (1) For noise recognition and resistance, previous methods usually directly filter out a noise subset, we propose to estimate the negative impact of each sample. It does not need additional correction mechanisms that may predict unreliable correction results, leading to self-reinforcing error. We assign a confidence weight to each sample according to its negative impact in the training process. This adaptively adjusts the contribution of each sample to avoid noisy accumulation. (2) For maintaining stable performance with increasing noise, we utilize the memorization effect of DNNs by maintaining a memory bank. Specifically, we apply GMM to select high-confident clean samples as the memory entry, where the memory entry is used to estimate the negative impact of each sample. Since clean samples are easier distinguished by GMM with increasing noise, the memory bank can still maintain high quality at a high noise ratio. Compared to the correction mechanism focusing on noise samples, memory bank-based estimation is more robust, which makes the model performance stable on noisy datasets. Extensive experiments demonstrate that our method significantly improves matching accuracy and performance stability at increasing noise ratio. Our approach also surpasses the state-of-the-art methods by a large margin. The code is available at: https://github.com/ZhangXu0963/NPC.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Negative Pre-aware for Noisy Cross-Modal Matching

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence

Lead the way for us

Journal: Proceedings of the AAAI Conference on Artificial Intelligence	Publication Date: Mar 24, 2024
Citations: 1

Similar Papers

Bias Mitigation and Representation Optimization for Noise-Robust Cross-modal Retrieval
Yu Liu ... Xun Yang
ACM Transactions on Multimedia Computing, Communications, and Applications | VOL. -
Yu Liu, et. al.Yu Liu ... Xun Yang
14 Oct 2024
ACM Transactions on Multimedia Computing, Communications, and Applications | VOL. -

Learning from Noisy Labeled Data via Sharpen Prediction Loss and Re-Correction
Juncheng Wang ... Jie Geng
-
Juncheng Wang, et. al.Juncheng Wang ... Jie Geng
14 Oct 2021
14 Oct 2021

Boosted RVM Algorithm for Imbalanced and Noisy Data
Wangchen Qin ... Quan Qi
-
Wangchen Qin, et. al.Wangchen Qin ... Quan Qi
01 Jul 2018
01 Jul 2018

3D Point Cloud Completion with Geometric-Aware Adversarial Augmentation
Mengxi Wu ... Yi Fang
-
Mengxi Wu, et. al.Mengxi Wu ... Yi Fang
21 Aug 2022
21 Aug 2022

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Negative Pre-aware for Noisy Cross-Modal Matching

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence