Scalable Penalized Regression for Noise Detection in Learning with Noisy Labels

Yikai Wang,Xinwei Sun,Yanwei Fu

doi:10.1109/cvpr52688.2022.00044

Abstract

Noisy training set usually leads to the degradation of generalization and robustness of neural networks. In this paper, we propose using a theoretically guaranteed noisy label detection framework to detect and remove noisy data for Learning with Noisy Labels (LNL). Specifically, we design a penalized regression to model the linear relation between network features and one-hot labels, where the noisy data are identified by the non-zero mean shift parameters solved in the regression model. To make the framework scalable to datasets that contain a large number of categories and training data, we propose a split algorithm to divide the whole training set into small pieces that can be solved by the penalized regression in parallel, leading to the Scalable Penalized Regression (SPR) framework. We provide the non-asymptotic probabilistic condition for SP R to correctly identify the noisy data. While SPR can be regarded as a sample selection module for standard supervised training pipeline, we further combine it with semi-supervised algorithm to further exploit the support of noisy data as unlabeled data. Experimental results on several benchmark datasets and real-world noisy datasets show the effectiveness of our framework. Our code and pretrained models are released at https://github.com/Yikai-Wang/SPR-LNL.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Scalable Penalized Regression for Noise Detection in Learning with Noisy Labels

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Knockoffs-SPR: Clean Sample Selection in Learning With Noisy Labels.
Yikai Wang ... Xinwei Sun
IEEE transactions on pattern analysis and machine intelligence | VOL. 46
Yikai Wang, et. al.Yikai Wang ... Xinwei Sun
01 May 2024
IEEE transactions on pattern analysis and machine intelligence | VOL. 46

Skin Lesion Segmentation in Dermoscopic Images with Noisy Data.
Norsang Lama ... William Van Stoecker
Journal of digital imaging | VOL. 36
Norsang Lama, et. al.Norsang Lama ... William Van Stoecker
05 Apr 2023
Journal of digital imaging | VOL. 36

Hyperspectral Classification With Noisy Label Detection via Superpixel-to-Pixel Weighting Distance
Bing Tu ... Siyuan Huang
IEEE Transactions on Geoscience and Remote Sensing | VOL. 58
Bing Tu, et. al.Bing Tu ... Siyuan Huang
01 Jun 2020
IEEE Transactions on Geoscience and Remote Sensing | VOL. 58

S-CUDA: Self-cleansing unsupervised domain adaptation for medical image segmentation.
Luyan Liu ... Yefeng Zheng
Medical Image Analysis | VOL. 74
Luyan Liu, et. al.Luyan Liu ... Yefeng Zheng
01 Dec 2021
Medical Image Analysis | VOL. 74

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Scalable Penalized Regression for Noise Detection in Learning with Noisy Labels

Abstract

Talk to us

Similar Papers