An Analysis of Boosted Linear Classifiers on Noisy Data with Applications to Multiple-Instance Learning

Rui Liu,Soumya Ray

doi:10.1109/icdm.2017.38

Abstract

An interesting observation about the well-known AdaBoost algorithm is that, though theory suggests it should overfit when applied to noisy data, experiments indicate it often does not do so in practice. In this paper, we study the behavior of AdaBoost on datasets with one-sided uniform class noise using linear classifiers as the base learner. We show analytically that, under some ideal conditions, this approach will not overfit, and can in fact recover a zero-error concept with respect to the true, uncorrupted instance labels. We also analytically show that AdaBoost increases the margins of predictions over boosting iterations, as has been previously suggested in the literature. We then compare the empirical behavior of AdaBoost using real world datasets with one-sided noise derived from multiple-instance data. Although our assumptions may not hold in a practical setting, our experiments show that standard AdaBoost still performs well, as suggested by our analysis, and often outperforms baseline variations in the literature that explicitly try to account for noise.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

An Analysis of Boosted Linear Classifiers on Noisy Data with Applications to Multiple-Instance Learning

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Supervised versus multiple instance learning
Soumya Ray ... Mark Craven
-
Soumya Ray, et. al.Soumya Ray ... Mark Craven
01 Jan 2004
01 Jan 2004

The Effectiveness of a New Negative Correlation Learning Algorithm for Classification Ensembles
Shuo Wang ... Xin Yao
-
Shuo Wang, et. al.Shuo Wang ... Xin Yao
01 Dec 2010
01 Dec 2010

Learning Instance Concepts from Multiple-Instance Data with Bags as Distributions
Gary Doran ... Soumya Ray
Proceedings of the AAAI Conference on Artificial Intelligence | VOL. 28
Gary Doran, et. al.Gary Doran ... Soumya Ray
21 Jun 2014
Proceedings of the AAAI Conference on Artificial Intelligence | VOL. 28

Multiple Instance Learning with multiple positive and negative target concepts
Andrew Karem ... Hichem Frigui
-
Andrew Karem, et. al.Andrew Karem ... Hichem Frigui
01 Dec 2016
01 Dec 2016

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

An Analysis of Boosted Linear Classifiers on Noisy Data with Applications to Multiple-Instance Learning

Abstract

Talk to us

Similar Papers