Distributed learning with bagging-like performance

Nitesh V Chawla,Thomas E Moore,Lawrence O Hall,Kevin W Bowyer,W.Philip Kegelmeyer,Clayton Springer

doi:10.1016/s0167-8655(02)00269-6

Abstract

Bagging forms a committee of classifiers by bootstrap aggregation of training sets from a pool of training data. A simple alternative to bagging is to partition the data into disjoint subsets. Experiments with decision tree and neural network classifiers on various datasets show that, given the same size partitions and bags, disjoint partitions result in performance equivalent to, or better than, bootstrap aggregates (bags). Many applications (e.g., protein structure prediction) involve use of datasets that are too large to handle in the memory of the typical computer. Hence, bagging with samples the size of the data is impractical. Our results indicate that, in such applications, the simple approach of creating a committee of n classifiers from disjoint partitions each of size 1/n (which will be memory resident during learning) in a distributed way results in a classifier which has a bagging-like performance gain. The use of distributed disjoint partitions in learning is significantly less complex and faster than bagging.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Distributed learning with bagging-like performance

Abstract

Talk to us

Similar Papers

More From: Pattern Recognition Letters

Lead the way for us

Journal: Pattern Recognition Letters	Publication Date: Oct 18, 2002
Citations: 72

Similar Papers

Bagging is a small-data-set phenomenon
N Chawla ... C Springer
-
N Chawla, et. al.N Chawla ... C Springer
01 Dec 2001
01 Dec 2001

A comparison of decision tree and backpropagation neural network classifiers for land use classification
M Pal ... P.M Mather
-
M Pal, et. al.M Pal ... P.M Mather
07 Nov 2002
07 Nov 2002

Extracting Land Use/Cover of Mountainous Area from Remote Sensing Images Using Artificial Neural Network and Decision Tree Classifications: A Case Study of Meizhou, China
Yong-Zhu Xiong ... Zhi Li
-
Yong-Zhu Xiong, et. al.Yong-Zhu Xiong ... Zhi Li
01 Oct 2010
01 Oct 2010

Weighted objective distance for the classification of elderly people with hypertension
Supansa Chaising ... Ramjee Prasad
Knowledge-Based Systems | VOL. 210
Supansa Chaising, et. al.Supansa Chaising ... Ramjee Prasad
28 Sep 2020
Knowledge-Based Systems | VOL. 210

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Distributed learning with bagging-like performance

Abstract

Talk to us

Similar Papers

More From: Pattern Recognition Letters