Multi-instance learning (MIL) is a critical area in machine learning, particularly for applications where data points are grouped into bags. Traditional methods, however, often face challenges in accurately classifying these bags. This paper presents the multi-instance partial decision tree (MIPART), a method that incorporates the partial decision tree (PART) algorithm within a Bagging framework, utilizing the simple multi-instance classifier (SimpleMI) as its base. MIPART was evaluated on 12 real-world multi-instance datasets using various performance metrics. Experimental results show that MIPART achieved an average accuracy of 84.27%, outperforming benchmarks in the literature. Notably, MIPART outperformed established methods such as Citation-KNN, MIBoost, MIEMDD, MILR, MISVM, and MITI, demonstrating a 15% improvement in average accuracy across the same datasets. The significance of these improvements was confirmed through rigorous non-parametric statistical tests, including Friedman aligned ranks and Wilcoxon signed-rank analyses. These findings suggest that the MIPART method is a significant advancement in multiple-instance classification, providing an effective tool for interpreting complex multi-instance datasets.
Read full abstract