Automatic document classification: the role of interclass similarity

Claudio Isaac Soriano-Burgos,Rafael Guzmán-Cabrera,Misael López-Ramírez

doi:10.35429/jedt.2022.10.8.33.39

Abstract

The continuous increase of information in digital format requires new methods and techniques to access, collect and organize these volumes of textual information. One of the most widely used techniques to organize information is the automatic classification of documents. Automatic text classification systems have a low efficiency when the classes are very similar, i.e. there is overlap between them, and in this case it is very important to be able to identify those attributes that allow us to separate one class from another. In this paper we present the relationship between overlap between classes and classification accuracy. A public corpus with four classes is used for the evaluation, and each class is further separated by positives and negatives. The results obtained from four subsets with different number of training instances are presented, for each case the similarity plots, the accuracy value and the confusion matrices obtained are presented. The results obtained are very illustrative and show that the higher the similarity between classes, the lower the classification accuracy.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Automatic document classification: the role of interclass similarity

Abstract

Talk to us

Similar Papers

More From: Journal Economic Development Technological Chance and Growth

Lead the way for us

Similar Papers

A Naive Bayes Based Tree Classification System
Xun Lu ... Yingyu Chen
-
Xun Lu, et. al.Xun Lu ... Yingyu Chen
01 Jan 2015
01 Jan 2015

Development of an Automatic Measurement and Classification System for a Robotic Arm Using Machine Vision
Ngoc Vu Ngo ... Van Cuong Duong
International Journal of Mechanical Engineering and Robotics Research | VOL. 13
Ngoc Vu Ngo, et. al.Ngoc Vu Ngo ... Van Cuong Duong
01 Jan 2024
International Journal of Mechanical Engineering and Robotics Research | VOL. 13

Evaluation of two semi-supervised learning methods and their combination for automatic classification of bone marrow cells
Iori Nakamura ... Hiroshi Kanai
Scientific Reports | VOL. 12
Iori Nakamura, et. al.Iori Nakamura ... Hiroshi Kanai
06 Oct 2022
Scientific Reports | VOL. 12

Improving Automatic Music Genre Classification Systems by Using Descriptive Statistical Features of Audio Signals
Ravindu Perera ... Manjusri Wickramasinghe
-
Ravindu Perera, et. al.Ravindu Perera ... Manjusri Wickramasinghe
01 Jan 2023
01 Jan 2023

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Automatic document classification: the role of interclass similarity

Abstract

Talk to us

Similar Papers

More From: Journal Economic Development Technological Chance and Growth