Chinese News Text Classification Method via Key Feature Enhancement

Bin Ge,Chunhui He,Jiuyang Tang,Hao Xu,Jibing Wu

doi:10.3390/app13095399

Bin Ge, Chunhui He + Show 3 more

Open Access

https://doi.org/10.3390/app13095399

Copy DOI

Journal: Applied Sciences	Publication Date: Apr 26, 2023
Citations: 1	License type: CC BY 4.0

Affiliation: National University of Defense Technology

Abstract

(1) Background: Chinese news text is a popular form of media communication, which can be seen everywhere in China. Chinese news text classification is an important direction in natural language processing (NLP). How to use high-quality text classification technology to help humans to efficiently organize and manage the massive amount of web news is an urgent problem to be solved. It is noted that the existing deep learning methods rely on a large-scale tagged corpus for news text classification tasks and this model is poorly interpretable because the size is large. (2) Methods: To solve the above problems, this paper proposes a Chinese news text classification method based on key feature enhancement named KFE-CNN. It can effectively expand the semantic information of key features to enhance sample data and then combine the zero–one binary vector representation to transform text features into binary vectors and input them into CNN model for training and implementation, thus improving the interpretability of the model and effectively compressing the size of the model. (3) Results: The experimental results show that our method can significantly improve the overall performance of the model and the average accuracy and F1-score of the THUCNews subset of the public dataset reached 97.84% and 98%. (4) Conclusions: this fully proved the effectiveness of the KFE-CNN method for the Chinese news text classification task and it also fully demonstrates that key feature enhancement can improve classification performance.

Full Text