Classification of Scientific Documents in the Kazakh Language Using Deep Neural Networks and a Fusion of Images and Text

Andrey Bogdanchikov,Iraklis Varlamis,Dauren Ayazbayev

doi:10.3390/bdcc6040123

Abstract

The rapid development of natural language processing and deep learning techniques has boosted the performance of related algorithms in several linguistic and text mining tasks. Consequently, applications such as opinion mining, fake news detection or document classification that assign documents to predefined categories have significantly benefited from pre-trained language models, word or sentence embeddings, linguistic corpora, knowledge graphs and other resources that are in abundance for the more popular languages (e.g., English, Chinese, etc.). Less represented languages, such as the Kazakh language, balkan languages, etc., still lack the necessary linguistic resources and thus the performance of the respective methods is still low. In this work, we develop a model that classifies scientific papers written in the Kazakh language using both text and image information and demonstrate that this fusion of information can be beneficial for cases of languages that have limited resources for machine learning models’ training. With this fusion, we improve the classification accuracy by 4.4499% compared to the models that use only text or only image information. The successful use of the proposed method in scientific documents’ classification paves the way for more complex classification models and more application in other domains such as news classification, sentiment analysis, etc., in the Kazakh language.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Big Data and Cognitive Computing	Publication Date: Oct 24, 2022
Citations: 4	License type: CC BY 4.0

R Discovery Prime

R Discovery Prime

Classification of Scientific Documents in the Kazakh Language Using Deep Neural Networks and a Fusion of Images and Text

Abstract

Talk to us

Similar Papers

More From: Big Data and Cognitive Computing

Lead the way for us

Similar Papers

Natural Language Processing with Optimal Deep Learning Based Fake News Classification
Sara A Althubiti ... Fayadh Alenezi
Computers, Materials & Continua | VOL. 73
Sara A Althubiti, et. al.Sara A Althubiti ... Fayadh Alenezi
01 Jan 2021
Computers, Materials & Continua | VOL. 73

Evaluation of Fake News Detection with Knowledge-Enhanced Language Models
Chenxi Whitehouse ... Tillman Weyde
Proceedings of the International AAAI Conference on Web and Social Media | VOL. 16
Chenxi Whitehouse, et. al.Chenxi Whitehouse ... Tillman Weyde
31 May 2022
Proceedings of the International AAAI Conference on Web and Social Media | VOL. 16

A veracity dissemination consistency-based few-shot fake news detection framework by synergizing adversarial and contrastive self-supervised learning
Weiqiang Jin ... Guang Yang
Scientific Reports | VOL. 14
Weiqiang Jin, et. al.Weiqiang Jin ... Guang Yang
22 Aug 2024
Scientific Reports | VOL. 14

Optimization techniques for sentiment analysis based on LLM (GPT-3)
Tong Zhan ... Huixiang Li
Applied and Computational Engineering | VOL. 67
Tong Zhan, et. al.Tong Zhan ... Huixiang Li
31 May 2024
Applied and Computational Engineering | VOL. 67

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Classification of Scientific Documents in the Kazakh Language Using Deep Neural Networks and a Fusion of Images and Text

Abstract

Talk to us

Similar Papers

More From: Big Data and Cognitive Computing