Bag of meta-words: A novel method to represent document for the sentiment classification

Mingsheng Fu,Hong Qu,Li Huang,Li Lu

doi:10.1016/j.eswa.2018.06.052

Abstract

It is crucial to represent the semantic information of a document in sentiment classification. Various semantic information representation models have been proposed, however existing approaches have their setbacks. Notable weaknesses among these are: (1) tradition VSM methods, completely ignore the semantic information; (2) averaging word embedding methods, cannot depict the synthetical semantic meaning of the given document; (3) neural network methods, require complex structure and are notoriously difficult to be trained. To overcome these limitations, we introduce a simple but novel method which we call bag of meta-words (BoMW). In our method, the semantic information of the document is indicated by a meta-words vector in which every single meta-word element denotes particular semantic information. Especially, these meta-words are extracted from pre-trained word embeddings through two different but complemental models, naive interval meta-words (NIM) and feature combination meta-words (FCM). In general, our new model BoMW is as simple as traditional VSM model but it can capture the synthetical semantic meanings of the document. Numerous experiments on two benchmarks (IMDB dataset and Pang’s dataset) are carried out to verify the effectiveness of the proposed method, and the results show that the performance of our method can exceed the traditional VSM methods and methods using pre-trained word embedding.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Bag of meta-words: A novel method to represent document for the sentiment classification

Abstract

Talk to us

Similar Papers

More From: Expert Systems with Applications

Lead the way for us

Journal: Expert Systems with Applications	Publication Date: Jun 30, 2018
Citations: 21

Similar Papers

Automatic text summarization of konkani texts using pre-trained word embeddings and deep learning
Jovi D’Silva ... Uzzal Sharma
International Journal of Electrical and Computer Engineering (IJECE) | VOL. 12
Jovi D’Silva, et. al.Jovi D’Silva ... Uzzal Sharma
01 Apr 2022
International Journal of Electrical and Computer Engineering (IJECE) | VOL. 12

Leveraging Pre-Trained Contextualized Word Embeddings to Enhance Sentiment Classification of Drug Reviews
Redouane Karsi ... Mounia Zaim
Revue d'Intelligence Artificielle | VOL. 35
Redouane Karsi, et. al.Redouane Karsi ... Mounia Zaim
31 Aug 2021
Revue d'Intelligence Artificielle | VOL. 35

A Deep Learning Architecture with Word Embeddings to Classify Sentiment in Twitter
Eman Hamdi ... Mostafa Aref
-
Eman Hamdi, et. al.Eman Hamdi ... Mostafa Aref
20 Sep 2020
20 Sep 2020

A knowledge-enriched ensemble method for word embedding and multi-sense embedding
Lanting Fang ... Kaiqi Zhao
IEEE Transactions on Knowledge and Data Engineering | VOL. -
Lanting Fang, et. al.Lanting Fang ... Kaiqi Zhao
01 Jan 2021
IEEE Transactions on Knowledge and Data Engineering | VOL. -

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Bag of meta-words: A novel method to represent document for the sentiment classification

Abstract

Talk to us

Similar Papers

More From: Expert Systems with Applications