Tweet sentiment quantification: An experimental re-evaluation.

Alejandro Moreo,Fabrizio Sebastiani

doi:10.1371/journal.pone.0263449

Alejandro Moreo, Fabrizio Sebastiani

Open Access

https://doi.org/10.1371/journal.pone.0263449

Copy DOI

Journal: PLOS ONE	Publication Date: Sep 16, 2022
Citations: 13	License type: CC BY 4.0

Affiliation: Institute of Information Science and Technologies

Abstract

Sentiment quantification is the task of training, by means of supervised learning, estimators of the relative frequency (also called “prevalence”) of sentiment-related classes (such as Positive, Neutral, Negative) in a sample of unlabelled texts. This task is especially important when these texts are tweets, since the final goal of most sentiment classification efforts carried out on Twitter data is actually quantification (and not the classification of individual tweets). It is well-known that solving quantification by means of “classify and count” (i.e., by classifying all unlabelled items by means of a standard classifier and counting the items that have been assigned to a given class) is less than optimal in terms of accuracy, and that more accurate quantification methods exist. Gao and Sebastiani 2016 carried out a systematic comparison of quantification methods on the task of tweet sentiment quantification. In hindsight, we observe that the experimentation carried out in that work was weak, and that the reliability of the conclusions that were drawn from the results is thus questionable. We here re-evaluate those quantification methods (plus a few more modern ones) on exactly the same datasets, this time following a now consolidated and robust experimental protocol (which also involves simulating the presence, in the test data, of class prevalence values very different from those of the training set). This experimental protocol (even without counting the newly added methods) involves a number of experiments 5,775 times larger than that of the original study. Due to the above-mentioned presence, in the test data, of samples characterised by class prevalence values very different from those of the training set, the results of our experiments are dramatically different from those obtained by Gao and Sebastiani, and provide a different, much more solid understanding of the relative strengths and weaknesses of different sentiment quantification methods.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Tweet sentiment quantification: An experimental re-evaluation.

Abstract

Talk to us

Similar Papers

More From: PLOS ONE

Lead the way for us

Similar Papers

Aspect-Based Sentiment Quantification
Vladyslav Matsiiako ... Flavius Frasincar
IEEE Transactions on Affective Computing | VOL. 13
Vladyslav Matsiiako, et. al.Vladyslav Matsiiako ... Flavius Frasincar
01 Oct 2022
IEEE Transactions on Affective Computing | VOL. 13

Cross-Lingual Sentiment Quantification
Andrea Esuli ... Alejandro Moreo
IEEE Intelligent Systems | VOL. 35
Andrea Esuli, et. al.Andrea Esuli ... Alejandro Moreo
16 Apr 2019
IEEE Intelligent Systems | VOL. 35

Comparing Twitter data to routine data sources in public health surveillance for the 2015 Pan/Parapan American Games: an ecological study.
Yasmin Khan ... Ian L Johnson
Canadian journal of public health = Revue canadienne de sante publique | VOL. 109
Yasmin Khan, et. al.Yasmin Khan ... Ian L Johnson
20 Apr 2018
Canadian journal of public health = Revue canadienne de sante publique | VOL. 109

Classification of Covid-19 Tweets Using Deep Learning Techniques
Pramod Sunagar ... V M Hemanth
-
Pramod Sunagar, et. al.Pramod Sunagar ... V M Hemanth
01 Jan 2020
01 Jan 2020

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Tweet sentiment quantification: An experimental re-evaluation.

Abstract

Talk to us

Similar Papers

More From: PLOS ONE