GPT-4 as an X data annotator: Unraveling its performance on a stance classification task.

Chandreen R Liyanage,Ravi Gokani,Vijay Mago

doi:10.1371/journal.pone.0307741

Chandreen R Liyanage, Ravi Gokani + Show 1 more

Open Access

https://doi.org/10.1371/journal.pone.0307741

Copy DOI

Export

Save

Cite

Journal: PloS one	Publication Date: Aug 15, 2024
Citations: 1	License type: CC BY 4.0

Abstract
Full-Text
Similar Papers

Abstract

Listen

Data annotation in NLP is a costly and time-consuming task, traditionally handled by human experts who require extensive training to enhance the task-related background knowledge. Besides, labeling social media texts is particularly challenging due to their brevity, informality, creativity, and varying human perceptions regarding the sociocultural context of the world. With the emergence of GPT models and their proficiency in various NLP tasks, this study aims to establish a performance baseline for GPT-4 as a social media text annotator. To achieve this, we employ our own dataset of tweets, expertly labeled for stance detection with full inter-rater agreement among three annotators. We experiment with three techniques: Zero-shot, Few-shot, and Zero-shot with Chain-of-Thoughts to create prompts for the labeling task. We utilize four training sets constructed with different label sets, including human labels, to fine-tune transformer-based large language models and various combinations of traditional machine learning models with embeddings for stance classification. Finally, all fine-tuned models undergo evaluation using a common testing set with human-generated labels. We use the results from models trained on human labels as the benchmark to assess GPT-4's potential as an annotator across the three prompting techniques. Based on the experimental findings, GPT-4 achieves comparable results through the Few-shot and Zero-shot Chain-of-Thoughts prompting methods. However, none of these labeling techniques surpass the top three models fine-tuned on human labels. Moreover, we introduce the Zero-shot Chain-of-Thoughts as an effective strategy for aspect-based social media text labeling, which performs better than the standard Zero-shot and yields results similar to the high-performing yet expensive Few-shot approach.

Full Text

Published Version

View

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

GPT-4 as an X data annotator: Unraveling its performance on a stance classification task.

Abstract

Published Version

Talk to us

Similar Papers

More From: PloS one

Lead the way for us

Similar Papers

Detecting Signs of Depression from Social Media Text using RoBERTa Pre-trained Language Models
...
-
, et. al. ...
12 May 2022
12 May 2022

Application of Transformer-Based Language Models to Detect Hate Speech in Social Media
Swapnanil Mukherjee ... Sujit Das
Journal of Computational and Cognitive Engineering | VOL. 2
Swapnanil Mukherjee, et. al.Swapnanil Mukherjee ... Sujit Das
17 Dec 2021
Journal of Computational and Cognitive Engineering | VOL. 2

Classifying Drug Ratings Using User Reviews with Transformer-Based Language Models
Akhil Shiju ... Zhe He
-
Akhil Shiju, et. al.Akhil Shiju ... Zhe He
01 Jun 2022
01 Jun 2022

Task-Specific Transformer-Based Language Models in Health Care: Scoping Review.
Ha Na Cho ... Soyoung Ko
JMIR medical informatics | VOL. 12
Ha Na Cho, et. al.Ha Na Cho ... Soyoung Ko
18 Nov 2024
JMIR medical informatics | VOL. 12

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

GPT-4 as an X data annotator: Unraveling its performance on a stance classification task.

Abstract

Published Version

Talk to us

Similar Papers

More From: PloS one