Unsupervised Domain Adaptation for Hate Speech Detection Using a Data Augmentation Approach

Sheikh Muhammad Sarwar,Vanessa Murdock

doi:10.1609/icwsm.v16i1.19340

Abstract

Online harassment in the form of hate speech has been on the rise in recent years. Addressing the issue requires a combination of content moderation by people, aided by automatic detection methods. As content moderation is itself harmful to the people doing it, we desire to reduce the burden by improving the automatic detection of hate speech. Hate speech presents a challenge as it is directed at different target groups using a completely different vocabulary. Further the authors of the hate speech are incentivized to disguise their behavior to avoid being removed from a platform. This makes it difficult to develop a comprehensive data set for training and evaluating hate speech detection models because the examples that represent one hate speech domain do not typically represent others, even within the same language or culture. We propose an unsupervised domain adaptation approach to augment labeled data for hate speech detection. We evaluate the approach with three different models (character CNNs, BiLSTMs and BERT) on three different collections. We show our approach improves Area under the Precision/Recall curve by as much as 42% and recall by as much as 278%, with no loss (and in some cases a significant gain) in precision.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Unsupervised Domain Adaptation for Hate Speech Detection Using a Data Augmentation Approach

Abstract

Talk to us

Similar Papers

More From: Proceedings of the International AAAI Conference on Web and Social Media

Lead the way for us

Journal: Proceedings of the International AAAI Conference on Web and Social Media	Publication Date: May 31, 2022
Citations: 7

Similar Papers

Domain Adaptive Hate Speech Detection
...
-
, et. al. ...
05 Jun 2022
05 Jun 2022

Hate Speech: A Pragmatic Assessment of the European Court of Human Rights’ Jurisprudence
Alessio Sardo
European Convention on Human Rights Law Review | VOL. 4
Alessio SardoAlessio Sardo
29 Nov 2022
European Convention on Human Rights Law Review | VOL. 4

Form of Hate Speech Comments on Najwa Shihab Youtube Channels in The General Election Campaign of President and Vice President of The Republic of Indonesia 2019
Rahayu Pristiwati ... Tsalisa Yuliyanti
Seloka: Jurnal Pendidikan Bahasa dan Sastra Indonesia | VOL. 9
Rahayu Pristiwati, et. al.Rahayu Pristiwati ... Tsalisa Yuliyanti
31 Dec 2020
Seloka: Jurnal Pendidikan Bahasa dan Sastra Indonesia | VOL. 9

UJARAN KEBENCIAN (KHITĀB AL-KARĀHIYAH) DALAM RUANG KONTESTASI SOSIAL POLITIK ARAB KONTEMPORER
Yoyo Yoyo
Adabiyyāt: Jurnal Bahasa dan Sastra | VOL. 3
Yoyo YoyoYoyo Yoyo
18 Jun 2019
Adabiyyāt: Jurnal Bahasa dan Sastra | VOL. 3

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Unsupervised Domain Adaptation for Hate Speech Detection Using a Data Augmentation Approach

Abstract

Talk to us

Similar Papers

More From: Proceedings of the International AAAI Conference on Web and Social Media