Creating a New Dataset for the Classification of Cyber Bullying

Çilem Koçak,Mehmet Bi̇len,Tuncay Yi̇ği̇t

doi:10.54569/aair.1206144

Abstract

Regardless of young or old, people have quickly stepped into the world of internet with today's communication technologies such as phones, tablets, computers and smart devices. As the place of the Internet in people's lives increases, social media platforms are diversifying and users want to take part in these platforms. With the increase in the number of social media users, some negativities are encountered. The most important problem encountered in social media platforms is cyber bullying. Although cyber bullying seems to be a daily dialogue between social media users or between groups, the situation of encountering is increasing day by day with the diversity of shared information, content and agenda social media environments. With the development of technology, it is necessary to develop a platform that detects bullying with artificial intelligence technologies. One of the biggest difficulties in text classification problems that we encounter during the development of these platforms is the need to train the artificial intelligence algorithm to be used with labeled data. In this study, 21 different people, including journalists, athletes, scientists, doctors, politicians, comedians, social media phenomena, and artists who actively use social media, were selected in order to create the necessary dataset for training the models to be developed to detect cyber bullying situations. The public messages (mentions) of these 21 people sent via Twitter were compiled. After filtering the repetitive and meaningless messages sent by bot accounts out of 10500 tweets compiled, the number of messages in the dataset decreased to 7706. The labeling process, which is necessary for the dataset to be used for training and testing purposes in classification processes, was carried out by three independent people who were given preliminary information about cyberbullying (1=Includes Cyber bullying, 0=Does not include Cyber bullying). The majority of the tags, which were read and assigned by 3 different people, were accepted as the final class of the relevant message. Afterwards, the dataset was preprocessed in accordance with the principles of natural language processing and made suitable for classification algorithms. The findings obtained after the classification processes performed with the basic classification algorithms are shared. When the findings are examined, it is understood that the data set created has the competence to be used in the detection and prevention of cyber bullying. In this context, it is predicted that training specially developed and optimized artificial intelligence algorithms with the relevant dataset for the detection of cyberbullying will greatly increase the success rate.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Creating a New Dataset for the Classification of Cyber Bullying

Abstract

Talk to us

Similar Papers

More From: Advances in Artificial Intelligence Research

Lead the way for us

Journal: Advances in Artificial Intelligence Research	Publication Date: Oct 29, 2023
License type: cc-by-nc

Similar Papers

Proactively Discouraging Cyberbullying Activities
Puneetha Kr
International Journal for Research in Applied Science and Engineering Technology | VOL. 9
Puneetha KrPuneetha Kr
31 Oct 2021
International Journal for Research in Applied Science and Engineering Technology | VOL. 9

Detection of Cyber bullying on Social Media using Deep Learning Algorithms
Afroz Abbas A ... Purushotham K
-
Afroz Abbas A, et. al.Afroz Abbas A ... Purushotham K
02 May 2024
02 May 2024

A Review Paper on Cyber Harassment Detection Using Machine Learning Algorithm on Social Networking Website
Miss Sneha Gajanan Sambare ... Dr Sanjay L Haridas
International Journal for Research in Applied Science and Engineering Technology | VOL. 10
Miss Sneha Gajanan Sambare, et. al.Miss Sneha Gajanan Sambare ... Dr Sanjay L Haridas
31 Oct 2022
International Journal for Research in Applied Science and Engineering Technology | VOL. 10

“I Need You All to Understand How Pervasive This Issue Is”: User Efforts to Regulate Child Sexual Offending on Social Media
Michael Salter ... Elly Hanson
-
Michael Salter, et. al.Michael Salter ... Elly Hanson
04 Jun 2021
04 Jun 2021

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Creating a New Dataset for the Classification of Cyber Bullying

Abstract

Talk to us

Similar Papers

More From: Advances in Artificial Intelligence Research