HumourHindiNet: Humour detection in Hindi web series using word embedding and convolutional neural network

Akshi Kumar,Sanjay Kumar,Abhishek Mallik

doi:10.1145/3661306

Abstract

Humour is a crucial aspect of human speech, and it is, therefore, imperative to create a system that can offer such detection. While data regarding humour in English speech is plentiful, the same cannot be said for a low-resource language like Hindi. Through this article, we introduce two multimodal datasets for humour detection in the Hindi web series. The dataset was collected from over 500 minutes of conversations amongst the characters of the Hindi web series Kota-Factory and Panchayat . Each dialogue is manually annotated as Humour or Non-Humour. Along with presenting a new Hindi language-based Humour detection dataset, we propose an improved framework for detecting humour in Hindi conversations. We start by preprocessing both datasets to obtain uniformity across the dialogues and datasets. The processed dialogues are then passed through the Skip-gram model for generating Hindi word embedding. The generated Hindi word embedding is then passed onto three convolutional neural network (CNN) architectures simultaneously, each having a different filter size for feature extraction. The extracted features are then passed through stacked Long Short-Term Memory (LSTM) layers for further processing and finally classifying the dialogues as Humour or Non-Humour. We conduct intensive experiments on both proposed Hindi datasets and evaluate several standard performance metrics. The performance of our proposed framework was also compared with several baselines and contemporary algorithms for Humour detection. The results demonstrate the effectiveness of our dataset to be used as a standard dataset for Humour detection in the Hindi web series. The proposed model yields an accuracy of 91.79 and 87.32 while an F1 score of 91.64 and 87.04 in percentage for the Kota-Factory and Panchayat datasets, respectively.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

HumourHindiNet: Humour detection in Hindi web series using word embedding and convolutional neural network

Abstract

Talk to us

Similar Papers

More From: ACM Transactions on Asian and Low-Resource Language Information Processing

Lead the way for us

Journal: ACM Transactions on Asian and Low-Resource Language Information Processing	Publication Date: Jun 26, 2024
License type: mit

Similar Papers

Gujarati Task Oriented Dialogue Slot Tagging Using Deep Neural Network Models
Rachana Parikh ... Hiren Joshi
-
Rachana Parikh, et. al.Rachana Parikh ... Hiren Joshi
01 Jan 2020
01 Jan 2020

Enhancing vessel arrival time prediction: A fusion-based deep learning approach
Asad Abdi ... Chintan Amrit
Expert Systems With Applications | VOL. 252
Asad Abdi, et. al.Asad Abdi ... Chintan Amrit
03 May 2024
Expert Systems With Applications | VOL. 252

Performance of Three Slim Variants of The Long Short-Term Memory (LSTM) Layer
Daniel Kent ... Fathi Salem
-
Daniel Kent, et. al.Daniel Kent ... Fathi Salem
01 Aug 2019
01 Aug 2019

Share Price Trend Prediction Using CRNN with LSTM Structure
Shyr-Shen Yu ... Chuin-Mu Wang
Smart Science | VOL. 7
Shyr-Shen Yu, et. al.Shyr-Shen Yu ... Chuin-Mu Wang
18 Apr 2019
Smart Science | VOL. 7

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

HumourHindiNet: Humour detection in Hindi web series using word embedding and convolutional neural network

Abstract

Talk to us

Similar Papers

More From: ACM Transactions on Asian and Low-Resource Language Information Processing