Latent Feature Word Representations to Enhance Topic Models for Text Mining Algorithms

Dr Thayyaba Khatoon Mohammed,Dr.Vsk Reddy,M Sandeep,M Gayatri

doi:10.35940/ijeat.b2503.129219

Dr Thayyaba Khatoon Mohammed, Dr.Vsk Reddy + Show 2 more

Open Access

https://doi.org/10.35940/ijeat.b2503.129219

Copy DOI

Export

Save

Cite

Abstract
Full-Text
Similar Papers

Abstract

Listen

Dealing with large number of textual documents needs proven models that leverage the efficiency in processing. Text mining needs such models to have meaningful approaches to extract latent features from document collection. Latent Dirichlet allocation (LDA) is one such probabilistic generative process model that helps in representing document collections in a systematic approach. In many text mining applications LDA is useful as it supports many models. One such model is known as Topic Model. However, topic models LDA needs to be improved in order to exploit latent feature vector representations of words trained on large corpora to improve word-topic mapping learnt on smaller corpus. With respect to document clustering and document classification, it is essential to have a novel topic models to improve performance. In this paper, an improved topic model is proposed and implemented using LDA which exploits the benefits of Word2Vec tool to have pre-trained word vectors so as to achieve the desired enhancement. A prototype application is built to demonstrate the proof of the concept with text mining operations like document clustering.

Full Text

Published Version

View

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

Latent Feature Word Representations to Enhance Topic Models for Text Mining Algorithms

Abstract

Published Version

Talk to us

Similar Papers

More From: International Journal of Engineering and Advanced Technology

Lead the way for us

Similar Papers

An intelligent literature review: adopting inductive approach to define machine learning applications in the clinical domain
Renu Sabharwal ... Shah J Miah
Journal of Big Data | VOL. 9
Renu Sabharwal, et. al.Renu Sabharwal ... Shah J Miah
28 Apr 2022
Journal of Big Data | VOL. 9

Sentiment Analysis of Consumer-Generated Online Reviews of Physical Bookstores Using Hybrid LSTM-CNN and LDA Topic Model
Yan Wang ... Xuteng Wang
-
Yan Wang, et. al.Yan Wang ... Xuteng Wang
01 Oct 2020
01 Oct 2020

Text Similarity Computing Based on LDA Topic Model and Word Co-occurrence
Liangxi Qin ... Minglai Shao
-
Liangxi Qin, et. al.Liangxi Qin ... Minglai Shao
01 Jan 2014
01 Jan 2014

Analyze IMDb movies by sentiment and topic analysis
Ningjing Ouyang
Environment and Social Psychology | VOL. 8
Ningjing OuyangNingjing Ouyang
25 Oct 2023
Environment and Social Psychology | VOL. 8

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

Latent Feature Word Representations to Enhance Topic Models for Text Mining Algorithms

Abstract

Published Version

Talk to us

Similar Papers

More From: International Journal of Engineering and Advanced Technology