Modified LDA vector and feedback analysis for short query Information Retrieval systems

Pedro Celard,Rubén Romero,Eva Lorenzo Iglesias,Adrián Seara Vieira,Lourdes Borrajo,José Manuel Sorribes-Fdez

doi:10.1093/jigpal/jzae044

Abstract

Abstract Information Retrieval systems benefit from the use of long queries containing a large volume of search-relevant information. This situation is not common, as users of such systems tend to use very short and precise queries with few keywords. In this work we propose a modification of the Latent Dirichlet Allocation (LDA) technique using data from the document collection and its vocabulary for a better representation of short queries. Additionally, a study is carried out on how the modification of the proposed LDA weighted vectors increase the performance using relevant documents as feedback. The work shown in this paper is tested using three biomedical corpora (TREC Genomics 2004, TREC Genomics 2005 and OHSUMED) and one legal corpus (FIRE 2017). Results prove that the application of the proposed representation technique, as well as the feedback adjustment, clearly outperforms the baseline methods (BM25 and non-modified LDA).

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Modified LDA vector and feedback analysis for short query Information Retrieval systems

Abstract

Talk to us

Similar Papers

More From: Logic Journal of the IGPL

Lead the way for us

Journal: Logic Journal of the IGPL	Publication Date: May 4, 2024
License type: CC BY 4.0

Similar Papers

Improving Short Query Representation in LDA Based Information Retrieval Systems
Pedro Celard ... Lourdes Borrajo
-
Pedro Celard, et. al.Pedro Celard ... Lourdes Borrajo
01 Jan 2021
01 Jan 2021

Latent word context model for information retrieval
Bernard Brosseau-Villeneuve ... Jian-Yun Nie
Information Retrieval | VOL. 17
Bernard Brosseau-Villeneuve, et. al.Bernard Brosseau-Villeneuve ... Jian-Yun Nie
05 Mar 2013
Information Retrieval | VOL. 17

Query length impact on misuse detection in information retrieval systems
Ling Ma ... Nazli Goharian
-
Ling Ma, et. al.Ling Ma ... Nazli Goharian
13 Mar 2005
13 Mar 2005

Models, Inference, and Implementation for Scalable Probabilistic Models of Text

-

01 Jan 2014
01 Jan 2014

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Modified LDA vector and feedback analysis for short query Information Retrieval systems

Abstract

Talk to us

Similar Papers

More From: Logic Journal of the IGPL