A new context‐dependent term weight computed by boost and discount using relevance information

E.K.F Dang,R.W.P Luk,S.C.F Chan,K.F.L Chung,K.S Ho,J Allan,D.L Lee

doi:10.1002/asi.21425

Abstract

AbstractWe studied the effectiveness of a new class of context‐dependent term weights for information retrieval. Unlike the traditional term frequency–inverse document frequency (TF–IDF), the new weighting of a term t in a document d depends not only on the occurrence statistics of t alone but also on the terms found within a text window (or “document‐context”) centered on t. We introduce a Boost and Discount (B&D) procedure which utilizes partial relevance information to compute the context‐dependent term weights of query terms according to a logistic regression model. We investigate the effectiveness of the new term weights compared with the context‐independent BM25 weights in the setting of relevance feedback. We performed experiments with title queries of the TREC‐6, ‐7, ‐8, and 2005 collections, comparing the residual Mean Average Precision (MAP) measures obtained using B&D term weights and those obtained by a baseline using BM25 weights. Given either 10 or 20 relevance judgments of the top retrieved documents, using the new term weights yields improvement over the baseline for all collections tested. The MAP obtained with the new weights has relative improvement over the baseline by 3.3 to 15.2%, with statistical significance at the 95% confidence level across all four collections.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A new context‐dependent term weight computed by boost and discount using relevance information

Abstract

Talk to us

Similar Papers

More From: Journal of the American Society for Information Science and Technology

Lead the way for us

Journal: Journal of the American Society for Information Science and Technology	Publication Date: Dec 1, 2010
Citations: 8

Similar Papers

(1+1)-Evolutionary Gradient Strategy to Evolve Global Term Weights in Information Retrieval
Osman Ali Sadek Ibrahim ... Dario Landa-Silva
-
Osman Ali Sadek Ibrahim, et. al.Osman Ali Sadek Ibrahim ... Dario Landa-Silva
07 Sep 2016
07 Sep 2016

Improved TFIDF in big news retrieval: An empirical study
Chien-Hsing Chen
Pattern Recognition Letters | VOL. 93
Chien-Hsing ChenChien-Hsing Chen
09 Nov 2016
Pattern Recognition Letters | VOL. 93

TAG term weight-based N gram Thesaurus generation for query expansion in information retrieval application
S.G Shaila ... A Vadivel
Journal of Information Science | VOL. 41
S.G Shaila, et. al.S.G Shaila ... A Vadivel
27 Apr 2015
Journal of Information Science | VOL. 41

A new neutrosophic TF-IDF term weighting for text mining tasks: text classification use case
Mariem Bounabi ... Karim Elmoutaouakil
International Journal of Web Information Systems | VOL. 17
Mariem Bounabi, et. al.Mariem Bounabi ... Karim Elmoutaouakil
08 Apr 2021
International Journal of Web Information Systems | VOL. 17

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A new context‐dependent term weight computed by boost and discount using relevance information

Abstract

Talk to us

Similar Papers

More From: Journal of the American Society for Information Science and Technology