A Flexible Supervised Term-Weighting Technique and its Application to Variable Extraction and Information Retrieval

Ana Gabriela Maguitman,Fernando Delbianco,Fernando Abel Tohmé,Mariano Maisonnave

doi:10.4114/intartif.vol22iss63pp61-80

Ana Gabriela Maguitman, Fernando Delbianco + Show 2 more

Open Access

https://doi.org/10.4114/intartif.vol22iss63pp61-80

Copy DOI

Journal: Inteligencia Artificial	Publication Date: Feb 27, 2019
Citations: 6	License type: CC BY-NC 4.0

Affiliation: Universidad Nacional del Sur

Abstract

Successful modeling and prediction depend on effective methods for the extraction of domain-relevant variables. This paper proposes a methodology for identifying domain-specific terms. The proposed methodology relies on a collection of documents labeled as relevant or irrelevant to the domain under analysis. Based on the labeled document collection, we propose a supervised technique that weights terms based on their descriptive and discriminating power. Finally, the descriptive and discriminating values are combined into a general measure that, through the use of an adjustable parameter, allows to independently favor different aspects of retrieval such as maximizing precision or recall, or achieving a balance between both of them. The proposed technique is applied to the economic domain and is empirically evaluated through a human-subject experiment involving experts and non-experts in Economy. It is also evaluated as a term-weighting technique for query-term selection showing promising results. We finally illustrate the applicability of the proposed technique to address diverse problems such as building prediction models, supporting knowledge modeling, and achieving total recall.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A Flexible Supervised Term-Weighting Technique and its Application to Variable Extraction and Information Retrieval

Abstract

Talk to us

Similar Papers

More From: Inteligencia Artificial

Lead the way for us

Similar Papers

Evaluation of clustering and summarizing in distributed latent semantic indexing
Mehdi Behshameh ... Salman Hooshmand
-
Mehdi Behshameh, et. al.Mehdi Behshameh ... Salman Hooshmand
01 Jan 2009
01 Jan 2009

Discriminative Feature Spamming Technique for Roman Urdu Sentiment Analysis
Khawar Mehmood ... Kamran Shafi
IEEE Access | VOL. 7
Khawar Mehmood, et. al.Khawar Mehmood ... Kamran Shafi
01 Jan 2019
IEEE Access | VOL. 7

Opinion Mining and Topic Categorization with Novel Term Weighting
Tatiana Gasanova ... Wolfgang Minker
-
Tatiana Gasanova, et. al.Tatiana Gasanova ... Wolfgang Minker
01 Jan 2014
01 Jan 2014

A Study of Applying Different Term Weighting Schemes on Arabic Text Classification
D S Guru ... Mahamad Suhil
-
D S Guru, et. al.D S Guru ... Mahamad Suhil
05 Nov 2018
05 Nov 2018

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A Flexible Supervised Term-Weighting Technique and its Application to Variable Extraction and Information Retrieval

Abstract

Talk to us

Similar Papers

More From: Inteligencia Artificial