Butcher, baker, or candlestick maker? Predicting occupations using predicate-argument relations

Kieran White,Richard F.E Sutcliffe

doi:10.1002/asi.21547

Abstract

In a previous question answering study, we identified nine semantic-relationship types, including synonyms, hypernyms, word chains, and holonyms, that exist between terms in Text Retrieval Conference queries and those in their supporting sentences in the Advanced Question Answering for Intelligence (Graff, 2002) corpus. The most frequently occurring relationship type was the hypernym (e.g., Katherine Hepburn is an actress). The aim of the present work, therefore, was to develop a method for determining a person's occupation from syntactic data in a text corpus. First, in the P-System, we compared predicate–argument data involving a proper name for different occupations using Okapi's BM25 weighting algorithm. When classifying actors and using sufficiently frequent names, an accuracy of 0.955 was attained. For evaluation purposes, we also implemented a standard apposition-based classifier (A-System). This performs well, but only if a particular name happens to appear in apposition with the corresponding occupation. Last, we created a hybrid (H-System) which combines the strengths of P with those of A. Using data with a minimum of 100 predicate–argument pairs, H performed best with an overall lenient accuracy of 0.750 while A and P scored 0.615 and 0.656, respectively. We therefore conclude that a hybrid approach combining information from different sources is the best way to predict occupations.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Butcher, baker, or candlestick maker? Predicting occupations using predicate-argument relations

Abstract

Talk to us

Similar Papers

More From: Journal of the American Society for Information Science and Technology

Lead the way for us

Journal: Journal of the American Society for Information Science and Technology	Publication Date: Apr 27, 2011
Citations: 1

Similar Papers

Combining Parts of Speech, Term Proximity, and Query Expansion for Document Retrieval
Eric Labouve ... Lubomir Stanchev
-
Eric Labouve, et. al.Eric Labouve ... Lubomir Stanchev
01 Jan 2019
01 Jan 2019

Passage-Based Text Summarization for Legal Information Retrieval
Ambedkar Kanapala ... Srikanth Jannu
Arabian Journal for Science and Engineering | VOL. 44
Ambedkar Kanapala, et. al.Ambedkar Kanapala ... Srikanth Jannu
18 Jul 2019
Arabian Journal for Science and Engineering | VOL. 44

Increasing the Visibility of Search using Genetic Algorithm
...
-
, et. al. ...
03 Sep 2015
03 Sep 2015

Towards a Mixed Approach to Extract Biomedical Terms from Text Corpus
Juan Antonio Lossio Ventura ... Maguelonne Teisseire
International Journal of Knowledge Discovery in Bioinformatics | VOL. 4
Juan Antonio Lossio Ventura, et. al.Juan Antonio Lossio Ventura ... Maguelonne Teisseire
01 Jan 2014
International Journal of Knowledge Discovery in Bioinformatics | VOL. 4

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Butcher, baker, or candlestick maker? Predicting occupations using predicate-argument relations

Abstract

Talk to us

Similar Papers

More From: Journal of the American Society for Information Science and Technology