Entity linking by focusing DBpedia candidate entities

Alex Olieman,Mostafa Dehghani,Hosein Azarbonyad,Maarten Marx,Jaap Kamps

doi:10.1145/2633211.2634353

Alex Olieman, Mostafa Dehghani + Show 3 more

Open Access

PDF Available

https://doi.org/10.1145/2633211.2634353

Copy DOI

Export

Save

Cite

Publication Date: Jan 1, 2014

Citations: 15

Affiliation: University of Amsterdam

Abstract
Full-Text PDF
Similar Papers

Abstract

Listen

Recently, Entity Linking and Retrieval turned out to be one of the most interesting tasks in Information Extraction due to its various applications. Entity Linking (EL) is the task of detecting mentioned entities in a text and linking them to the corresponding entries of a Knowledge Base. EL is traditionally composed of three major parts: i)spotting, ii)candidate generation, and iii)candidate disambiguation. The performance of an EL system is highly dependent on the accuracy of each individual part. In this paper, we on these three main building blocks of EL systems and try to improve on the results of one of the open source EL systems, namely DBpedia Spotlight. We propose to use text pre-processing and parameter tuning to focus a general-purpose EL system to perform better on different kinds of input text. Also, one of the main drawbacks of EL systems is identifying where a name does not refer to any known entity. To improve this so-called NIL-detection, we define different features using a set of texts and their known entities and design a classifier to automatically classify DBpedia Spotlight's output entities as NIL or Not NIL. The proposed system has participated in the SIGIR ERD Challenge 2014 and the performance analysis of this system on the challenge's datasets shows that the proposed approaches successfully improve the accuracy of the baseline system.

Full Text