Geotagging with local lexicons to build indexes for textually-specified spatial data

Michael D Lieberman,Hanan Samet,Jagan Sankaranarayanan

doi:10.1109/icde.2010.5447903

Abstract

The successful execution of location-based and feature-based queries on spatial databases requires the construction of spatial indexes on the spatial attributes. This is not simple when the data is unstructured as is the case when the data is a collection of documents such as news articles, which is the domain of discourse, where the spatial attribute consists of text that can be (but is not required to be) interpreted as the names of locations. In other words, spatial data is specified using text (known as a toponym) instead of geometry, which means that there is some ambiguity involved. The process of identifying and disambiguating references to geographic locations is known as geotagging and involves using a combination of internal document structure and external knowledge, including a document-independent model of the audience's vocabulary of geographic locations, termed its spatial lexicon. In contrast to previous work, a new spatial lexicon model is presented that distinguishes between a global lexicon of locations known to all audiences, and an audience-specific local lexicon. Generic methods for inferring audiences' local lexicons are described. Evaluations of this inference method and the overall geotagging procedure indicate that establishing local lexicons cannot be overlooked, especially given the increasing prevalence of highly local data sources on the Internet, and will enable the construction of more accurate spatial indexes.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Geotagging with local lexicons to build indexes for textually-specified spatial data

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Determining the spatial reader scopes of news sources using local lexicons
Gianluca Quercini ... Jagan Sankaranarayanan
-
Gianluca Quercini, et. al.Gianluca Quercini ... Jagan Sankaranarayanan
02 Nov 2010
02 Nov 2010

Uncovering the spatial relatedness in Wikipedia
Gianluca Quercini ... Hanan Samet
-
Gianluca Quercini, et. al.Gianluca Quercini ... Hanan Samet
04 Nov 2014
04 Nov 2014

Spatio-textual spreadsheets
Michael D Lieberman ... Jon Sperling
-
Michael D Lieberman, et. al.Michael D Lieberman ... Jon Sperling
04 Nov 2009
04 Nov 2009

Analysis of syntactic and semantic features for fine-grained event-spatial understanding in outbreak news reports
Hutchatai Chanlekha ... Nigel Collier
Journal of Biomedical Semantics | VOL. 1
Hutchatai Chanlekha, et. al.Hutchatai Chanlekha ... Nigel Collier
01 Jan 2009
Journal of Biomedical Semantics | VOL. 1

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Geotagging with local lexicons to build indexes for textually-specified spatial data

Abstract

Talk to us

Similar Papers