Visual extraction of information from web pages

Giuseppe Della Penna,Sergio Orefice,Daniele Magazzeni

doi:10.1016/j.jvlc.2009.06.001

Visual extraction of information from web pages

Giuseppe Della Penna, Sergio Orefice + Show 1 more

https://doi.org/10.1016/j.jvlc.2009.06.001

Copy DOI

Export

Save

Cite

Journal: Journal of Visual Languages & Computing	Publication Date: Jun 27, 2009
Citations: 17

#Web Page Structure #Graphical Attributes #Spatial Relations #Visual Appearance #Web Page #Semantic Attributes #Query Tool #Spatial Query #Graphical System #Graphical Software

Abstract
Full-Text
Similar Papers

Abstract

Listen

In this paper we present a graphical software system that provides an automatic support to the extraction of information from web pages. The underlying extraction technique exploits the visual appearance of the information in the document, and is driven by the spatial relations occurring among the elements in the page. However, the usual information extraction modalities based on the web page structure can be used in our framework, too. The technique has been integrated within the Spatial Relation Query (SRQ) tool. The tool is provided with a graphical front-end which allows one to define and manage a library of spatial relations, and to use a SQL-like language for composing queries driven by these relations and by further semantic and graphical attributes.

Full Text

Published Version

Check institute access

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: Journal of Visual Languages & Computing

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.

R Discovery Prime

R Discovery Prime

Visual extraction of information from web pages