Automated detection and segmentation of table of contents page and index pages from document images

S Mandal,A.K Das,S.P Chowdhury,B Chanda

doi:10.1109/iciap.2003.1234052

Automated detection and segmentation of table of contents page and index pages from document images

S Mandal, A.K Das + Show 2 more

https://doi.org/10.1109/iciap.2003.1234052

Copy DOI

Publication Date: Jan 1, 2003

Citations: 17

Affiliation: Indian Institute of Engineering Science and Technology, Shibpur, Indian Statistical Institute

#Index Pages #Table Of Contents Pages + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

The requirement of identifying and segmenting the table of contents (TOC) and index pages in the development of a digital library is obvious. A digital document library is created to provide a non-labour intensive, cheap and flexible way of storing, representing and managing paper documents in electronic form to facilitate indexing, viewing, printing and extracting the intended portions. Information from the TOC and index pages is extracted to use in a document database for effective retrieval of the required pieces of information. We present fully automatic identification and segmentation of TOC and index pages from a scanned document.

Full Text