Graph Ranking on Maximal Frequent Sequences for Single Extractive Text Summarization

Yulia Ledeneva,René Arnulfo García-Hernández,Alexander Gelbukh

doi:10.1007/978-3-642-54903-8_39

Yulia Ledeneva, René Arnulfo García-Hernández + Show 1 more

Open Access

https://doi.org/10.1007/978-3-642-54903-8_39

Copy DOI

Abstract

AbstractWe suggest a new method for the task of extractive text summarization using graph-based ranking algorithms. The main idea of this paper is to rank Maximal Frequent Sequences (MFS) in order to identify the most important information in a text. MFS are considered as nodes of a graph in term selection step, and then are ranked in term weighting step using a graph-based algorithm. We show that the proposed method produces results superior to the-state-of-the-art methods; in addition, the best sentences were found with this method. We prove that MFS are better than other terms. Moreover, we show that the longer is MFS, the better are the results. If the stop-words are excluded, we lose the sense of MFS, and the results are worse. Other important aspect of this method is that it does not require deep linguistic knowledge, nor domain or language specific annotated corpora, which makes it highly portable to other domains, genres, and languages.KeywordsTerm SelectionTerm WeightingWord Sense DisambiguationText SummarizationDocument SummarizationThese keywords were added by machine and not by the authors. This process is experimental and the keywords may be updated as the learning algorithm improves.

Full Text