The CNN-Corpus in Spanish

Rafael Dueire Lins,Steven J Simske,Bruno Tenorio,Rafael Ferreira,Luciano Cabral,Jamilson Batista,Diego A Salcedo,Gabriel De França Pereira E Silva,Hilario Oliveira,Rinaldo Lima

doi:10.1145/3342558.3345423

The CNN-Corpus in Spanish

Rafael Dueire Lins, Steven J Simske + Show 8 more

https://doi.org/10.1145/3342558.3345423

Copy DOI

Publication Date: Sep 23, 2019

Citations: 2

Affiliation: Hospital das Clínicas da Universidade Federal de Pernambuco, Universidade Federal Rural de Pernambuco, Colorado State University

#Spanish Language #Texts In Spanish + Show 8 more

Abstract
Full-Text
Similar Papers

Abstract

This paper details the development and features of the CNN-corpus in Spanish, possibly the largest test corpus for single document extractive text summarization in the Spanish language. Its current version encompasses 1,117 well-written texts in Spanish, each of them has an abstractive and an extractive summary. The development methodology adopted allows good-quality qualitative and quantitative assessments of summarization strategies for tools developed in the Spanish language.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.