Tensor Space Models for Authorship Identification

Spyridon Plakias,Efstathios Stamatatos

doi:10.1007/978-3-540-87881-0_22

Abstract

Authorship identification can be viewed as a text categorization task. However, in this task the most frequent features appear to be the most important discriminators, there is usually a shortage of training texts, and the training texts are rarely evenly distributed over the authors. To cope with these problems, we propose tensors of second order for representing the stylistic properties of texts. Our approach requires the calculation of much fewer parameters in comparison to the traditional vector space representation. We examine various methods for building appropriate tensors taking into account that similar features should be placed in the same neighborhood. Based on an existing generalization of SVM able to handle tensors we perform experiments on corpora controlled for genre and topic and show that the proposed approach can effectively handle cases where only limited training texts are available.KeywordsAuthorship identificationTensor space representationText categorization

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Tensor Space Models for Authorship Identification

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Author identification: Using text sampling to handle the class imbalance problem
Efstathios Stamatatos
Information Processing & Management | VOL. 44
Efstathios StamatatosEfstathios Stamatatos
09 Jul 2007
Information Processing & Management | VOL. 44

A sequential addition and migration method for generating microstructures of short fibers with prescribed length distribution
Alok Mehta ... Matti Schneider
Computational Mechanics | VOL. 70
Alok Mehta, et. al.Alok Mehta ... Matti Schneider
03 Jul 2022
Computational Mechanics | VOL. 70

An anisotropic brittle damage model with a damage tensor of second order using a micromorphic approach
Marek Fassin ... Stefanie Reese
PAMM | VOL. 19
Marek Fassin, et. al.Marek Fassin ... Stefanie Reese
01 Nov 2019
PAMM | VOL. 19

Sketch of the mesoscopic description of nematic liquid crystals
W Muschik ... H Ehrentraut
Journal of Non-Newtonian Fluid Mechanics | VOL. 119
W Muschik, et. al.W Muschik ... H Ehrentraut
01 May 2004
Journal of Non-Newtonian Fluid Mechanics | VOL. 119

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Tensor Space Models for Authorship Identification

Abstract

Talk to us

Similar Papers