Stylometric similarity in literary corpora: Non-authorship clustering andDeutscher Novellenschatz

Simon Päpcke,Katharina Herget,Ulrik Brandes,Anastasia Glawion,Thomas Weitin

doi:10.1093/llc/fqac039

Simon Päpcke, Katharina Herget + Show 3 more

Open Access

https://doi.org/10.1093/llc/fqac039

Copy DOI

Abstract

AbstractA distant-reading task in literary corpus analysis is to group stylometrically similar texts. Since there are many ways to define writing style, the result not only depends on the clustering method but even more so on the measure of similarity. With authorship attribution, the predominant application of stylometry, as its benchmark much research has addressed the utility of methods for measuring similarity. We use a corpus of German-language novellas to demonstrate that one may be interested in very different meaningful groups of texts simultaneously, and that these can be recovered from stylometric clustering if the measure is chosen accordingly. As can be expected, different measures do better at recovering groups associated with, for instance, subgenre, author gender, or narrative perspective. As a consequence, it is suggested that corpus analyses should not be based on what is currently considered the most refined measure of stylometric similarity, but rather break down the decisions that yield a specific measure and provide substantively justified arguments for them.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Digital Scholarship in the Humanities	Publication Date: Aug 9, 2022
Citations: 2	License type: CC BY 4.0

R Discovery Prime

R Discovery Prime

Stylometric similarity in literary corpora: Non-authorship clustering andDeutscher Novellenschatz

Abstract

Talk to us

Similar Papers

More From: Digital Scholarship in the Humanities

Lead the way for us

Similar Papers

군집분석 방법들을 비교하기 위한 상사그림
Dae-Heung Jang
Korean Journal of Applied Statistics | VOL. 26
Dae-Heung JangDae-Heung Jang
30 Apr 2013
Korean Journal of Applied Statistics | VOL. 26

Similarity
Samer Hassan ... Rada Mihalcea
-
Samer Hassan, et. al.Samer Hassan ... Rada Mihalcea
06 Mar 2017
06 Mar 2017

Measuring Countriess Human Rights Positions in UN Universal Periodic Review
Ehsaneddin Asgari ... Ali Sanaei
SSRN Electronic Journal | VOL. -
Ehsaneddin Asgari, et. al.Ehsaneddin Asgari ... Ali Sanaei
01 Sep 2017
SSRN Electronic Journal | VOL. -

2値変量に基づく教師無し分類における類似係数の選択
Minoru Ishida ... Hiroe Tsubaki
Kodo Keiryogaku (The Japanese Journal of Behaviormetrics) | VOL. 38
Minoru Ishida, et. al.Minoru Ishida ... Hiroe Tsubaki
01 Jan 2010
Kodo Keiryogaku (The Japanese Journal of Behaviormetrics) | VOL. 38

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Stylometric similarity in literary corpora: Non-authorship clustering andDeutscher Novellenschatz

Abstract

Talk to us

Similar Papers

More From: Digital Scholarship in the Humanities