Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:On the State of the Art in Authorship Attribution and Authorship Verification

Sep 14, 2022

Jacob Tyo, Bhuwan Dhingra, Zachary C. Lipton

Figure 1 for On the State of the Art in Authorship Attribution and Authorship Verification

Figure 2 for On the State of the Art in Authorship Attribution and Authorship Verification

Figure 3 for On the State of the Art in Authorship Attribution and Authorship Verification

Figure 4 for On the State of the Art in Authorship Attribution and Authorship Verification

Share this with someone who'll enjoy it:

Abstract:Despite decades of research on authorship attribution (AA) and authorship verification (AV), inconsistent dataset splits/filtering and mismatched evaluation methods make it difficult to assess the state of the art. In this paper, we present a survey of the fields, resolve points of confusion, introduce Valla that standardizes and benchmarks AA/AV datasets and metrics, provide a large-scale empirical evaluation, and provide apples-to-apples comparisons between existing methods. We evaluate eight promising methods on fifteen datasets (including distribution-shifted challenge sets) and introduce a new large-scale dataset based on texts archived by Project Gutenberg. Surprisingly, we find that a traditional Ngram-based model performs best on 5 (of 7) AA tasks, achieving an average macro-accuracy of $76.50\%$ (compared to $66.71\%$ for a BERT-based model). However, on the two AA datasets with the greatest number of words per author, as well as on the AV datasets, BERT-based models perform best. While AV methods are easily applied to AA, they are seldom included as baselines in AA papers. We show that through the application of hard-negative mining, AV methods are competitive alternatives to AA methods. Valla and all experiment code can be found here: https://github.com/JacobTyo/Valla

View paper on

Share this with someone who'll enjoy it:

Title:On the State of the Art in Authorship Attribution and Authorship Verification

Paper and Code