Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Subjective Assessments of Legibility in Ancient Manuscript Images -- The SALAMI Dataset

Feb 19, 2021

Simon Brenner, Robert Sablatnig

Figure 1 for Subjective Assessments of Legibility in Ancient Manuscript Images -- The SALAMI Dataset

Figure 2 for Subjective Assessments of Legibility in Ancient Manuscript Images -- The SALAMI Dataset

Figure 3 for Subjective Assessments of Legibility in Ancient Manuscript Images -- The SALAMI Dataset

Figure 4 for Subjective Assessments of Legibility in Ancient Manuscript Images -- The SALAMI Dataset

Share this with someone who'll enjoy it:

Abstract:The research field concerned with the digital restoration of degraded written heritage lacks a quantitative metric for evaluating its results, which prevents the comparison of relevant methods on large datasets. Thus, we introduce a novel dataset of Subjective Assessments of Legibility in Ancient Manuscript Images (SALAMI) to serve as a ground truth for the development of quantitative evaluation metrics in the field of digital text restoration. This dataset consists of 250 images of 50 manuscript regions with corresponding spatial maps of mean legibility and uncertainty, which are based on a study conducted with 20 experts of philology and paleography. As this study is the first of its kind, the validity and reliability of its design and the results obtained are motivated statistically: we report a high intra- and inter-rater agreement and show that the bulk of variation in the scores is introduced by the images regions observed and not by controlled or uncontrolled properties of participants and test environments, thus concluding that the legibility scores measured are valid attributes of the underlying images.

* This is an extended version of a paper presented at the PatReCH 2020 (ICPR) workshop and currently in the publication process for Springer LNCS

View paper on

Share this with someone who'll enjoy it:

Title:Subjective Assessments of Legibility in Ancient Manuscript Images -- The SALAMI Dataset

Paper and Code