Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Stefan Fiel

TU Wien - Computer Vision Lab

Few-Shot Learning by Dimensionality Reduction in Gradient Space

Jun 07, 2022

Martin Gauch, Maximilian Beck, Thomas Adler, Dmytro Kotsur, Stefan Fiel, Hamid Eghbal-zadeh, Johannes Brandstetter, Johannes Kofler, Markus Holzleitner, Werner Zellinger(+3 more)

Figure 1 for Few-Shot Learning by Dimensionality Reduction in Gradient Space

Figure 2 for Few-Shot Learning by Dimensionality Reduction in Gradient Space

Figure 3 for Few-Shot Learning by Dimensionality Reduction in Gradient Space

Figure 4 for Few-Shot Learning by Dimensionality Reduction in Gradient Space

Abstract:We introduce SubGD, a novel few-shot learning method which is based on the recent finding that stochastic gradient descent updates tend to live in a low-dimensional parameter subspace. In experimental and theoretical analyses, we show that models confined to a suitable predefined subspace generalize well for few-shot learning. A suitable subspace fulfills three criteria across the given tasks: it (a) allows to reduce the training error by gradient flow, (b) leads to models that generalize well, and (c) can be identified by stochastic gradient descent. SubGD identifies these subspaces from an eigendecomposition of the auto-correlation matrix of update directions across different tasks. Demonstrably, we can identify low-dimensional suitable subspaces for few-shot learning of dynamical systems, which have varying properties described by one or few parameters of the analytical system description. Such systems are ubiquitous among real-world applications in science and engineering. We experimentally corroborate the advantages of SubGD on three distinct dynamical systems problem settings, significantly outperforming popular few-shot learning methods both in terms of sample efficiency and performance.

* Accepted at Conference on Lifelong Learning Agents (CoLLAs) 2022. Code: https://github.com/ml-jku/subgd Blog post: https://ml-jku.github.io/subgd

Via

Access Paper or Ask Questions

READ-BAD: A New Dataset and Evaluation Scheme for Baseline Detection in Archival Documents

Dec 11, 2017

Tobias Grüning, Roger Labahn, Markus Diem, Florian Kleber, Stefan Fiel

Figure 1 for READ-BAD: A New Dataset and Evaluation Scheme for Baseline Detection in Archival Documents

Figure 2 for READ-BAD: A New Dataset and Evaluation Scheme for Baseline Detection in Archival Documents

Figure 3 for READ-BAD: A New Dataset and Evaluation Scheme for Baseline Detection in Archival Documents

Figure 4 for READ-BAD: A New Dataset and Evaluation Scheme for Baseline Detection in Archival Documents

Abstract:Text line detection is crucial for any application associated with Automatic Text Recognition or Keyword Spotting. Modern algorithms perform good on well-established datasets since they either comprise clean data or simple/homogeneous page layouts. We have collected and annotated 2036 archival document images from different locations and time periods. The dataset contains varying page layouts and degradations that challenge text line segmentation methods. Well established text line segmentation evaluation schemes such as the Detection Rate or Recognition Accuracy demand for binarized data that is annotated on a pixel level. Producing ground truth by these means is laborious and not needed to determine a method's quality. In this paper we propose a new evaluation scheme that is based on baselines. The proposed scheme has no need for binarization and it can handle skewed as well as rotated text lines. The ICDAR 2017 Competition on Baseline Detection and the ICDAR 2017 Competition on Layout Analysis for Challenging Medieval Manuscripts used this evaluation scheme. Finally, we present results achieved by a recently published text line detection algorithm.

* Submitted to DAS2018

Via

Access Paper or Ask Questions

Unsupervised Feature Learning for Writer Identification and Writer Retrieval

Aug 18, 2017

Vincent Christlein, Martin Gropp, Stefan Fiel, Andreas Maier

Figure 1 for Unsupervised Feature Learning for Writer Identification and Writer Retrieval

Figure 2 for Unsupervised Feature Learning for Writer Identification and Writer Retrieval

Figure 3 for Unsupervised Feature Learning for Writer Identification and Writer Retrieval

Figure 4 for Unsupervised Feature Learning for Writer Identification and Writer Retrieval

Abstract:Deep Convolutional Neural Networks (CNN) have shown great success in supervised classification tasks such as character classification or dating. Deep learning methods typically need a lot of annotated training data, which is not available in many scenarios. In these cases, traditional methods are often better than or equivalent to deep learning methods. In this paper, we propose a simple, yet effective, way to learn CNN activation features in an unsupervised manner. Therefore, we train a deep residual network using surrogate classes. The surrogate classes are created by clustering the training dataset, where each cluster index represents one surrogate class. The activations from the penultimate CNN layer serve as features for subsequent classification tasks. We evaluate the feature representations on two publicly available datasets. The focus lies on the ICDAR17 competition dataset on historical document writer identification (Historical-WI). We show that the activation features trained without supervision are superior to descriptors of state-of-the-art writer identification methods. Additionally, we achieve comparable results in the case of handwriting classification using the ICFHR16 competition dataset on historical Latin script types (CLaMM16).

* ICDAR2017 camera ready (fixed p@2 values, missing table references)

Via

Access Paper or Ask Questions