Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Fabienne Cap

An Evaluation of Neural Machine Translation Models on Historical Spelling Normalization

Aug 04, 2018

Gongbo Tang, Fabienne Cap, Eva Pettersson, Joakim Nivre

Figure 1 for An Evaluation of Neural Machine Translation Models on Historical Spelling Normalization

Figure 2 for An Evaluation of Neural Machine Translation Models on Historical Spelling Normalization

Figure 3 for An Evaluation of Neural Machine Translation Models on Historical Spelling Normalization

Figure 4 for An Evaluation of Neural Machine Translation Models on Historical Spelling Normalization

Abstract:In this paper, we apply different NMT models to the problem of historical spelling normalization for five languages: English, German, Hungarian, Icelandic, and Swedish. The NMT models are at different levels, have different attention mechanisms, and different neural network architectures. Our results show that NMT models are much better than SMT models in terms of character error rate. The vanilla RNNs are competitive to GRUs/LSTMs in historical spelling normalization. Transformer models perform better only when provided with more training data. We also find that subword-level models with a small subword vocabulary are better than character-level models for low-resource languages. In addition, we propose a hybrid method which further improves the performance of historical spelling normalization.

* 12 pages, accepted by COLING 2018, added subword-level Transformer models

Via

Access Paper or Ask Questions