Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Divide and Rule: Training Context-Aware Multi-Encoder Translation Models with Little Resources

Mar 31, 2021

Lorenzo Lupo, Marco Dinarelli, Laurent Besacier

Figure 1 for Divide and Rule: Training Context-Aware Multi-Encoder Translation Models with Little Resources

Figure 2 for Divide and Rule: Training Context-Aware Multi-Encoder Translation Models with Little Resources

Figure 3 for Divide and Rule: Training Context-Aware Multi-Encoder Translation Models with Little Resources

Figure 4 for Divide and Rule: Training Context-Aware Multi-Encoder Translation Models with Little Resources

Share this with someone who'll enjoy it:

Abstract:Multi-encoder models are a broad family of context-aware Neural Machine Translation (NMT) systems that aim to improve translation quality by encoding document-level contextual information alongside the current sentence. The context encoding is undertaken by contextual parameters, trained on document-level data. In this work, we show that training these parameters takes large amount of data, since the contextual training signal is sparse. We propose an efficient alternative, based on splitting sentence pairs, that allows to enrich the training signal of a set of parallel sentences by breaking intra-sentential syntactic links, and thus frequently pushing the model to search the context for disambiguating clues. We evaluate our approach with BLEU and contrastive test sets, showing that it allows multi-encoder models to achieve comparable performances to a setting where they are trained with $\times10$ document-level data. We also show that our approach is a viable option to context-aware NMT for language pairs with zero document-level parallel data.

View paper on

Share this with someone who'll enjoy it:

Title:Divide and Rule: Training Context-Aware Multi-Encoder Translation Models with Little Resources

Paper and Code