Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:On the Evaluation of Contextual Embeddings for Zero-Shot Cross-Lingual Transfer Learning

Apr 30, 2020

Phillip Keung, Yichao Lu, Julian Salazar, Vikas Bhardwaj

Figure 1 for On the Evaluation of Contextual Embeddings for Zero-Shot Cross-Lingual Transfer Learning

Figure 2 for On the Evaluation of Contextual Embeddings for Zero-Shot Cross-Lingual Transfer Learning

Figure 3 for On the Evaluation of Contextual Embeddings for Zero-Shot Cross-Lingual Transfer Learning

Figure 4 for On the Evaluation of Contextual Embeddings for Zero-Shot Cross-Lingual Transfer Learning

Share this with someone who'll enjoy it:

Abstract:Pre-trained multilingual contextual embeddings have demonstrated state-of-the-art performance in zero-shot cross-lingual transfer learning, where multilingual BERT is fine-tuned on some source language (typically English) and evaluated on a different target language. However, published results for baseline mBERT zero-shot accuracy vary as much as 17 points on the MLDoc classification task across four papers. We show that the standard practice of using English dev accuracy for model selection in the zero-shot setting makes it difficult to obtain reproducible results on the MLDoc and XNLI tasks. English dev accuracy is often uncorrelated (or even anti-correlated) with target language accuracy, and zero-shot cross-lingual performance varies greatly within the same fine-tuning run and between different fine-tuning runs. We recommend providing oracle scores alongside the zero-shot results: still fine-tune using English, but choose a checkpoint with the target dev set. Reporting this upper bound makes results more consistent by avoiding the variation from bad checkpoints.

View paper on

Share this with someone who'll enjoy it:

Title:On the Evaluation of Contextual Embeddings for Zero-Shot Cross-Lingual Transfer Learning

Paper and Code