Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Finding Universal Grammatical Relations in Multilingual BERT

May 20, 2020

Ethan A. Chi, John Hewitt, Christopher D. Manning

Figure 1 for Finding Universal Grammatical Relations in Multilingual BERT

Figure 2 for Finding Universal Grammatical Relations in Multilingual BERT

Figure 3 for Finding Universal Grammatical Relations in Multilingual BERT

Figure 4 for Finding Universal Grammatical Relations in Multilingual BERT

Share this with someone who'll enjoy it:

Abstract:Recent work has found evidence that Multilingual BERT (mBERT), a transformer-based multilingual masked language model, is capable of zero-shot cross-lingual transfer, suggesting that some aspects of its representations are shared cross-lingually. To better understand this overlap, we extend recent work on finding syntactic trees in neural networks' internal representations to the multilingual setting. We show that subspaces of mBERT representations recover syntactic tree distances in languages other than English, and that these subspaces are approximately shared across languages. Motivated by these results, we present an unsupervised analysis method that provides evidence mBERT learns representations of syntactic dependency labels, in the form of clusters which largely agree with the Universal Dependencies taxonomy. This evidence suggests that even without explicit supervision, multilingual masked language models learn certain linguistic universals.

* To appear in ACL 2020; Farsi typo corrected

View paper on

Share this with someone who'll enjoy it:

Title:Finding Universal Grammatical Relations in Multilingual BERT

Paper and Code