Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Tsung-Yuan Hsu

Language Representation in Multilingual BERT and its applications to improve Cross-lingual Generalization

Oct 23, 2020

Chi-Liang Liu, Tsung-Yuan Hsu, Yung-Sung Chuang, Chung-Yi Li, Hung-yi Lee

Figure 1 for Language Representation in Multilingual BERT and its applications to improve Cross-lingual Generalization

Figure 2 for Language Representation in Multilingual BERT and its applications to improve Cross-lingual Generalization

Figure 3 for Language Representation in Multilingual BERT and its applications to improve Cross-lingual Generalization

Figure 4 for Language Representation in Multilingual BERT and its applications to improve Cross-lingual Generalization

Abstract:A token embedding in multilingual BERT (m-BERT) contains both language and semantic information. We find that representation of a language can be obtained by simply averaging the embeddings of the tokens of the language. With the language representation, we can control the output languages of multilingual BERT by manipulating the token embeddings and achieve unsupervised token translation. We further propose a computationally cheap but effective approach to improve the cross-lingual ability of m-BERT based on the observation.

* preprint

Via

Access Paper or Ask Questions

What makes multilingual BERT multilingual?

Oct 20, 2020

Chi-Liang Liu, Tsung-Yuan Hsu, Yung-Sung Chuang, Hung-yi Lee

Figure 1 for What makes multilingual BERT multilingual?

Figure 2 for What makes multilingual BERT multilingual?

Figure 3 for What makes multilingual BERT multilingual?

Figure 4 for What makes multilingual BERT multilingual?

* arXiv admin note: substantial text overlap with arXiv:2004.09205

Via

Access Paper or Ask Questions

Investigation of Sentiment Controllable Chatbot

Jul 11, 2020

Hung-yi Lee, Cheng-Hao Ho, Chien-Fu Lin, Chiung-Chih Chang, Chih-Wei Lee, Yau-Shian Wang, Tsung-Yuan Hsu, Kuan-Yu Chen

Figure 1 for Investigation of Sentiment Controllable Chatbot

Figure 2 for Investigation of Sentiment Controllable Chatbot

Figure 3 for Investigation of Sentiment Controllable Chatbot

Figure 4 for Investigation of Sentiment Controllable Chatbot

Abstract:Conventional seq2seq chatbot models attempt only to find sentences with the highest probabilities conditioned on the input sequences, without considering the sentiment of the output sentences. In this paper, we investigate four models to scale or adjust the sentiment of the chatbot response: a persona-based model, reinforcement learning, a plug and play model, and CycleGAN, all based on the seq2seq model. We also develop machine-evaluated metrics to estimate whether the responses are reasonable given the input. These metrics, together with human evaluation, are used to analyze the performance of the four models in terms of different aspects; reinforcement learning and CycleGAN are shown to be very attractive.

* arXiv admin note: text overlap with arXiv:1804.02504

Via

Access Paper or Ask Questions

A Study of Cross-Lingual Ability and Language-specific Information in Multilingual BERT

Apr 20, 2020

Chi-Liang Liu, Tsung-Yuan Hsu, Yung-Sung Chuang, Hung-Yi Lee

Figure 1 for A Study of Cross-Lingual Ability and Language-specific Information in Multilingual BERT

Figure 2 for A Study of Cross-Lingual Ability and Language-specific Information in Multilingual BERT

Figure 3 for A Study of Cross-Lingual Ability and Language-specific Information in Multilingual BERT

Figure 4 for A Study of Cross-Lingual Ability and Language-specific Information in Multilingual BERT

Abstract:Recently, multilingual BERT works remarkably well on cross-lingual transfer tasks, superior to static non-contextualized word embeddings. In this work, we provide an in-depth experimental study to supplement the existing literature of cross-lingual ability. We compare the cross-lingual ability of non-contextualized and contextualized representation model with the same data. We found that datasize and context window size are crucial factors to the transferability. We also observe the language-specific information in multilingual BERT. By manipulating the latent representations, we can control the output languages of multilingual BERT, and achieve unsupervised token translation. We further show that based on the observation, there is a computationally cheap but effective approach to improve the cross-lingual ability of multilingual BERT.

Via

Access Paper or Ask Questions

Scalable Sentiment for Sequence-to-sequence Chatbot Response with Performance Analysis

Apr 07, 2018

Chih-Wei Lee, Yau-Shian Wang, Tsung-Yuan Hsu, Kuan-Yu Chen, Hung-Yi Lee, Lin-shan Lee

Figure 1 for Scalable Sentiment for Sequence-to-sequence Chatbot Response with Performance Analysis

Figure 2 for Scalable Sentiment for Sequence-to-sequence Chatbot Response with Performance Analysis

Figure 3 for Scalable Sentiment for Sequence-to-sequence Chatbot Response with Performance Analysis

Figure 4 for Scalable Sentiment for Sequence-to-sequence Chatbot Response with Performance Analysis

Abstract:Conventional seq2seq chatbot models only try to find the sentences with the highest probabilities conditioned on the input sequences, without considering the sentiment of the output sentences. Some research works trying to modify the sentiment of the output sequences were reported. In this paper, we propose five models to scale or adjust the sentiment of the chatbot response: persona-based model, reinforcement learning, plug and play model, sentiment transformation network and cycleGAN, all based on the conventional seq2seq model. We also develop two evaluation metrics to estimate if the responses are reasonable given the input. These metrics together with other two popularly used metrics were used to analyze the performance of the five proposed models on different aspects, and reinforcement learning and cycleGAN were shown to be very attractive. The evaluation metrics were also found to be well correlated with human evaluation.

Via

Access Paper or Ask Questions