Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Breaking Down Word Semantics from Pre-trained Language Models through Layer-wise Dimension Selection

Oct 08, 2023

Nayoung Choi

Figure 1 for Breaking Down Word Semantics from Pre-trained Language Models through Layer-wise Dimension Selection

Figure 2 for Breaking Down Word Semantics from Pre-trained Language Models through Layer-wise Dimension Selection

Figure 3 for Breaking Down Word Semantics from Pre-trained Language Models through Layer-wise Dimension Selection

Figure 4 for Breaking Down Word Semantics from Pre-trained Language Models through Layer-wise Dimension Selection

Share this with someone who'll enjoy it:

Abstract:Contextual word embeddings obtained from pre-trained language model (PLM) have proven effective for various natural language processing tasks at the word level. However, interpreting the hidden aspects within embeddings, such as syntax and semantics, remains challenging. Disentangled representation learning has emerged as a promising approach, which separates specific aspects into distinct embeddings. Furthermore, different linguistic knowledge is believed to be stored in different layers of PLM. This paper aims to disentangle semantic sense from BERT by applying a binary mask to middle outputs across the layers, without updating pre-trained parameters. The disentangled embeddings are evaluated through binary classification to determine if the target word in two different sentences has the same meaning. Experiments with cased BERT$_{\texttt{base}}$ show that leveraging layer-wise information is effective and disentangling semantic sense further improve performance.

View paper on

Share this with someone who'll enjoy it:

Title:Breaking Down Word Semantics from Pre-trained Language Models through Layer-wise Dimension Selection

Paper and Code