Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Tensorized Embedding Layers for Efficient Model Compression

Jan 30, 2019

Valentin Khrulkov, Oleksii Hrinchuk, Leyla Mirvakhabova, Ivan Oseledets

Figure 1 for Tensorized Embedding Layers for Efficient Model Compression

Figure 2 for Tensorized Embedding Layers for Efficient Model Compression

Figure 3 for Tensorized Embedding Layers for Efficient Model Compression

Figure 4 for Tensorized Embedding Layers for Efficient Model Compression

Share this with someone who'll enjoy it:

Abstract:The embedding layers transforming input words into real vectors are the key components of deep neural networks used in natural language processing. However, when the vocabulary is large (e.g., 800k unique words in the One-Billion-Word dataset), the corresponding weight matrices can be enormous, which precludes their deployment in a limited resource setting. We introduce a novel way of parametrizing embedding layers based on the Tensor Train (TT) decomposition, which allows compressing the model significantly at the cost of a negligible drop or even a slight gain in performance. Importantly, our method does not take the pre-trained model and compress its weights but rather supplants the standard embedding layers with their TT-based counterparts. The resulting model is then trained end-to-end, however, it can capitalize on larger batches due to the reduced memory requirements. We evaluate our method on a wide range of benchmarks in sentiment analysis, neural machine translation, and language modeling, and analyze the trade-off between performance and compression ratios for a wide range of architectures, from MLPs to LSTMs and Transformers.

View paper on

OpenReview

Share this with someone who'll enjoy it:

Title:Tensorized Embedding Layers for Efficient Model Compression

Paper and Code