Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Young-yoon Lee

Iterative Compression of End-to-End ASR Model using AutoML

Aug 06, 2020

Abhinav Mehrotra, Łukasz Dudziak, Jinsu Yeo, Young-yoon Lee, Ravichander Vipperla, Mohamed S. Abdelfattah, Sourav Bhattacharya, Samin Ishtiaq, Alberto Gil C. P. Ramos, SangJeong Lee(+2 more)

Figure 1 for Iterative Compression of End-to-End ASR Model using AutoML

Figure 2 for Iterative Compression of End-to-End ASR Model using AutoML

Figure 3 for Iterative Compression of End-to-End ASR Model using AutoML

Figure 4 for Iterative Compression of End-to-End ASR Model using AutoML

Abstract:Increasing demand for on-device Automatic Speech Recognition (ASR) systems has resulted in renewed interests in developing automatic model compression techniques. Past research have shown that AutoML-based Low Rank Factorization (LRF) technique, when applied to an end-to-end Encoder-Attention-Decoder style ASR model, can achieve a speedup of up to 3.7x, outperforming laborious manual rank-selection approaches. However, we show that current AutoML-based search techniques only work up to a certain compression level, beyond which they fail to produce compressed models with acceptable word error rates (WER). In this work, we propose an iterative AutoML-based LRF approach that achieves over 5x compression without degrading the WER, thereby advancing the state-of-the-art in ASR compression.

* INTERSPEECH 2020

Via

Access Paper or Ask Questions