Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:BiFSMNv2: Pushing Binary Neural Networks for Keyword Spotting to Real-Network Performance

Nov 13, 2022

Haotong Qin, Xudong Ma, Yifu Ding, Xiaoyang Li, Yang Zhang, Zejun Ma, Jiakai Wang, Jie Luo, Xianglong Liu

Figure 1 for BiFSMNv2: Pushing Binary Neural Networks for Keyword Spotting to Real-Network Performance

Figure 2 for BiFSMNv2: Pushing Binary Neural Networks for Keyword Spotting to Real-Network Performance

Figure 3 for BiFSMNv2: Pushing Binary Neural Networks for Keyword Spotting to Real-Network Performance

Figure 4 for BiFSMNv2: Pushing Binary Neural Networks for Keyword Spotting to Real-Network Performance

Share this with someone who'll enjoy it:

Abstract:Deep neural networks, such as the Deep-FSMN, have been widely studied for keyword spotting (KWS) applications while suffering expensive computation and storage. Therefore, network compression technologies like binarization are studied to deploy KWS models on edge. In this paper, we present a strong yet efficient binary neural network for KWS, namely BiFSMNv2, pushing it to the real-network accuracy performance. First, we present a Dual-scale Thinnable 1-bit-Architecture to recover the representation capability of the binarized computation units by dual-scale activation binarization and liberate the speedup potential from an overall architecture perspective. Second, we also construct a Frequency Independent Distillation scheme for KWS binarization-aware training, which distills the high and low-frequency components independently to mitigate the information mismatch between full-precision and binarized representations. Moreover, we implement BiFSMNv2 on ARMv8 real-world hardware with a novel Fast Bitwise Computation Kernel, which is proposed to fully utilize registers and increase instruction throughput. Comprehensive experiments show our BiFSMNv2 outperforms existing binary networks for KWS by convincing margins across different datasets and even achieves comparable accuracy with the full-precision networks (e.g., only 1.59% drop on Speech Commands V1-12). We highlight that benefiting from the compact architecture and optimized hardware kernel, BiFSMNv2 can achieve an impressive 25.1x speedup and 20.2x storage-saving on edge hardware.

* arXiv admin note: text overlap with arXiv:2202.06483

View paper on

Share this with someone who'll enjoy it:

Title:BiFSMNv2: Pushing Binary Neural Networks for Keyword Spotting to Real-Network Performance

Paper and Code