Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:A Study of Different Ways to Use The Conformer Model For Spoken Language Understanding

Apr 08, 2022

Nick J. C. Wang, Shaojun Wang, Jing Xiao

Figure 1 for A Study of Different Ways to Use The Conformer Model For Spoken Language Understanding

Figure 2 for A Study of Different Ways to Use The Conformer Model For Spoken Language Understanding

Figure 3 for A Study of Different Ways to Use The Conformer Model For Spoken Language Understanding

Figure 4 for A Study of Different Ways to Use The Conformer Model For Spoken Language Understanding

Share this with someone who'll enjoy it:

Abstract:SLU combines ASR and NLU capabilities to accomplish speech-to-intent understanding. In this paper, we compare different ways to combine ASR and NLU, in particular using a single Conformer model with different ways to use its components, to better understand the strengths and weaknesses of each approach. We find that it is not necessarily a choice between two-stage decoding and end-to-end systems which determines the best system for research or application. System optimization still entails carefully improving the performance of each component. It is difficult to prove that one direction is conclusively better than the other. In this paper, we also propose a novel connectionist temporal summarization (CTS) method to reduce the length of acoustic encoding sequences while improving the accuracy and processing speed of end-to-end models. This method achieves the same intent accuracy as the best two-stage SLU recognition with complicated and time-consuming decoding but does so at lower computational cost. This stacked end-to-end SLU system yields an intent accuracy of 93.97% for the SmartLights far-field set, 95.18% for the close-field set, and 99.71% for FluentSpeech.

* Submitted to INTERSPEECH 2022. (5 pages, 1 figure.)

View paper on

Share this with someone who'll enjoy it:

Title:A Study of Different Ways to Use The Conformer Model For Spoken Language Understanding

Paper and Code