Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Quality-Net: An End-to-End Non-intrusive Speech Quality Assessment Model based on BLSTM

Aug 17, 2018

Szu-Wei Fu, Yu Tsao, Hsin-Te Hwang, Hsin-Min Wang

Figure 1 for Quality-Net: An End-to-End Non-intrusive Speech Quality Assessment Model based on BLSTM

Figure 2 for Quality-Net: An End-to-End Non-intrusive Speech Quality Assessment Model based on BLSTM

Figure 3 for Quality-Net: An End-to-End Non-intrusive Speech Quality Assessment Model based on BLSTM

Figure 4 for Quality-Net: An End-to-End Non-intrusive Speech Quality Assessment Model based on BLSTM

Share this with someone who'll enjoy it:

Abstract:Nowadays, most of the objective speech quality assessment tools (e.g., perceptual evaluation of speech quality (PESQ)) are based on the comparison of the degraded/processed speech with its clean counterpart. The need of a "golden" reference considerably restricts the practicality of such assessment tools in real-world scenarios since the clean reference usually cannot be accessed. On the other hand, human beings can readily evaluate the speech quality without any reference (e.g., mean opinion score (MOS) tests), implying the existence of an objective and non-intrusive (no clean reference needed) quality assessment mechanism. In this study, we propose a novel end-to-end, non-intrusive speech quality evaluation model, termed Quality-Net, based on bidirectional long short-term memory. The evaluation of utterance-level quality in Quality-Net is based on the frame-level assessment. Frame constraints and sensible initializations of forget gate biases are applied to learn meaningful frame-level quality assessment from the utterance-level quality label. Experimental results show that Quality-Net can yield high correlation to PESQ (0.9 for the noisy speech and 0.84 for the speech processed by speech enhancement). We believe that Quality-Net has potential to be used in a wide variety of applications of speech signal processing.

* Accepted in Interspeech2018

View paper on

Share this with someone who'll enjoy it:

Title:Quality-Net: An End-to-End Non-intrusive Speech Quality Assessment Model based on BLSTM

Paper and Code