Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time

Nov 04, 2020

Sashi Novitasari, Andros Tjandra, Tomoya Yanagita, Sakriani Sakti, Satoshi Nakamura

Figure 1 for Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time

Figure 2 for Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time

Figure 3 for Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time

Share this with someone who'll enjoy it:

Abstract:Inspired by a human speech chain mechanism, a machine speech chain framework based on deep learning was recently proposed for the semi-supervised development of automatic speech recognition (ASR) and text-to-speech synthesis TTS) systems. However, the mechanism to listen while speaking can be done only after receiving entire input sequences. Thus, there is a significant delay when encountering long utterances. By contrast, humans can listen to what hey speak in real-time, and if there is a delay in hearing, they won't be able to continue speaking. In this work, we propose an incremental machine speech chain towards enabling machine to listen while speaking in real-time. Specifically, we construct incremental ASR (ISR) and incremental TTS (ITTS) by letting both systems improve together through a short-term loop. Our experimental results reveal that our proposed framework is able to reduce delays due to long utterances while keeping a comparable performance to the non-incremental basic machine speech chain.

* Accepted in INTERSPEECH 2020

View paper on

Share this with someone who'll enjoy it:

Title:Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time

Paper and Code