Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Semi-supervised multi-channel speaker diarization with cross-channel attention

Jul 17, 2023

Shilong Wu, Jun Du, Maokui He, Shutong Niu, Hang Chen, Haitao Tang, Chin-Hui Lee

Figure 1 for Semi-supervised multi-channel speaker diarization with cross-channel attention

Figure 2 for Semi-supervised multi-channel speaker diarization with cross-channel attention

Figure 3 for Semi-supervised multi-channel speaker diarization with cross-channel attention

Figure 4 for Semi-supervised multi-channel speaker diarization with cross-channel attention

Share this with someone who'll enjoy it:

Abstract:Most neural speaker diarization systems rely on sufficient manual training data labels, which are hard to collect under real-world scenarios. This paper proposes a semi-supervised speaker diarization system to utilize large-scale multi-channel training data by generating pseudo-labels for unlabeled data. Furthermore, we introduce cross-channel attention into the Neural Speaker Diarization Using Memory-Aware Multi-Speaker Embedding (NSD-MA-MSE) to learn channel contextual information of speaker embeddings better. Experimental results on the CHiME-7 Mixer6 dataset which only contains partial speakers' labels of the training set, show that our system achieved 57.01% relative DER reduction compared to the clustering-based model on the development set. We further conducted experiments on the CHiME-6 dataset to simulate the scenario of missing partial training set labels. When using 80% and 50% labeled training data, our system performs comparably to the results obtained using 100% labeled data for training.

* 8 pages,3 figures

View paper on

Share this with someone who'll enjoy it:

Title:Semi-supervised multi-channel speaker diarization with cross-channel attention

Paper and Code