Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Audio Deep Fake Detection System with Neural Stitching for ADD 2022

Apr 20, 2022

Rui Yan, Cheng Wen, Shuran Zhou, Tingwei Guo, Wei Zou, Xiangang Li

Figure 1 for Audio Deep Fake Detection System with Neural Stitching for ADD 2022

Figure 2 for Audio Deep Fake Detection System with Neural Stitching for ADD 2022

Figure 3 for Audio Deep Fake Detection System with Neural Stitching for ADD 2022

Figure 4 for Audio Deep Fake Detection System with Neural Stitching for ADD 2022

Share this with someone who'll enjoy it:

Abstract:This paper describes our best system and methodology for ADD 2022: The First Audio Deep Synthesis Detection Challenge\cite{Yi2022ADD}. The very same system was used for both two rounds of evaluation in Track 3.2 with a similar training methodology. The first round of Track 3.2 data is generated from Text-to-Speech(TTS) or voice conversion (VC) algorithms, while the second round of data consists of generated fake audio from other participants in Track 3.1, aiming to spoof our systems. Our systems use a standard 34-layer ResNet, with multi-head attention pooling \cite{india2019self} to learn the discriminative embedding for fake audio and spoof detection. We further utilize neural stitching to boost the model's generalization capability in order to perform equally well in different tasks, and more details will be explained in the following sessions. The experiments show that our proposed method outperforms all other systems with a 10.1% equal error rate(EER) in Track 3.2.

* Accepted to ICASSP 2022

View paper on

Share this with someone who'll enjoy it:

Title:Audio Deep Fake Detection System with Neural Stitching for ADD 2022

Paper and Code