flâneur — a map of the web's best reading

[2209.14098] Deepfake audio detection by speaker verification

arxiv.org · 645 words · saved by 1 readers

Thanks to recent advances in deep learning, sophisticated generation tools exist, nowadays, that produce extremely realistic synthetic speech. However, malicious uses of such tools are possible and likely, posing a serious threat to our society. Hence, synthetic voice detection has become a pressing research topic, and a large variety of detection methods have been recently proposed. Unfortunately, they hardly generalize to synthetic audios generated by tools never seen in the training phase, which makes them unfit to face real-world scenarios. In this work, we aim at overcoming this issue by proposing a new detection approach that leverages only the biometric characteristics of the speaker, with no reference to specific manipulations. Since the detector is trained only on real data, generalization is automatically ensured. The proposed approach can be implemented based on off-the-shelf speaker verification tools. We test several such solutions on three popular test sets, obtaining good performance, high generalization ability, and high robustness to audio impairment.

[2209.14098] Deepfake audio detection by speaker verification Skip to main content arXiv is now an independent nonprofit! Learn more × Search arXiv Press Enter to search · Advanced search --> Computer Science > Sound arXiv:2209.14098 (cs) [Submitted on 28 Sep 2022] Title: Deepfake audio detection by speaker verification Authors: Alessandro Pianese , Davide Cozzolino , Giovanni Poggi , Luisa Verdoliva View a PDF of the paper titled Deepfake audio detection by speaker verification, by Alessandro Pianese and Davide Cozzolino and Giovanni Poggi and Luisa Verdoliva View PDF Abstract: T

Explore this link on the map →

related reading