返回
Audio Deepfake Approaches
DOI:10.1109/ACCESS.2023.3333866.png)
摘要
En 中文
This paper presents a review of techniques involved in the creation and detection of audio deepfakes, the first section provides information about general deep fakes. In the second section, the main methods for audio deepfakes are outlined and subsequently compared. The results discuss various methods for detecting audio deepfakes, including analyzing statistical properties, examining media consistency, and utilizing machine learning and deep learning algorithms. Major methods used to detect fake audio in these studies included Support Vector Machines (SVMs), Decision Trees (DTs), Convolutional Neural Networks (CNNs), Siamese CNNs, Deep Neural Networks (DNNs), and a combination of CNNs and Recurrent Neural Networks (RNNs). The accuracy of these methods varied, with the highest accuracy being 99% for SVM and the lowest being 73.33% for DT. The Equal Error Rate (EER) was reported in a few of the studies, with the lowest being 2% for Deep-Sonar and the highest being 12.24 for DNN-HLLs. The t-DCF was also reported in some of the studies, with the Siamese CNN performing the best with a 55% improvement in min-t-DCF and EER compared to other methods.
Keyword:
Deepfakes
artificial intelligence
deep learning
audio deepfakes
forensics
datasets
survey
期刊
IF:
3.6
论文数:
9.8W
被引数:
29.4W
机构
引用论文
Constructing interface engineering and tailoring a nanoflower-like FeP/CoP heterostructure for enhanced oxygen evolution reaction
RSC Advances
IF0

