arrow
返回

A Robust Region-Aware Framework for Audio Forgery Localization

delete2026-01-01
delete0
PRE
AI
J
Jiale Luo
H
Hongxia Wang *
H
Hanqing Liu
H
Huang, Xiao
K
Kaile Wang
DOI:10.1109/TASLPRO.2026.3661237delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
含多个精细伪造段的音频定位对音频伪造定位对策构成了显著挑战。当此类音频受到压缩引入的噪声影响,并可能包含未知模型生成的内容时,该挑战进一步加剧。为应对此挑战,我们提出了一种新颖框架,通过捕获不同伪造音频区域间的差异并针对各区域应用不同策略来增强音频伪造定位。具体而言,我们引入了一种基于聚类的专家混合(CMoE)框架,用于动态分配专业专家至伪造音频的不同区域。此外,我们引入了一种可学习掩码模块,以有效从大规模预训练模型中提取鲁棒且泛化的特征。我们还设计了一种组对比损失函数,以增强区分不同音频区域差异的能力。实验结果表明,我们的模型在泛化和鲁棒性方面表现优异,同时能够有效定位不同压缩格式和未知生成模型中的伪造内容。
Keyword:
Feature extraction
Forgery
Location awareness
Robustness
Noise
Reviews
Propagation losses
Graph neural networks
Data mining
Contrastive learning
Audio forgery localization
audio forensics
deep clustering
contrastive learning

期刊

I
IEEE Transactions on Audio Speech and Language Processing
IF:
0
论文数:
151
被引数:
0

机构

S
sichuan university
学者数:
12.1W
论文数: 7.8W
被引数: 100
引用论文

引用论文

ASVspoof 2021: Towards Spoofed and Deepfake Speech Detection in the Wild
err2023-01-01
err42
errOAAI
errLiu, Xuechen; Wang, Xin; Sahidullah, Md; Patino, Jose; Delgado, Hector; Kinnunen, Tomi; Todisco, Massimiliano; Yamagishi, Junichi; Evans, Nicholas; Nautsch, Andreas; Lee, Kong Aik
err分享
err收藏
err分享
err收藏
Robust copy-move detection and localization of digital audio based CFCC feature
err
err0
PREAI
errWang,Dongyu; Li,Xiaojie; Shi,Canghong; Niu,Xianhua; Xiong,Ling; Wu,Hanzhou; Qian,Qing; Qi,Chao
err分享
err收藏
Human detection of political speech deepfakes across transcripts, audio, and video
err2024-09-02
err3
errOAAI
errGroh, Matthew; Sankaranarayanan, Aruna; Singh, Nikhil; Kim, Dong Young; Lippman, Andrew; Picard, Rosalind
err分享
err收藏
Does Audio Deepfake Detection Generalize?
err2022-09-18
err0
errOAAI
errNicolas Müller; Pavel Czempin; Franziska Diekmann; Adam Froghyar; Konstantin Böttinger
err分享
err收藏
学者 查看更多内容