对抗制
计算机科学
扩散
计算机安全
人工智能
热力学
物理
作者
Jiayuan Chen,Yunshu Dai,Fangjun Huang
出处
期刊:
日期:2025-03-12
卷期号:: 1-5
被引量:1
标识
DOI:10.1109/icassp49660.2025.10889191
摘要
Recently, adversarial attacks on speaker recognition systems have garnered significant interest. However, existing methods focus on injecting subtle perturbations into audio, which may compromise auditory quality. To address this problem, we propose a novel approach named DiffAttack, which employs a diffusion model for generating high-quality adversarial samples. Firstly, we extract the Mel spectrogram of the original audio. Subsequently, the Mel spectrogram is optimized to fool the speaker recognition system while preserving the high auditory quality of the attacked audio. Lastly, a conditional diffusion model is used to reconstruct the adversarial audio from the optimized Mel spectrogram. Experimental evaluations on ECAPA and ResNet, two advanced speaker recognition systems, demonstrate that our method exceeds those state-of-the-art methods in terms of attack success rate, transferability, and auditory quality.
科研通智能强力驱动
Strongly Powered by AbleSci AI