DGS-MLDG: Domain Gradient Surgery Guided Meta-Learning for Domain Generalization in Speech Deepfake Detection

📅 2026-09-27
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the generalization bottlenecks in audio deepfake detection caused by domain shift and gradient conflicts inherent in meta-learning. To overcome these challenges, we propose Domain Gradient Surgery (DGS). The core innovation lies in introducing an asymmetric gradient projection strategy that eliminates conflicts between meta-training and meta-testing phases, alongside a layer-wise dynamic intervention mechanism (LW-DGS) designed to achieve conflict-free optimization trajectories. Experimental results demonstrate that the proposed method reduces the average relative Equal Error Rate (EER) by 5.29% and 4.04%, respectively, across standard benchmarks. These improvements significantly enhance detection robustness in cross-domain scenarios, establishing DGS as an effective solution for mitigating gradient interference during meta-optimization in audio spoofing detection tasks.
📝 Abstract
Speech deepfake detection faces significant challenges due to domain shifts. Domain generalization (DG), particularly meta-learning for domain generalization (MLDG), offers a promising solution by simulating and mitigating domain shifts. However, MLDG is often hindered by conflicting gradients between its meta-train and meta-test objectives, leading to suboptimal performance. To address this problem, we propose domain gradient surgery (DGS), a meta-learning method that resolves conflicts through an asymmetric projection strategy. DGS removes the destructive component from the meta-test gradient, ensuring a conflict-free optimization trajectory versus the meta-train gradient. Furthermore, we introduce layer-wise DGS (LW-DGS), an efficient variant of DGS that dynamically identifies and intervenes only conflict-prone layers. Extensive experiments on challenging benchmarks demonstrate that DGS-MLDG and LW-DGS-MLDG achieve an average relative EER reduction of 5.29% and 4.04%, respectively.
Problem

Research questions and friction points this paper is trying to address.

Speech Deepfake Detection
Domain Generalization
Meta-Learning
Gradient Conflict
Domain Shift
Innovation

Methods, ideas, or system contributions that make the work stand out.

Domain Gradient Surgery
Meta-Learning for Domain Generalization
Speech Deepfake Detection
Asymmetric Projection
Layer-wise Intervention
🔎 Similar Papers
S
Siqing Qin
Dept. of Electrical and Electronic Engineering, The Hong Kong Polytechnic University
Kong Aik Lee
Kong Aik Lee
The Hong Kong Polytechnic University, Hong Kong
Speaker and Spoken Language RecognitionSpeech ProcessingDigital Signal ProcessingSubband
Y
Youzhi Tu
Dept. of Electrical and Electronic Engineering, The Hong Kong Polytechnic University
E
Eng Siong Chng
College of Computing and Data Science, Nanyang Technological University
M
Man-Wai Mak
Dept. of Electrical and Electronic Engineering, The Hong Kong Polytechnic University