arXiv:2503. 17577v2 Announce Type: replace-cross Abstract: Deepfakes have emerged as a widespread and rapidly escalating concern in generative AI, spanning images, audio, and videos.
By Xiang Li, Pin-Yu Chen, Wenqi Wei
arXiv:2607. 25543v1 Announce Type: cross Abstract: Generative AI has rapidly expanded audio-visual forgery beyond human-centric deepfakes into general scenes.
By Jielun Peng, Yabin Wang, Yaqi Li, Jincheng Liu, Xiaopeng Hong, Athanasios V. Vasilakos
ToolDF is a tool‑integrated reasoning framework designed for detecting mixed‑authenticity audio deepfakes, where genuine and manipulated audio cues coexist across time or overlapping sources. It uses an audio large language model to orchestrate tasks such as source separation and routing to domain‑specific experts, aggregating their evidence into an interpretable verdict. The authors also introduce a mixed‑authenticity ADD benchmark and report that ToolDF outperforms monolithic baselines, achieving significant macro‑F1 gains while localizing evidence to specific temporal regions and acoustic sources.
By Taewoo Kim, Young Han Lee, Nam In Park, Chanwoo Kim
The paper introduces a training‑free proactive defense for detecting partial deepfake speech by using self‑embedding steganography. It embeds a compressed version of the clean audio within itself, allowing post‑hoc extraction of reference content and enabling detection of spoofed segments via codec‑based restoration. Experiments on a benchmark dataset show that this method complements passive detectors and operates without any training, offering a robust, data‑efficient alternative for partial deepfake detection.
By Yigitcan \"Ozer, Zhe Zhang, Wanying Ge, Xin Wang, Junichi Yamagishi
The paper introduces ROGUE, a framework that builds robust audio deepfake detection workflows by combining multiple detection tools. ROGUE treats workflow creation as a sequential decision problem and uses a dual-agent system: a perturbation agent generates audio distortions while a policy agent selects and executes detection tools that are resilient to those perturbations. Experiments on several datasets and real-world corruptions show that ROGUE consistently outperforms strong baselines in robustness and generalization, demonstrating the value of adversarially optimized workflow generation for reliable deployment.
By Xiang Li, Pin-Yu Chen, Wenqi Wei
arXiv:2606. 29544v1 Announce Type: cross Abstract: We present Proteus, a framework developed at Resemble AI for automated robustness testing of our audio deepfake detection system.
By Nicolas M. M\"uller, Aditya Tirumala Bukkapatnam, Zohaib Ahmed