最后更新时间: 2026-08-27
| 发布日期 | 文章名 | 第一作者 | 链接 |
|---|---|---|---|
| 2026-08-26 | A Training-Free Proactive Defense Against Partial Speech Manipulation via Self-Embedding Steganography | Yigitcan Özer | arXiv |
| 2026-08-25 | Investigating voiced and unvoiced regions of speech for audio deepfake detection | Ganesh Sivaraman | arXiv |
| 2026-08-25 | On the Robustness of Audio Deepfake Detection under Audio Watermarking | Zi Qian Yong | arXiv |
| 2026-08-24 | AT-ADD: A Benchmark and Challenge for Robust and All-Type Audio Deepfake Detection | Yuankun Xie | arXiv |
| 2026-08-20 | Tracking the Trend in How Speech Synthesizers Deceive People | Milan Šalko | arXiv |
| 2026-08-18 | The Last Mile of Deepfake Speech Detection: An Industry-Academia Experience Report | Anton Firc | arXiv |
| 2026-08-14 | AT-ADD: All-Type Audio Deepfake Detection Challenge Summary | Yuankun Xie | arXiv |
| 2026-08-13 | Trajectory Dynamics in Self-Supervised Learning Latent Space for Audio Deepfake Detection | Tomás Andrade Weber | arXiv |
| 2026-08-10 | MADBench: A Benchmark for Modality-Aware Audio Deepfake Detection | Yanqiu Li | arXiv |
| 2026-08-06 | AffectDF: The Most Comprehensive Benchmark for Speech Deepfake Detection against Emotionally Expressive Attacks | Aurosweta Mahapatra | arXiv |
| 2026-08-03 | Multi-Backbone Self-Supervised Ensembles for Audio Deepfake Detection and a Cross-Track Analysis of Generation-Detection Asymmetry | Seunghyun Kim | arXiv |
| 2026-08-01 | Hidden-Domain Routing for All-Type Audio Deepfake Detection | Yifan Gao | arXiv |
| 2026-08-01 | REIMU: Efficient Heterogeneous Hierarchical Reasoning for SSL-Based Speech Deepfake Detection | Kwok-Ho Ng | arXiv |
| 2026-07-30 | Cloned Voices, Real Consequences: Evaluating Bias in Political Deepfake Detection for Electoral Integrity in Brazil | Lucas Rafael Stefanel Gris | arXiv |
| 2026-07-30 | Teffic-Audio: Tell Fact from Fiction | Wan Lin | arXiv |
| 2026-07-30 | Teffic-Audio: Tell Fact from Fiction | Wan Lin | arXiv |
| 2026-07-29 | Audio-Anchored Fusion of Multi-Ratio DiT Reconstruction Residuals for Cross-Domain Audio Deepfake Detection | Haotian Mo | arXiv |
| 2026-07-27 | Leveraging Gradient Reversal Loss and Multitask Learning for Datasets-Aware Audio Deepfake Detection | Mingrui Liang | arXiv |
| 2026-07-24 | How Meta-Learning Shapes LoRA Adapter Geometry in Speech Deepfake Detection | Ivan Kukanov | arXiv |
| 2026-07-23 | Probing Speaker Identity Sensitivity in Audio Deepfake Detectors | Daniyal Kabir Dar | arXiv |
| 2026-07-23 | Toward Interpretable Speech Deepfake Detection using Artifact-Specific Experts and Calibrated Detection Scores | Viola Negroni | arXiv |
| 2026-07-22 | Layer-Wise Decision Fusion for Fake Audio Detection Using XLS-R | Yixuan Xiao | arXiv |
| 2026-07-20 | Time-Frequency Consistency Learning for Robust Speech Deepfake Detection | Jun Xue | arXiv |
| 2026-07-20 | Time-Frequency Consistency Learning for Robust Speech Deepfake Detection | Jun Xue | arXiv |
| 2026-07-17 | Component-Level Ensemble Fusion for Speech and Environmental Sound Deepfake Detection | André Runewicz | arXiv |
| 2026-07-14 | Explainable-by-Design Audio Deepfake Detection via Wiener-Hopf Linear Prediction | Mattia Tamiazzo | arXiv |
| 2026-07-14 | Traceback Translators Against Forgetting in Continual Fake Speech Detection | Enrico Gottardis | arXiv |
| 2026-07-13 | Evidence Subspace Projection: Measuring How Much Evidence Explains Deepfake Detection in Self-Supervised Speech Models | Yixuan Xiao | arXiv |
| 2026-07-11 | PC-Mix: Partial-Component Audio Spoofing Detection under Mixed Speech and Environmental Sound Conditions | Zhenshan Zhang | arXiv |
| 2026-07-10 | What You Train Is What You Get: Gender Bias, Training Composition, and Post-Hoc Mitigation in Audio Deepfake Detection | Aishwarya R. Fursule | arXiv |
| 2026-07-09 | Why Do You Say It Like That? A Phoneme-Level Framework for Explainable Speech Deepfake Detection | Anna Taylor | arXiv |
| 2026-07-06 | SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation | Linxi Li | arXiv |
| 2026-07-03 | DETECT-3B-Omni is Agnostic of Content and Demographics | Nicolas M. Müller | arXiv |
| 2026-07-03 | An Intervention-Based Framework for Shortcut Diagnosis in Spoofing Countermeasures | Santiago Rubio | arXiv |
| 2026-06-29 | Probing-Guided Layer Selection from Self-Supervised Speech Models for Generalizable Audio Deepfake Detection | Marjan Beheshti | arXiv |
| 2026-06-29 | Detecting Audio Deepfakes on the Edge:Lightweight SSL-Based Detection in a Browser Plugin | Octavian Pascu | arXiv |
| 2026-06-28 | Proteus: Automated Adversarial Robustness Testing for Audio Deepfake Detectors | Nicolas M. Müller | arXiv |
| 2026-06-24 | What Do Deepfake Benchmarks Measure? An Audit Using Frozen Self-Supervised Representations | Samuel Pagon | arXiv |
| 2026-06-24 | Supervised Post-training of Speech Foundation Models for Robust Adaptation in Speech Deepfake Detection | Zihan Pan | arXiv |
| 2026-06-22 | The Watermark Shortcut: How Provenance Marking Sabotages Audio Deepfake Detection | Nicolas M. Müller | arXiv |
| 2026-06-17 | FlowFake: Liquid Networks for Audio Deepfake Detection | Shivaay Dhondiyal | arXiv |
| 2026-06-17 | SingFox: A Multi-Lingual Singfake Detection Corpus | Arth J. Shah | arXiv |
| 2026-06-15 | Dual-Granularity Orthogonal Disentanglement for Generalizable Audio Deepfake Detection | Zhuodong Liu | arXiv |
| 2026-06-15 | XAI-Grounded Explanation Generation for Speech Deepfake Detection with Training-Free Multimodal Large Language Models | Yupei Li | arXiv |
| 2026-06-15 | Dual-Granularity Orthogonal Disentanglement for Generalizable Audio Deepfake Detection | Zhuodong Liu | arXiv |
| 2026-06-13 | Phonetically Explainable Speech Deepfake Detection | Manasi Chhibber | arXiv |
| 2026-06-12 | The Perceived Fragility of Explanations in Audio Models: Manipulation of Attribution with Unchanged Predictions | Piotr Kitłowski | arXiv |
| 2026-06-09 | Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge | Xueping Zhang | arXiv |
| 2026-06-09 | Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge | Xueping Zhang | arXiv |
| 2026-06-08 | Linguistically Augmented Audio Speech Data (LinguAS) | Ashley R. Keaton | arXiv |
| 2026-06-08 | Dual-Branch Gated Fusion for Open-Set Audio Deepfake Source Tracing | Awais Khan | arXiv |
| 2026-06-03 | FoeGlass: Simple In-Context Learning Is Enough for Red Teaming Audio Deepfake Detectors | Sepehr Dehdashtian | arXiv |
| 2026-05-28 | Audio Deepfake Detection with Half-Truth Localisation Using Cross-Attentive Feature Fusion | S. Sutharya | arXiv |
| 2026-05-28 | Audio Deepfake Detection with Half-Truth Localisation Using Cross-Attentive Feature Fusion | S. Sutharya | arXiv |
| 2026-05-22 | MixFake: Benchmarking and Enhancing Audio Deepfake Detection in Diverse Real-world Mixed Audio | Qingcao Li | arXiv |
| 2026-05-21 | Eroding Trust in Real Speech: A Large-Scale Study of Human Audio Deepfake Perception | Nicolas M. Müller | arXiv |
| 2026-05-18 | Profiling the Voice: Speaker-Specific Phoneme Fingerprinting for Speech Deepfake Detection | Jun Xue | arXiv |
| 2026-05-18 | Escaping the Linearity Trap: Manifold Detours for Black-Box Adversarial Attacks on Singing Audio Deepfake Detection | Yifan Liao | arXiv |
| 2026-05-10 | RADAR Challenge 2026: Robust Audio Deepfake Recognition under Media Transformations | Hieu-Thi Luong | arXiv |
| 2026-05-10 | RADAR Challenge 2026: Robust Audio Deepfake Recognition under Media Transformations | Hieu-Thi Luong | arXiv |
| 2026-05-09 | Towards Trustworthy Audio Deepfake Detection: A Systematic Framework for Diagnosing and Mitigating Gender Bias | Aishwarya Fursule | arXiv |
| 2026-05-08 | Asymmetric Phase Coding Audio Watermarking | Guang Yang | arXiv |
| 2026-05-07 | Quantum Kernels for Audio Deepfake Detection Using Spectrogram Patch Features | Lisan Al Amin | arXiv |
| 2026-05-05 | Deepfake Audio Detection Using Self-supervised Fusion Representations | Khalid Zaman | arXiv |
| 2026-05-04 | Phoneme-Level Deepfake Detection Across Emotional Conditions Using Self-Supervised Embeddings | Vamshi Nallaguntla | arXiv |
| 2026-05-04 | Toward Fine-Grained Speech Inpainting Forensics:A Dataset, Method, and Metric for Multi-Region Tampering Localization | Tung Vu | arXiv |
| 2026-04-29 | Diffusion Reconstruction towards Generalizable Audio Deepfake Detection | Bo Cheng | arXiv |
| 2026-04-28 | Similarity Choice and Negative Scaling in Supervised Contrastive Learning for Deepfake Audio Detection | Jaskirat Sudan | arXiv |
| 2026-04-26 | RTCFake: Speech Deepfake Detection in Real-Time Communication | Jun Xue | arXiv |
| 2026-04-21 | Environmental Sound Deepfake Detection Using Deep-Learning Framework | Lam Pham | arXiv |
| 2026-04-21 | Audio Spoof Detection with GaborNet | Waldek Maciejko | arXiv |
| 2026-04-21 | Environmental Sound Deepfake Detection Using Deep-Learning Framework | Lam Pham | arXiv |
| 2026-04-19 | HCFD: A Benchmark for Audio Deepfake Detection in Healthcare | Mohd Mujtaba Akhtar | arXiv |
| 2026-04-17 | ICLAD: In-Context Learning with Comparison-Guidance for Audio Deepfake Detection | Benjamin Chou | arXiv |
| 2026-04-15 | Classical Machine Learning Baselines for Deepfake Audio Detection on the Fake-or-Real Dataset | Faheem Ahmad | arXiv |
| 2026-04-14 | ProSDD: Learning Prosodic Representations for Speech Deepfake Detection against Expressive and Emotional Attacks | Aurosweta Mahapatra | arXiv |
| 2026-04-13 | StreamMark: A Deep Learning-Based Semi-Fragile Audio Watermarking for Proactive Deepfake Detection | Zhentao Liu | arXiv |
| 2026-04-09 | Quantum Vision Theory Applied to Audio Classification for Deepfake Speech Detection | Khalid Zaman | arXiv |
| 2026-04-09 | AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan | Yuankun Xie | arXiv |
| 2026-04-09 | DeepFense: A Unified, Modular, and Extensible Framework for Robust Deepfake Audio Detection | Yassine El Kheir | arXiv |
| 2026-04-06 | Joint Fullband-Subband Modeling for High-Resolution SingFake Detection | Xuanjun Chen | arXiv |
| 2026-04-03 | Split and Conquer Partial Deepfake Speech | Inbal Rimon | arXiv |
| 2026-04-03 | If It's Good Enough for You, It's Good Enough for Me: Transferability of Audio Sufficiencies across Models | David A. Kelly | arXiv |
| 2026-04-01 | TRACE: Training-Free Partial Audio Deepfake Detection via Embedding Trajectory Analysis of Speech Foundation Models | Awais Khan | arXiv |
| 2026-03-30 | Audio Language Model for Deepfake Detection Grounded in Acoustic Chain-of-Thought | Runkun Chen | arXiv |
| 2026-03-29 | A General Model for Deepfake Speech Detection: Diverse Bonafide Resources or Diverse AI-Based Generators | Lam Pham | arXiv |
| 2026-03-29 | A General Model for Deepfake Speech Detection: Diverse Bonafide Resources or Diverse AI-Based Generators | Lam Pham | arXiv |
| 2026-03-27 | AFSS: Artifact-Focused Self-Synthesis for Mitigating Bias in Audio Deepfake Detection | Hai-Son Nguyen-Le | arXiv |
| 2026-03-25 | Enhancing Efficiency and Performance in Deepfake Audio Detection through Neuron-level Dropin & Neuroplasticity Mechanisms | Yupei Li | arXiv |
| 2026-03-24 | Echoes: A semantically-aligned music deepfake detection dataset | Octavian Pascu | arXiv |
| 2026-03-21 | SNAP: Speaker Nulling for Artifact Projection in Speech Deepfake Detection | Kyudan Jung | arXiv |
| 2026-03-20 | Audio Avatar Fingerprinting: An Approach for Authorized Use of Voice Cloning in the Era of Synthetic Audio | Candice R. Gerstner | arXiv |
| 2026-03-19 | Enhancing Multi-Corpus Training in SSL-Based Anti-Spoofing Models: Domain-Invariant Feature Extraction | Anh-Tuan Dao | arXiv |
| 2026-03-16 | PhonemeDF: A Synthetic Speech Dataset for Audio Deepfake Detection and Naturalness Evaluation | Vamshi Nallaguntla | arXiv |
| 2026-03-16 | Investigating the Impact of Speech Enhancement on Audio Deepfake Detection in Noisy Environments | Anacin | arXiv |
| 2026-03-13 | Understanding the strengths and weaknesses of SSL models for audio deepfake model attribution | Gabriel Pîrlogeanu | arXiv |
| 2026-03-11 | Towards Robust Speech Deepfake Detection via Human-Inspired Reasoning | Artem Dvirniak | arXiv |
| 2026-03-11 | Probabilistic Verification of Voice Anti-Spoofing Models | Evgeny Kushnir | arXiv |
| 2026-03-10 | Quantizer-Aware Hierarchical Neural Codec Modeling for Speech Deepfake Detection | Jinyang Wu | arXiv |
| 2026-03-09 | Gender Fairness in Audio Deepfake Detection: Performance and Disparity Analysis | Aishwarya Fursule | arXiv |
| 2026-03-09 | Unsupervised Domain Adaptation for Audio Deepfake Detection with Modular Statistical Transformations | Urawee Thani | arXiv |
| 2026-03-06 | Do Compact SSL Backbones Matter for Audio Deepfake Detection? A Controlled Study with RAPTOR | Ajinkya Kulkarni | arXiv |
| 2026-03-06 | How Well Do Current Speech Deepfake Detection Methods Generalize to the Real World? | Daixian Li | arXiv |
| 2026-03-05 | The First Environmental Sound Deepfake Detection Challenge: Benchmarking Robustness, Evaluation, and Insights | Han Yin | arXiv |
| 2026-03-05 | The First Environmental Sound Deepfake Detection Challenge: Benchmarking Robustness, Evaluation, and Insights | Han Yin | arXiv |
| 2026-03-04 | Cyclostationarity Analysis as a Complement to Self-Supervised Representations for Speech Deepfake Detection | Cemal Hanilçi | arXiv |
| 2026-03-03 | Does Fine-tuning by Reinforcement Learning Improve Generalization in Binary Speech Deepfake Detection? | Xin Wang | arXiv |
| 2026-03-02 | A SUPERB-Style Benchmark of Self-Supervised Speech Models for Audio Deepfake Detection | Hashim Ali | arXiv |
| 2026-02-24 | Assessing the Impact of Speaker Identity in Speech Spoofing Detection | Anh-Tuan Dao | arXiv |
| 2026-02-18 | How to Label Resynthesized Audio: The Dual Role of Neural Audio Codecs in Audio Deepfake Detection | Yixuan Xiao | arXiv |
| 2026-02-14 | BreathNet: Generalizable Audio Deepfake Detection via Breath-Cue-Guided Feature Refinement | Zhe Ye | arXiv |
| 2026-02-05 | HyperPotter: Spell the Charm of High-Order Interactions in Audio Deepfake Detection | Qing Wen | arXiv |
| 2026-02-04 | Fine-Grained Frame Modeling in Multi-head Self-Attention for Speech Deepfake Detection | Tuan Dat Phuong | arXiv |
| 2026-02-04 | HoliAntiSpoof: Audio LLM for Holistic Speech Anti-Spoofing | Xuenan Xu | arXiv |
| 2026-02-03 | WST-X Series: Wavelet Scattering Transform for Interpretable Speech Deepfake Detection | Xi Xuan | arXiv |
| 2026-02-01 | HierCon: Hierarchical Contrastive Attention for Audio Deepfake Detection | Zhili Nicholas Liang | arXiv |
| 2026-01-30 | Multi-Speaker Conversational Audio Deepfake: Taxonomy, Dataset and Pilot Study | Alabi Ahmed | arXiv |
| 2026-01-30 | Towards Explicit Acoustic Evidence Perception in Audio LLMs for Speech Deepfake Detection | Xiaoxuan Guo | arXiv |
| 2026-01-30 | Divide and Conquer: Multimodal Video Deepfake Detection via Cross-Modal Fusion and Localization | Qingcao Li | arXiv |
| 2026-01-29 | Localizing Speech Deepfakes Beyond Transitions via Segment-Aware Learning | Yuchen Mao | arXiv |
| 2026-01-28 | Audio Deepfake Detection in the Age of Advanced Text-to-Speech models | Robin Singh | arXiv |
| 2026-01-27 | Audio Deepfake Detection at the First Greeting: "Hi!" | Haohan Shi | arXiv |
| 2026-01-24 | Revealing the Truth with ConLLM for Detecting Multi-Modal Deepfakes | Gautam Siddharth Kashyap | arXiv |
| 2026-01-21 | WeDefense: A Toolkit to Defend Against Fake Audio | Lin Zhang | arXiv |
| 2026-01-21 | Multi-Task Transformer for Explainable Speech Deepfake Detection via Formant Modeling | Viola Negroni | arXiv |
| 2026-01-20 | Emotion and Acoustics Should Agree: Cross-Level Inconsistency Analysis for Audio Deepfake Detection | Jinhua Zhang | arXiv |
| 2026-01-19 | Context and Transcripts Improve Detection of Deepfake Audios of Public Figures | Chongyang Gao | arXiv |
| 2026-01-12 | ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Evaluation Plan | Xueping Zhang | arXiv |
| 2026-01-10 | Lightweight Resolution-Aware Audio Deepfake Detection via Cross-Scale Attention and Consistency Learning | K. A. Shahriar | arXiv |
| 2026-01-07 | Analyzing Reasoning Shifts in Audio Deepfake Detection under Adversarial Attacks: The Reasoning Tax versus Shield Bifurcation | Binh Nguyen | arXiv |
| 2026-01-06 | Interpretable All-Type Audio Deepfake Detection with Audio LLMs via Frequency-Time Reinforcement Learning | Yuankun Xie | arXiv |
| 2026-01-06 | XLSR-MamBo: Scaling the Hybrid Mamba-Attention Backbone for Audio Deepfake Detection | Kwok-Ho Ng | arXiv |
| 2026-01-06 | Vulnerabilities of Audio-Based Biometric Authentication Systems Against Deepfake Speech Synthesis | Mengze Hong | arXiv |
| 2026-01-02 | Investigating the Viability of Employing Multi-modal Large Language Models in the Context of Audio Deepfake Detection | Akanksha Chuchra | arXiv |
| 2025-12-31 | Defense Against Synthetic Speech: Real-Time Detection of RVC Voice Conversion Attacks | Prajwal Chinchmalatpure | arXiv |
| 2025-12-30 | Environmental Sound Deepfake Detection Challenge: An Overview | Han Yin | arXiv |
| 2025-12-25 | Zero-Shot to Zero-Lies: Detecting Bengali Deepfake Audio through Transfer Learning | Most. Sharmin Sultana Samu | arXiv |
| 2025-12-23 | EnvSSLAM-FFN: Lightweight Layer-Fused System for ESDD 2026 Challenge | Xiaoxuan Guo | arXiv |
| 2025-12-21 | Reliable Audio Deepfake Detection in Variable Conditions via Quantum-Kernel SVMs | Lisan Al Amin | arXiv |
| 2025-12-20 | A Data-Centric Approach to Generalizable Speech Deepfake Detection | Wen Huang | arXiv |
| 2025-12-17 | BEAT2AASIST model with layer fusion for ESDD 2026 Challenge | Sanghyeok Chung | arXiv |
| 2025-12-15 | HQ-MPSD: A Multilingual Artifact-Controlled Benchmark for Partial Deepfake Speech Detection | Menglu Li | arXiv |
| 2025-12-15 | Toward Noise-Aware Audio Deepfake Detection: Survey, SNR-Benchmarks, and Practical Recipes | Udayon Sen | arXiv |
| 2025-12-12 | The Affective Bridge: Unifying Feature Representations for Speech Deepfake Detection | Yupei Li | arXiv |
| 2025-12-12 | The Affective Bridge: Preserving Speech Representations while Enhancing Deepfake Detection vian emotional Constraints | Yupei Li | arXiv |
| 2025-12-10 | Human perception of audio deepfakes: the role of language and speaking style | Eugenia San Segundo | arXiv |
| 2025-12-09 | DFALLM: Achieving Generalizable Multitask Deepfake Detection by Optimizing Audio LLM Components | Yupei Li | arXiv |
| 2025-12-09 | BUT Systems for Environmental Sound Deepfake Detection in the ESDD 2026 Challenge | Junyi Peng | arXiv |
| 2025-12-08 | MultiAPI Spoof: A Multi-API Dataset and Local-Attention Network for Speech Anti-spoofing Detection | Xueping Zhang | arXiv |
| 2025-12-05 | Technical Report of Nomi Team in the Environmental Sound Deepfake Detection Challenge 2026 | Candy Olivia Mawalim | arXiv |
| 2025-11-25 | Continual Audio Deepfake Detection via Universal Adversarial Perturbation | Wangjie Li | arXiv |
| 2025-11-13 | Curved Worlds, Clear Boundaries: Generalizing Speech Deepfake Detection using Hyperbolic and Spherical Geometry Spaces | Farhan Sheth | arXiv |
| 2025-10-27 | TwinShift: Benchmarking Audio Deepfake Detection across Synthesizer and Speaker Shifts | Jiyoung Hong | arXiv |
| 2025-10-23 | Can Current Detectors Catch Face-to-Voice Deepfake Attacks? | Nguyen Linh Bao Nguyen | arXiv |
| 2025-10-22 | EchoFake: A Replay-Aware Dataset for Practical Speech Deepfake Detection | Tong Zhang | arXiv |
| 2025-10-14 | FakeMark: Deepfake Speech Attribution With Watermarked Artifacts | Wanying Ge | arXiv |
| 2025-10-06 | WaveSP-Net: Learnable Wavelet-Domain Sparse Prompt Tuning for Speech Deepfake Detection | Xi Xuan | arXiv |
| 2025-10-03 | Forensic Similarity for Speech Deepfakes | Viola Negroni | arXiv |
| 2025-09-30 | On Deepfake Voice Detection -- It's All in the Presentation | Héctor Delgado | arXiv |
| 2025-09-29 | Advancing Zero-Shot Open-Set Speech Deepfake Source Tracing | Manasi Chhibber | arXiv |
| 2025-09-28 | Generalizable Speech Deepfake Detection via Information Bottleneck Enhanced Adversarial Alignment | Pu Huang | arXiv |
| 2025-09-26 | Zero-Day Audio DeepFake Detection via Retrieval Augmentation and Profile Matching | Xuechen Liu | arXiv |
| 2025-09-25 | Addressing Gradient Misalignment in Data-Augmented Training for Robust Speech Deepfake Detection | Duc-Tuan Truong | arXiv |
| 2025-09-25 | QAMO: Quality-aware Multi-centroid One-class Learning For Speech Deepfake Detection | Duc-Tuan Truong | arXiv |
| 2025-09-25 | AUDDT: Audio Unified Deepfake Detection Benchmark Toolkit | Yi Zhu | arXiv |
| 2025-09-24 | SEA-Spoof: Bridging The Gap in Multilingual Audio Deepfake Detection for South-East Asian | Jinyang Wu | arXiv |
| 2025-09-23 | Why Speech Deepfake Detectors Won't Generalize: The Limits of Detection in an Open World | Visar Berisha | arXiv |
| 2025-09-22 | Attention-based Mixture of Experts for Robust Speech Deepfake Detection | Viola Negroni | arXiv |
| 2025-09-21 | FakeSound2: A Benchmark for Explainable and Generalizable Deepfake Sound Detection | Zeyu Xie | arXiv |
| 2025-09-18 | How Does Instrumental Music Help SingFake Detection? | Xuanjun Chen | arXiv |
| 2025-09-17 | Mixture of Low-Rank Adapter Experts in Generalizable Audio Deepfake Detection | Janne Laakkonen | arXiv |
| 2025-09-15 | Improving Out-of-Domain Audio Deepfake Detection via Layer Selection and Fusion of SSL-Based Countermeasures | Pierre Serrano | arXiv |
| 2025-09-13 | Emoanti: audio anti-deepfake with refined emotion-guided representations | Xiaokang Li | arXiv |
| 2025-09-12 | Towards Data Drift Monitoring for Speech Deepfake Detection in the context of MLOps | Xin Wang | arXiv |
| 2025-09-11 | Bona fide Cross Testing Reveals Weak Spot in Audio Deepfake Detection Systems | Chin Yuen Kwok | arXiv |
| 2025-09-11 | MoLEx: Mixture of LoRA Experts in Speech Self-Supervised Models for Audio Deepfake Detection | Zihan Pan | arXiv |
| 2025-09-10 | Audio Deepfake Verification | Li Wang | arXiv |
| 2025-09-09 | When Fine-Tuning is Not Enough: Lessons from HSAD on Hybrid and Adversarial Audio Spoof Detection | Bin Hu | arXiv |
| 2025-09-08 | Adversarial Attacks on Audio Deepfake Detection: A Benchmark and Comparative Study | Kutub Uddin | arXiv |
| 2025-09-08 | Speaker Privacy and Security in the Big Data Era: Protection and Defense against Deepfake | Liping Chen | arXiv |
| 2025-09-05 | XMUspeech Systems for the ASVspoof 5 Challenge | Wangjie Li | arXiv |
| 2025-09-04 | Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection | Yunqi Hao | arXiv |
| 2025-09-04 | NE-PADD: Leveraging Named Entity Knowledge for Robust Partial Audio Deepfake Detection via Attention Aggregation | Huhong Xian | arXiv |
| 2025-09-04 | AUDETER: A Large-scale Dataset for Deepfake Audio Detection in Open Worlds | Qizhou Wang | arXiv |
| 2025-09-03 | Multi-level SSL Feature Gating for Audio Deepfake Detection | Hoan My Tran | arXiv |
| 2025-09-02 | Speech DF Arena: A Leaderboard for Speech DeepFake Detection Models | Sandipana Dowerah | arXiv |
| 2025-08-29 | Generalizable Audio Spoofing Detection using Non-Semantic Representations | Arnab Das | arXiv |
| 2025-08-28 | Multilingual Dataset Integration Strategies for Robust Audio Deepfake Detection: A SAFE Challenge System | Hashim Ali | arXiv |
| 2025-08-14 | Fake Speech Wild: Detecting Deepfake Speech on Social Media Platform | Yuankun Xie | arXiv |
| 2025-08-13 | Perturbed Public Voices (P$^{2}$V): A Dataset for Robust Audio Deepfake Detection | Chongyang Gao | arXiv |
| 2025-08-12 | Fake-Mamba: Real-Time Speech Deepfake Detection Using Bidirectional Mamba as Self-Attention's Alternative | Xi Xuan | arXiv |
| 2025-08-11 | SCDF: A Speaker Characteristics DeepFake Speech Dataset for Bias Analysis | Vojtěch Staněk | arXiv |
| 2025-08-06 | ESDD 2026: Environmental Sound Deepfake Detection Challenge Evaluation Plan | Han Yin | arXiv |
| 2025-08-06 | Multilingual Source Tracing of Speech Deepfakes: A First Benchmark | Xi Xuan | arXiv |
| 2025-08-04 | Towards Reliable Audio Deepfake Attribution and Model Recognition: A Multi-Level Autoencoder-Based Framework | Andrea Di Pierno | arXiv |
| 2025-08-03 | Generalizable Audio Deepfake Detection via Hierarchical Structure Learning and Feature Whitening in Poincaré sphere | Mingru Yang | arXiv |
| 2025-08-02 | Multi-Granularity Adaptive Time-Frequency Attention Framework for Audio Deepfake Detection under Real-World Communication Degradations | Haohan Shi | arXiv |
| 2025-08-01 | Fusion of Modulation Spectrogram and SSL with Multi-head Attention for Fake Speech Detection | Rishith Sadashiv T N | arXiv |
| 2025-07-29 | SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods | Wen Huang | arXiv |
| 2025-07-27 | Two Views, One Truth: Spectral and Self-Supervised Features Fusion for Robust Speech Deepfake Detection | Yassine El Kheir | arXiv |
| 2025-07-22 | LENS-DF: Deepfake Detection and Temporal Localization for Long-Form Noisy Speech | Xuechen Liu | arXiv |
| 2025-07-20 | Frame-level Temporal Difference Learning for Partial Deepfake Speech Detection | Menglu Li | arXiv |
| 2025-07-17 | Enkidu: Universal Frequential Perturbation for Real-Time Audio Privacy Protection against Voice Deepfakes | Zhou Feng | arXiv |
| 2025-07-17 | SHIELD: A Secure and Highly Enhanced Integrated Learning for Robust Deepfake Detection against Adversarial Attacks | Kutub Uddin | arXiv |
| 2025-07-15 | Towards Scalable AASIST: Refining Graph Attention for Speech Deepfake Detection | Ivan Viakhirev | arXiv |
| 2025-07-11 | Phoneme-Level Analysis for Person-of-Interest Speech Deepfake Detection | Davide Salvi | arXiv |
| 2025-07-11 | RawTFNet: A Lightweight CNN Architecture for Speech Anti-spoofing | Yang Xiao | arXiv |
| 2025-07-09 | Open-Set Source Tracing of Audio Deepfake Systems | Nicholas Klein | arXiv |
| 2025-07-07 | Evaluating Fake Music Detection Performance Under Audio Augmentations | Tomasz Sroka | arXiv |
| 2025-07-04 | Robust Localization of Partially Fake Speech: Metrics and Out-of-Domain Evaluation | Hieu-Thi Luong | arXiv |
| 2025-07-02 | Generalizable Detection of Audio Deepfakes | Jose A. Lopez | arXiv |
| 2025-06-30 | Collecting, Curating, and Annotating Good Quality Speech deepfake dataset for Famous Figures: Process and Challenges | Hashim Ali | arXiv |
| 2025-06-26 | Post-training for Deepfake Speech Detection | Wanying Ge | arXiv |
| 2025-06-23 | IndieFake Dataset: A Benchmark Dataset for Audio Deepfake Detection | Abhay Kumar | arXiv |
| 2025-06-17 | A Comparative Study on Proactive and Passive Detection of Deepfake Speech | Chia-Hua Wu | arXiv |
| 2025-06-17 | Manipulated Regions Localization For Partially Deepfake Audio: A Survey | Jiayi He | arXiv |
| 2025-06-14 | Towards Neural Audio Codec Source Parsing | Orchid Chetia Phukan | arXiv |
| 2025-06-13 | From Sharpness to Better Generalization for Speech Deepfake Detection | Wen Huang | arXiv |
| 2025-06-11 | Unmasking real-world audio deepfakes: A data-centric approach | David Combei | arXiv |
| 2025-06-10 | Context-aware TFL: A Universal Context-aware Contrastive Learning Framework for Temporal Forgery Localization | Qilin Yin | arXiv |
| 2025-06-10 | Multimodal Zero-Shot Framework for Deepfake Hate Speech Detection in Low-Resource Languages | Rishabh Ranjan | arXiv |
| 2025-06-08 | Towards Generalized Source Tracing for Codec-Based Deepfake Speech | Xuanjun Chen | arXiv |
| 2025-06-07 | SynHate: Detecting Hate Speech in Synthetic Deepfake Audio | Rishabh Ranjan | arXiv |
| 2025-06-07 | Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? | Bikash Dutta | arXiv |
| 2025-06-06 | TADA: Training-free Attribution and Out-of-Domain Detection of Audio Deepfakes | Adriana Stan | arXiv |
| 2025-06-03 | A Data-Driven Diffusion-based Approach for Audio Deepfake Explanations | Petr Grinberg | arXiv |
| 2025-06-03 | Trusted Fake Audio Detection Based on Dirichlet Distribution | Chi Ding | arXiv |
| 2025-06-03 | PartialEdit: Identifying Partial Deepfakes in the Era of Neural Speech Editing | You Zhang | arXiv |
| 2025-06-02 | Unveiling Audio Deepfake Origins: A Deep Metric learning And Conformer Network Approach With Ensemble Fusion | Ajinkya Kulkarni | arXiv |
| 2025-05-31 | XMAD-Bench: Cross-Domain Multilingual Audio Deepfake Benchmark | Ioan-Paul Ciobanu | arXiv |
| 2025-05-31 | RPRA-ADD: Forgery Trace Enhancement-Driven Audio Deepfake Detection | Ruibo Fu | arXiv |
| 2025-05-30 | Rehearsal with Auxiliary-Informed Sampling for Audio Deepfake Detection | Falih Gozi Febrinanto | arXiv |
| 2025-05-29 | Few-Shot Speech Deepfake Detection Adaptation with Gaussian Processes | Neta Glazer | arXiv |
| 2025-05-26 | STOPA: A Database of Systematic VariaTion Of DeePfake Audio for Open-Set Source Tracing and Attribution | Anton Firc | arXiv |
| 2025-05-25 | EnvSDD: Benchmarking Environmental Sound Deepfake Detection | Han Yin | arXiv |
| 2025-05-23 | What You Read Isn't What You Hear: Linguistic Sensitivity in Deepfake Speech Detection | Binh Nguyen | arXiv |
| 2025-05-20 | Replay Attacks Against Audio Deepfake Detection | Nicolas Müller | arXiv |
| 2025-05-20 | Listen, Analyze, and Adapt to Learn New Attacks: An Exemplar-Free Class Incremental Learning Method for Audio Deepfake Source Tracing | Yang Xiao | arXiv |
| 2025-05-20 | Naturalness-Aware Curriculum Learning with Dynamic Temperature for Speech Deepfake Detection | Taewoo Kim | arXiv |
| 2025-05-20 | BiCrossMamba-ST: Speech Deepfake Detection with Bidirectional Mamba Spectro-Temporal Cross-Attention | Yassine El Kheir | arXiv |
| 2025-05-20 | Forensic deepfake audio detection using segmental speech features | Tianle Yang | arXiv |
| 2025-05-20 | Source Verification for Speech Deepfakes | Viola Negroni | arXiv |
| 2025-05-19 | Codec-Based Deepfake Source Tracing via Neural Audio Codec Taxonomy | Xuanjun Chen | arXiv |
| 2025-05-16 | ALLM4ADD: Unlocking the Capabilities of Audio Large Language Models for Audio Deepfake Detection | Hao Gu | arXiv |
| 2025-05-16 | BanglaFake: Constructing and Evaluating a Specialized Bengali Deepfake Audio Dataset | Istiaq Ahmed Fahad | arXiv |
| 2025-05-10 | Beyond Identity: A Generalizable Approach for Deepfake Audio Detection | Yasaman Ahmadiadli | arXiv |
| 2025-04-29 | End-to-end Audio Deepfake Detection from RAW Waveforms: a RawNet-Based Approach with Cross-Dataset Evaluation | Andrea Di Pierno | arXiv |
| 2025-04-22 | FADEL: Uncertainty-aware Fake Audio Detection with Evidential Deep Learning | Ju Yeon Kang | arXiv |
| 2025-04-16 | Benchmarking Audio Deepfake Detection Robustness in Real-world Communication Scenarios | Haohan Shi | arXiv |
| 2025-04-15 | Generalized Audio Deepfake Detection Using Frame-level Latent Information Entropy | Botao Zhao | arXiv |
| 2025-04-14 | SafeSpeech: Robust and Universal Voice Protection Against Malicious Speech Synthesis | Zhisheng Zhang | arXiv |
| 2025-04-09 | Detect All-Type Deepfake Audio: Wavelet Prompt Tuning for Enhanced Auditory Perception | Yuankun Xie | arXiv |
| 2025-04-08 | Nes2Net: A Lightweight Nested Architecture for Foundation Model Driven Speech Anti-spoofing | Tianchi Liu | arXiv |
| 2025-03-23 | Anomaly Detection and Localization for Speech Deepfakes via Feature Pyramid Matching | Emma Coletta | arXiv |
| 2025-03-21 | Measuring the Robustness of Audio Deepfake Detectors | Xiang Li | arXiv |
| 2025-03-15 | Adaptive Mixture of Low-Rank Experts for Robust Audio Spoofing Detection | Qixian Chen | arXiv |
| 2025-02-27 | DIN-CTS: Low-Complexity Depthwise-Inception Neural Network with Contrastive Training Strategy for Deepfake Speech Detection | Lam Pham | arXiv |
| 2025-02-27 | DeePen: Penetration Testing for Audio Deepfake Detection | Nicolas Müller | arXiv |
| 2025-02-20 | Pitch Imperfect: Detecting Audio Deepfakes Through Acoustic Prosodic Analysis | Kevin Warren | arXiv |
| 2025-02-15 | Generalizable speech deepfake detection via meta-learned LoRA | Janne Laakkonen | arXiv |
| 2025-02-14 | A Preliminary Exploration with GPT-4o Voice Mode | Yu-Xiang Lin | arXiv |
| 2025-02-13 | SyntheticPop: Attacking Speaker Verification Systems With Synthetic VoicePops | Eshaq Jamdar | arXiv |
| 2025-02-13 | ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech | Xin Wang | arXiv |
| 2025-02-06 | XAttnMark: Learning Robust Audio Watermarking with Cross-Attention | Yixin Liu | arXiv |
| 2025-02-05 | Comprehensive Layer-wise Analysis of SSL Models for Audio Deepfake Detection | Yassine El Kheir | arXiv |
| 2025-01-24 | Generalizable Audio Deepfake Detection via Latent Space Refinement and Augmentation | Wen Huang | arXiv |
| 2025-01-23 | What Does an Audio Deepfake Detector Focus on? A Study in the Time Domain | Petr Grinberg | arXiv |
| 2025-01-21 | Transferable Adversarial Attacks on Audio Deepfake Detection | Muhammad Umar Farooq | arXiv |
| 2025-01-14 | CodecFake+: A Large-Scale Neural Audio Codec-Based Deepfake Speech Dataset | Xuanjun Chen | arXiv |
| 2025-01-11 | Neural Codec Source Tracing: Toward Comprehensive Attribution in Open-Set Condition | Yuankun Xie | arXiv |
| 2025-01-09 | Unmasking Deepfakes: Leveraging Augmentations and Features Variability for Deepfake Speech Detection | Inbal Rimon | arXiv |
| 2025-01-09 | SIGNL: A Label-Efficient Audio Deepfake Detection System via Spectral-Temporal Graph Non-Contrastive Learning | Falih Gozi Febrinanto | arXiv |
| 2025-01-09 | DiffAttack: Diffusion-based Timbre-reserved Adversarial Attack in Speaker Identification | Qing Wang | arXiv |
| 2024-12-24 | Explaining Speaker and Spoof Embeddings via Probing | Xuechen Liu | arXiv |
| 2024-12-23 | Are audio DeepFake detection models polyglots? | Bartłomiej Marek | arXiv |
| 2024-12-23 | Investigating Prosodic Signatures via Speech Pre-Trained Models for Audio Deepfake Source Attribution | Orchid Chetia Phukan | arXiv |
| 2024-12-17 | Phoneme-Level Feature Discrepancies: A Key to Detecting Sophisticated Speech Deepfakes | Kuiyuan Zhang | arXiv |
| 2024-12-16 | Region-Based Optimization in Continual Learning for Audio Deepfake Detection | Yujie Chen | arXiv |
| 2024-12-12 | Audios Don't Lie: Multi-Frequency Channel Attention Mechanism for Audio Deepfake Detection | Yangguang Feng | arXiv |
| 2024-12-02 | Reject Threshold Adaptation for Open-Set Model Attribution of Deepfake Audio | Xinrui Yan | arXiv |
| 2024-11-30 | From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview | Yupei Li | arXiv |
| 2024-11-29 | Parallel Stacked Aggregated Network for Voice Authentication in IoT-Enabled Smart Devices | Awais Khan | arXiv |
| 2024-11-26 | Comparative Analysis of ASR Methods for Speech Deepfake Detection | Davide Salvi | arXiv |
| 2024-11-22 | VQalAttent: a Transparent Speech Generation Pipeline based on Transformer-learned VQ-VAE Latent Space | Armani Rodriguez | arXiv |
| 2024-11-21 | Listening for Expert Identified Linguistic Features: Assessment of Audio Deepfake Discernment among Undergraduate Students | Noshaba N. Bhalli | arXiv |
| 2024-11-14 | Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation | Kuiyuan Zhang | arXiv |
| 2024-11-08 | Toward Transdisciplinary Approaches to Audio Deepfake Discernment | Vandana P. Janeja | arXiv |
| 2024-10-31 | I Can Hear You: Selective Robust Training for Deepfake Audio Detection | Zirui Zhang | arXiv |
| 2024-10-28 | Mitigating Unauthorized Speech Synthesis for Voice Protection | Zhisheng Zhang | arXiv |
| 2024-10-27 | Meta-Learning Approaches for Improving Detection of Unseen Speech Deepfakes | Ivan Kukanov | arXiv |
| 2024-10-21 | ALDAS: Audio-Linguistic Data Augmentation for Spoofed Audio Detection | Zahra Khanjani | arXiv |
| 2024-10-13 | Prompt Tuning for Audio Deepfake Detection: Computationally Efficient Test-time Domain Adaptation with Limited Target Dataset | Hideyuki Oiso | arXiv |
| 2024-10-11 | Quantum-Trained Convolutional Neural Network for Deepfake Audio Detection | Chu-Hsuan Abraham Lin | arXiv |
| 2024-10-09 | Toward Robust Real-World Audio Deepfake Detection: Closing the Explainability Gap | Georgia Channing | arXiv |
| 2024-10-09 | Learn from Real: Reality Defender's Submission to ASVspoof5 Challenge | Yi Zhu | arXiv |
| 2024-10-09 | Diffuse or Confuse: A Diffusion Deepfake Speech Dataset | Anton Firc | arXiv |
| 2024-10-09 | Can DeepFake Speech be Reliably Detected? | Hongbin Liu | arXiv |
| 2024-10-06 | Where are we in audio deepfake detection? A systematic analysis over generative and detection models | Xiang Li | arXiv |
| 2024-10-04 | A Multimodal Framework for Deepfake Detection | Kashish Gandhi | arXiv |
| 2024-10-01 | Augmentation through Laundering Attacks for Audio Spoof Detection | Hashim Ali | arXiv |
| 2024-09-26 | Freeze and Learn: Continual Learning with Selective Freezing for Speech Deepfake Detection | Davide Salvi | arXiv |
| 2024-09-24 | Representation Loss Minimization with Randomized Selection Strategy for Efficient Environmental Fake Audio Detection | Orchid Chetia Phukan | arXiv |
| 2024-09-24 | Leveraging Mixture of Experts for Improved Speech Deepfake Detection | Viola Negroni | arXiv |
| 2024-09-23 | A Comprehensive Survey with Critical Analysis for Deepfake Speech Detection | Lam Pham | arXiv |
| 2024-09-23 | LlamaPartialSpoof: An LLM-Driven Fake Speech Dataset Simulating Disinformation Generation | Hieu-Thi Luong | arXiv |
| 2024-09-23 | Room Impulse Responses help attackers to evade Deep Fake Detection | Hieu-Thi Luong | arXiv |
| 2024-09-21 | Strong Alone, Stronger Together: Synergizing Modality-Binding Foundation Models with Optimal Transport for Non-Verbal Emotion Recognition | Orchid Chetia Phukan | arXiv |
| 2024-09-18 | Mixture of Experts Fusion for Fake Audio Detection Using Frozen wav2vec 2.0 | Zhiyong Wang | arXiv |
| 2024-09-18 | SpoofCeleb: Speech Deepfake Detection and SASV In The Wild | Jee-weon Jung | arXiv |
| 2024-09-14 | SafeEar: Content Privacy-Preserving Audio Deepfake Detection | Xinfeng Li | arXiv |
| 2024-09-13 | DFADD: The Diffusion and Flow-Matching Based Audio Deepfake Dataset | Jiawei Du | arXiv |
| 2024-09-11 | D-CAPTCHA++: A Study of Resilience of Deepfake CAPTCHA under Transferable Imperceptible Adversarial Attack | Hong-Hanh Nguyen-Le | arXiv |
| 2024-09-09 | Continuous Learning of Transformer-based Audio Deepfake Detection | Tuan Duy Nguyen Le | arXiv |
| 2024-09-08 | Exploring WavLM Back-ends for Speech Spoofing and Deepfake Detection | Theophile Stourbe | arXiv |
| 2024-09-03 | USTC-KXDIGIT System Description for ASVspoof5 Challenge | Yihao Chen | arXiv |
| 2024-08-30 | Utilizing Speaker Profiles for Impersonation Audio Detection | Hao Gu | arXiv |
| 2024-08-30 | AASIST3: KAN-Enhanced AASIST Speech Deepfake Detection using SSL Features and Additional Regularization for the ASVspoof 2024 Challenge | Kirill Borodin | arXiv |
| 2024-08-27 | Is Audio Spoof Detection Robust to Laundering Attacks? | Hashim Ali | arXiv |
| 2024-08-26 | A Preliminary Case Study on Long-Form In-the-Wild Audio Spoofing Detection | Xuechen Liu | arXiv |
| 2024-08-25 | Analyzing the Impact of Splicing Artifacts in Partially Fake Speech Signals | Viola Negroni | arXiv |
| 2024-08-23 | Toward Improving Synthetic Audio Spoofing Detection Robustness via Meta-Learning and Disentangled Training With Adversarial Examples | Zhenyu Wang | arXiv |
| 2024-08-20 | Does Current Deepfake Audio Detection Model Effectively Detect ALM-based Deepfake Audio? | Yuankun Xie | arXiv |
| 2024-08-20 | A Noval Feature via Color Quantisation for Fake Audio Detection | Zhiyong Wang | arXiv |
| 2024-08-19 | ASASVIcomtech: The Vicomtech-UGR Speech Deepfake Detection and SASV Systems for the ASVspoof5 Challenge | Juan M. Martín-Doñas | arXiv |
| 2024-08-14 | WavLM model ensemble for audio deepfake detection | David Combei | arXiv |
| 2024-08-13 | Temporal Variability and Multi-Viewed Self-Supervised Representations to Tackle the ASVspoof5 Deepfake Challenge | Yuankun Xie | arXiv |
| 2024-08-09 | ADD 2023: Towards Audio Deepfake Detection and Analysis in the Wild | Jiangyan Yi | arXiv |
| 2024-07-26 | SLIM: Style-Linguistics Mismatch Model for Generalized Audio Deepfake Detection | Yi Zhu | arXiv |
| 2024-07-14 | Advancing Continual Learning for Robust Deepfake Audio Classification | Feiyi Dong | arXiv |
| 2024-07-11 | An Unsupervised Domain Adaptation Method for Locating Manipulated Region in partially fake Audio | Siding Zeng | arXiv |
| 2024-07-10 | Source Tracing of Audio Deepfake Systems | Nicholas Klein | arXiv |
| 2024-07-10 | Targeted Augmented Data for Audio Deepfake Detection | Marcella Astrid | arXiv |
| 2024-07-03 | Towards Attention-based Contrastive Learning for Audio Spoof Detection | Chirag Goel | arXiv |
| 2024-07-01 | Deepfake Audio Detection Using Spectrogram-based Feature and Ensemble of Deep Learning Models | Lam Pham | arXiv |
| 2024-06-24 | One-Class Learning with Adaptive Centroid Shift for Audio Deepfake Detection | Hyun Myung Kim | arXiv |
| 2024-06-20 | A Multi-Stream Fusion Approach with One-Class Learning for Audio-Visual Deepfake Detection | Kyungbok Lee | arXiv |
| 2024-06-14 | Frequency-mix Knowledge Distillation for Fake Speech Detection | Cunhang Fan | arXiv |
| 2024-06-13 | Interpretable Temporal Class Activation Representation for Audio Spoofing Detection | Menglu Li | arXiv |
| 2024-06-12 | FakeSound: Deepfake General Audio Detection | Zeyu Xie | arXiv |
| 2024-06-12 | Codecfake: An Initial Dataset for Detecting LLM-based Deepfake Audio | Yi Lu | arXiv |
| 2024-06-11 | CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems | Haibin Wu | arXiv |
| 2024-06-10 | RawBMamba: End-to-End Bidirectional State Space Model for Audio Deepfake Detection | Yujie Chen | arXiv |
| 2024-06-05 | Generalized Source Tracing: Detecting Novel Audio Deepfake Algorithm with Real Emphasis and Fake Dispersion Strategy | Yuankun Xie | arXiv |
| 2024-06-05 | Harder or Different? Understanding Generalization of Audio Deepfake Detection | Nicolas M. Müller | arXiv |
| 2024-06-05 | Genuine-Focused Learning using Mask AutoEncoder for Generalized Fake Audio Detection | Xiaopeng Wang | arXiv |
| 2024-06-05 | Generalized Fake Audio Detection via Deep Stable Learning | Zhiyong Wang | arXiv |
| 2024-06-05 | Singing Voice Graph Modeling for SingFake Detection | Xuanjun Chen | arXiv |
| 2024-05-14 | Towards Robust Audio Deepfake Detection: A Evolving Benchmark for Continual Learning | Xiaohui Zhang | arXiv |
| 2024-05-08 | The Codecfake Dataset and Countermeasures for the Universally Detection of Deepfake Audio | Yuankun Xie | arXiv |
| 2024-05-03 | Training-Free Deepfake Voice Recognition by Leveraging Large-Scale Pre-Trained Models | Alessandro Pianese | arXiv |
| 2024-04-26 | An RFP dataset for Real, Fake, and Partially fake audio detection | Abdulazeez AlAli | arXiv |
| 2024-04-24 | CLAD: Robust Audio Deepfake Detection Against Manipulation Attacks with Contrastive Learning | Haolin Wu | arXiv |
| 2024-04-23 | Every Breath You Don't Take: Deepfake Speech Detection Using Breath | Seth Layton | arXiv |
| 2024-04-22 | Retrieval-Augmented Audio Deepfake Detection | Zuheng Kang | arXiv |
| 2024-04-22 | A Survey on Speech Deepfake Detection | Menglu Li | arXiv |
| 2024-04-19 | Enhancing Generalization in Audio Deepfake Detection: A Neural Collapse based Sampling and Training Approach | Mohammed Yousif | arXiv |
| 2024-04-18 | HyDiscGAN: A Hybrid Distributed cGAN for Audio-Visual Privacy Preservation in Multimodal Sentiment Analysis | Zhuojia Wu | arXiv |
| 2024-04-07 | Cross-Domain Audio Deepfake Detection: Dataset and Analysis | Yuang Li | arXiv |
| 2024-03-31 | Heterogeneity over Homogeneity: Investigating Multilingual Speech Pre-Trained Models for Detecting Audio Deepfake | Orchid Chetia Phukan | arXiv |
| 2024-03-26 | Detection of Deepfake Environmental Audio | Hafsa Ouajdi | arXiv |
| 2024-03-21 | Exploring Green AI for Audio Deepfake Detection | Subhajit Saha | arXiv |
| 2024-03-18 | Towards the Development of a Real-Time Deepfake Audio Detection System in Communication Platforms | Jonat John Mathew | arXiv |
| 2024-03-04 | A robust audio deepfake detection system via multi-view feature | Yujie Yang | arXiv |
| 2024-02-28 | PITCH: AI-assisted Tagging of Deepfake Audio Calls using Challenge-Response | Govind Mittal | arXiv |
| 2024-02-22 | Human Brain Exhibits Distinct Patterns When Listening to Fake Versus Real Audio: Preliminary Evidence | Mahsa Salehi | arXiv |
| 2024-02-09 | A New Approach to Voice Authenticity | Nicolas M. Müller | arXiv |
| 2024-01-24 | MOS-FAD: Improving Fake Audio Detection Via Automatic Mean Opinion Score Prediction | Wangjin Zhou | arXiv |
| 2024-01-17 | MLAAD: The Multi-Language Audio Anti-Spoofing Dataset | Nicolas M. Müller | arXiv |
| 2024-01-11 | Self-Attention and Hybrid Features for Replay and Deep-Fake Audio Detection | Lian Huang | arXiv |
| 2024-01-04 | AntiDeepFake: AI for Deep Fake Speech Recognition | Enkhtogtokh Togootogtokh | arXiv |
| 2023-12-15 | What to Remember: Self-Adaptive Continual Learning for Audio Deepfake Detection | Xiaohui Zhang | arXiv |
| 2023-12-13 | Audio Deepfake Detection with Self-Supervised WavLM and Multi-Fusion Attentive Classifier | Yinlin Guo | arXiv |
| 2023-11-29 | Vulnerability of Automatic Identity Recognition to Audio-Visual Deepfakes | Pavel Korshunov | arXiv |
| 2023-11-26 | AV-Deepfake1M: A Large-Scale LLM-Driven Audio-Visual Deepfake Dataset | Zhixi Cai | arXiv |
| 2023-11-06 | MFAAN: Unveiling Audio Deepfakes with a Multi-Feature Authenticity Network | Karthik Sivarama Krishnan | arXiv |
| 2023-10-05 | Securing Voice Biometrics: One-Shot Learning Approach for Audio Deepfake Detection | Awais Khan | arXiv |
| 2023-09-22 | Deepfake audio as a data augmentation technique for training automatic speech to text transcription models | Alexandre R. Ferreira | arXiv |
| 2023-09-15 | Characterizing the temporal dynamics of universal speech representations for generalizable deepfake detection | Yi Zhu | arXiv |
| 2023-09-15 | HM-Conformer: A Conformer-based audio deepfake detection system with hierarchical pooling and multi-level classification token aggregation methods | Hyun-seo Shin | arXiv |
| 2023-09-11 | Towards generalisable and calibrated synthetic speech detection with self-supervised representations | Octavian Pascu | arXiv |
| 2023-09-05 | FSD: An Initial Chinese Dataset for Fake Song Detection | Yuankun Xie | arXiv |
| 2023-09-02 | Timbre-reserved Adversarial Attack in Speaker Identification | Qing Wang | arXiv |
| 2023-08-29 | Audio Deepfake Detection: A Survey | Jiangyan Yi | arXiv |
| 2023-08-22 | Complex-valued neural networks for voice anti-spoofing | Nicolas M. Müller | arXiv |
| 2023-08-20 | The DKU-DUKEECE System for the Manipulation Region Location Task of ADD 2023 | Zexin Cai | arXiv |
| 2023-08-19 | Spatial Reconstructed Local Attention Res2Net with F0 Subband for Fake Speech Detection | Cunhang Fan | arXiv |
| 2023-08-07 | Do You Remember? Overcoming Catastrophic Forgetting for Fake Audio Detection | Xiaohui Zhang | arXiv |
| 2023-07-28 | All-for-One and One-For-All: Deep learning-based feature fusion for Synthetic Speech Detection | Daniele Mari | arXiv |
| 2023-07-13 | Uncovering the Deceptions: An Analysis on Audio Spoofing Detection and Future Prospects | Rishabh Ranjan | arXiv |
| 2023-07-03 | An End-to-End Multi-Module Audio Deepfake Generation System for ADD Challenge 2023 | Sheng Zhao | arXiv |
| 2023-06-27 | TranssionADD: A multi-frame reinforcement based sequence tagging model for audio deepfake detection | Jie Liu | arXiv |
| 2023-06-27 | Multi-perspective Information Fusion Res2Net with RandomSpecmix for Fake Speech Detection | Shunbo Dong | arXiv |
| 2023-06-09 | Low-rank Adaptation Method for Wav2vec2-based Fake Audio Detection | Chenglong Wang | arXiv |
| 2023-06-08 | Adaptive Fake Audio Detection with Low-Rank Model Squeezing | Xiaohui Zhang | arXiv |
| 2023-06-02 | Improved DeepFake Detection Using Whisper Features | Piotr Kawa | arXiv |
| 2023-05-30 | Pseudo-Siamese Network based Timbre-reserved Black-box Adversarial Attack in Speaker Identification | Qing Wang | arXiv |
| 2023-05-25 | Betray Oneself: A Novel Audio DeepFake Detection Model via Mono-to-Stereo Conversion | Rui Liu | arXiv |
| 2023-05-23 | ADD 2023: the Second Audio Deepfake Detection Challenge | Jiangyan Yi | arXiv |
| 2023-05-23 | TO-Rawnet: Improving RawNet with TCN and Orthogonal Regularization for Fake Audio Detection | Chenglong Wang | arXiv |
| 2023-05-23 | Detection of Cross-Dataset Fake Audio Based on Prosodic and Pronunciation Features | Chenglong Wang | arXiv |
| 2023-05-22 | The defender's perspective on automatic speaker verification: An overview | Haibin Wu | arXiv |
| 2023-05-18 | Improving Generalization Ability of Countermeasures for New Mismatch Scenario by Combining Multiple Advanced Regularization Terms | Chang Zeng | arXiv |
| 2023-04-25 | AI-Synthesized Voice Detection Using Neural Vocoder Artifacts | Chengzhe Sun | arXiv |
| 2023-03-02 | Learning From Yourself: A Self-Distillation Method for Fake Speech Detection | Jun Xue | arXiv |
| 2023-02-20 | Hello Me, Meet the Real Me: Audio Deepfake Attacks on Voice Assistants | Domna Bilika | arXiv |
| 2023-02-18 | Exposing AI-Synthesized Human Voices Using Neural Vocoder Artifacts | Chengzhe Sun | arXiv |
| 2023-01-19 | Warning: Humans Cannot Reliably Detect Speech Deepfakes | Kimberly T. Mai | arXiv |
| 2023-01-08 | Deepfake CAPTCHA: A Method for Preventing Fake Calls | Lior Yasur | arXiv |
| 2022-12-30 | Defense Against Adversarial Attacks on Audio DeepFake Detection | Piotr Kawa | arXiv |
| 2022-12-16 | Source Tracing: Detecting Voice Spoofing | Tinglong Zhu | arXiv |
| 2022-11-15 | Improved disentangled speech representations using contrastive learning in factorized hierarchical variational autoencoder | Yuying Xie | arXiv |
| 2022-11-11 | SceneFake: An Initial Dataset and Benchmarks for Scene Fake Audio Detection | Jiangyan Yi | arXiv |
| 2022-11-10 | EmoFake: An Initial Dataset for Emotion Fake Audio Detection | Yan Zhao | arXiv |
| 2022-11-01 | Waveform Boundary Detection for Partially Spoofed Audio | Zexin Cai | arXiv |
| 2022-10-31 | Combining Automatic Speaker Verification and Prosody Analysis for Synthetic Speech Detection | Luigi Attorresi | arXiv |
| 2022-10-21 | Adaptive re-calibration of channel-wise features for Adversarial Audio Classification | Vardhan Dongre | arXiv |
| 2022-10-13 | Deepfake Detection System for the ADD Challenge Track 3.2 Based on Score Fusion | Yuxiang Zhang | arXiv |
| 2022-10-12 | SpecRNet: Towards Faster and More Accessible Audio DeepFake Detection | Piotr Kawa | arXiv |
| 2022-10-11 | Deep Spectro-temporal Artifacts for Detecting Synthesized Speech | Xiaohui Liu | arXiv |
| 2022-10-05 | ASVspoof 2021: Towards Spoofed and Deepfake Speech Detection in the Wild | Xuechen Liu | arXiv |
| 2022-09-28 | Deepfake audio detection by speaker verification | Alessandro Pianese | arXiv |
| 2022-09-26 | Faked Speech Detection with Zero Prior Knowledge | Sahar Al Ajmi | arXiv |
| 2022-09-16 | TIMIT-TTS: a Text-to-Speech Dataset for Multimodal Synthetic Media Detection | Davide Salvi | arXiv |
| 2022-08-21 | Audio Deepfake Attribution: An Initial Dataset and Investigation | Xinrui Yan | arXiv |
| 2022-08-20 | An Initial Investigation for Detecting Vocoder Fingerprints of Fake Audio | Xinrui Yan | arXiv |
| 2022-08-20 | Fully Automated End-to-End Fake Audio Detection | Chenglong Wang | arXiv |
| 2022-08-02 | Audio Deepfake Detection Based on a Combination of F0 Information and Real Plus Imaginary Spectrogram Features | Jun Xue | arXiv |
| 2022-07-12 | CFAD: A Chinese Dataset for Fake Audio Detection | Haoxin Ma | arXiv |
| 2022-06-27 | Attack Agnostic Dataset: Towards Generalization and Stabilization of Audio DeepFake Detection | Piotr Kawa | arXiv |
| 2022-04-19 | Audio Deep Fake Detection System with Neural Stitching for ADD 2022 | Rui Yan | arXiv |
| 2022-04-19 | Time Domain Adversarial Voice Conversion for ADD 2022 | Cheng Wen | arXiv |
| 2022-04-11 | The PartialSpoof Database and Countermeasures for the Detection of Short Fake Speech Segments Embedded in an Utterance | Lin Zhang | arXiv |
| 2022-03-30 | Does Audio Deepfake Detection Generalize? | Nicolas M. Müller | arXiv |
| 2022-03-28 | Attacker Attribution of Audio Deepfakes | Nicolas M. Müller | arXiv |
| 2022-03-17 | Towards a New Science of Disinformation | Claudio S. Pinhanez | arXiv |
| 2022-03-03 | The Vicomtech Audio Deepfake Detection System based on Wav2Vec2 for the 2022 ADD Challenge | Juan M. Martín-Doñas | arXiv |
| 2022-02-25 | Human Detection of Political Speech Deepfakes across Transcripts, Audio, and Video | Matthew Groh | arXiv |
| 2022-02-17 | ADD 2022: the First Audio Deep Synthesis Detection Challenge | Jiangyan Yi | arXiv |
| 2022-02-14 | Partially Fake Audio Detection by Self-attention-based Fake Span Discovery | Haibin Wu | arXiv |
| 2022-02-09 | CAU_KU team's submission to ADD 2022 Challenge task 1: Low-quality fake audio detection through frequency feature masking | Il-Youp Kwak | arXiv |
| 2022-01-29 | The HCCL-DKU system for fake audio generation task of the 2022 ICASSP ADD Challenge | Ziyi Chen | arXiv |
| 2021-12-06 | Audio Deepfake Perceptions in College Going Populations | Gabrielle Watson | arXiv |
| 2021-11-28 | How Deep Are the Fakes? Focusing on Audio Deepfake: A Survey | Zahra Khanjani | arXiv |
| 2021-11-04 | WaveFake: A Data Set to Facilitate Audio Deepfake Detection | Joel Frank | arXiv |
| 2021-09-07 | Evaluation of an Audio-Video Multimodal Deepfake Dataset using Unimodal and Multimodal Detectors | Hasam Khalid | arXiv |
| 2021-09-01 | ASVspoof 2021: accelerating progress in spoofed and deepfake speech detection | Junichi Yamagishi | arXiv |
| 2021-09-01 | ASVspoof 2021: Automatic Speaker Verification Spoofing and Countermeasures Challenge Evaluation Plan | Héctor Delgado | arXiv |
| 2021-08-11 | FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset | Hasam Khalid | arXiv |
| 2021-07-27 | End-to-End Spectro-Temporal Graph Attention Networks for Speaker Verification Anti-Spoofing and Speech Deepfake Detection | Hemlata Tak | arXiv |
| 2021-07-26 | Raw Differentiable Architecture Search for Speech Deepfake and Spoofing Detection | Wanying Ge | arXiv |
| 2021-07-26 | UR Channel-Robust Synthetic Speech Detection System for ASVspoof 2021 | Xinhui Chen | arXiv |
| 2021-07-20 | Human Perception of Audio Deepfakes | Nicolas M. Müller | arXiv |
| 2021-04-20 | Identification of fake stereo audio | Tianyun Liu | arXiv |
| 2021-04-15 | Continual Learning for Fake Audio Detection | Haoxin Ma | arXiv |
| 2021-04-08 | Generalized Spoofing Detection Inspired from Audio Generation Artifacts | Yang Gao | arXiv |
| 2021-04-08 | Half-Truth: A Partially Fake Audio Detection Dataset | Jiangyan Yi | arXiv |
| 2020-11-07 | Detection and Evaluation of human and machine generated speech in spoofing attacks on automatic speaker verification systems | Yang Gao | arXiv |
| 2019-07-01 | Analysis by Adversarial Synthesis -- A Novel Approach for Speech Vocoding | Ahmed Mustafa | arXiv |
| 2019-04-13 | Towards Vulnerability Analysis of Voice-Driven Interfaces and Countermeasures for Replay | Khalid Mahmood Malik | arXiv |
| 2019-04-09 | ASVspoof 2019: Future Horizons in Spoofed and Fake Audio Detection | Massimiliano Todisco | arXiv |