A SUPERB-Style Benchmark of Self-Supervised Speech Models for Audio Deepfake Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ali, Hashim, Adupa, Nithin Sai, Subramani, Surya, Malik, Hafiz |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multilingual Dataset Integration Strategies for Robust Audio Deepfake Detection: A SAFE Challenge System
von: Ali, Hashim, et al.
Veröffentlicht: (2025)
von: Ali, Hashim, et al.
Veröffentlicht: (2025)
Augmentation through Laundering Attacks for Audio Spoof Detection
von: Ali, Hashim, et al.
Veröffentlicht: (2024)
von: Ali, Hashim, et al.
Veröffentlicht: (2024)
Collecting, Curating, and Annotating Good Quality Speech deepfake dataset for Famous Figures: Process and Challenges
von: Ali, Hashim, et al.
Veröffentlicht: (2025)
von: Ali, Hashim, et al.
Veröffentlicht: (2025)
LJ-Spoof: A Generatively Varied Corpus for Audio Anti-Spoofing and Synthesis Source Tracing
von: Subramani, Surya, et al.
Veröffentlicht: (2026)
von: Subramani, Surya, et al.
Veröffentlicht: (2026)
Is Audio Spoof Detection Robust to Laundering Attacks?
von: Ali, Hashim, et al.
Veröffentlicht: (2024)
von: Ali, Hashim, et al.
Veröffentlicht: (2024)
Similarity Choice and Negative Scaling in Supervised Contrastive Learning for Deepfake Audio Detection
von: Sudan, Jaskirat, et al.
Veröffentlicht: (2026)
von: Sudan, Jaskirat, et al.
Veröffentlicht: (2026)
SLIM: Style-Linguistics Mismatch Model for Generalized Audio Deepfake Detection
von: Zhu, Yi, et al.
Veröffentlicht: (2024)
von: Zhu, Yi, et al.
Veröffentlicht: (2024)
Mitigating Intra-Speaker Variability in Diarization with Style-Controllable Speech Augmentation
von: Kim, Miseul, et al.
Veröffentlicht: (2025)
von: Kim, Miseul, et al.
Veröffentlicht: (2025)
Benchmarking Audio Deepfake Detection Robustness in Real-world Communication Scenarios
von: Shi, Haohan, et al.
Veröffentlicht: (2025)
von: Shi, Haohan, et al.
Veröffentlicht: (2025)
On the Parameter Estimation of Sinusoidal Models for Speech and Audio Signals
von: Kafentzis, George P.
Veröffentlicht: (2024)
von: Kafentzis, George P.
Veröffentlicht: (2024)
A Hybrid Model for Weakly-Supervised Speech Dereverberation
von: Bahrman, Louis, et al.
Veröffentlicht: (2025)
von: Bahrman, Louis, et al.
Veröffentlicht: (2025)
Neural Speech and Audio Coding: Modern AI Technology Meets Traditional Codecs
von: Kim, Minje, et al.
Veröffentlicht: (2024)
von: Kim, Minje, et al.
Veröffentlicht: (2024)
Privacy-Preserving End-to-End Full-Duplex Speech Dialogue Models
von: Kuzmin, Nikita, et al.
Veröffentlicht: (2026)
von: Kuzmin, Nikita, et al.
Veröffentlicht: (2026)
Neural Spectral Band Generation for Audio Coding
von: Choi, Woongjib, et al.
Veröffentlicht: (2025)
von: Choi, Woongjib, et al.
Veröffentlicht: (2025)
Relational Proxy Loss for Audio-Text based Keyword Spotting
von: Jung, Youngmoon, et al.
Veröffentlicht: (2024)
von: Jung, Youngmoon, et al.
Veröffentlicht: (2024)
Crowdsourced Multilingual Speech Intelligibility Testing
von: Lechler, Laura, et al.
Veröffentlicht: (2024)
von: Lechler, Laura, et al.
Veröffentlicht: (2024)
AudioLDM 2: Learning Holistic Audio Generation with Self-supervised Pretraining
von: Liu, Haohe, et al.
Veröffentlicht: (2023)
von: Liu, Haohe, et al.
Veröffentlicht: (2023)
Speech Enhancement Based on Drifting Models
von: Xu, Liang, et al.
Veröffentlicht: (2026)
von: Xu, Liang, et al.
Veröffentlicht: (2026)
Beyond Identity: A Generalizable Approach for Deepfake Audio Detection
von: Ahmadiadli, Yasaman, et al.
Veröffentlicht: (2025)
von: Ahmadiadli, Yasaman, et al.
Veröffentlicht: (2025)
Single-stage TTS with Masked Audio Token Modeling and Semantic Knowledge Distillation
von: Gállego, Gerard I., et al.
Veröffentlicht: (2024)
von: Gállego, Gerard I., et al.
Veröffentlicht: (2024)
Voice Mapping of Text-to-Speech Systems: A Metric-Based Approach for Voice Quality Assessment
von: Cai, Huanchen, et al.
Veröffentlicht: (2026)
von: Cai, Huanchen, et al.
Veröffentlicht: (2026)
Audio Deepfake Detection in the Age of Advanced Text-to-Speech models
von: Singh, Robin, et al.
Veröffentlicht: (2026)
von: Singh, Robin, et al.
Veröffentlicht: (2026)
Language Bias in Self-Supervised Learning For Automatic Speech Recognition
von: Storey, Edward, et al.
Veröffentlicht: (2025)
von: Storey, Edward, et al.
Veröffentlicht: (2025)
Speech Boosting: Low-Latency Live Speech Enhancement for TWS Earbuds
von: Bae, Hanbin, et al.
Veröffentlicht: (2024)
von: Bae, Hanbin, et al.
Veröffentlicht: (2024)
Laugh Now Cry Later: Controlling Time-Varying Emotional States of Flow-Matching-Based Zero-Shot Text-to-Speech
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
Does Current Deepfake Audio Detection Model Effectively Detect ALM-based Deepfake Audio?
von: Xie, Yuankun, et al.
Veröffentlicht: (2024)
von: Xie, Yuankun, et al.
Veröffentlicht: (2024)
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation
von: Kim, Ji-Hoon, et al.
Veröffentlicht: (2024)
von: Kim, Ji-Hoon, et al.
Veröffentlicht: (2024)
Exploring Disentangled Neural Speech Codecs from Self-Supervised Representations
von: Aihara, Ryo, et al.
Veröffentlicht: (2025)
von: Aihara, Ryo, et al.
Veröffentlicht: (2025)
Compressing Quaternion Convolutional Neural Networks for Audio Classification
von: Singh, Arshdeep, et al.
Veröffentlicht: (2025)
von: Singh, Arshdeep, et al.
Veröffentlicht: (2025)
Joint Semantic Knowledge Distillation and Masked Acoustic Modeling for Full-band Speech Restoration with Improved Intelligibility
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2024)
EEG-Based Speech Decoding: A Novel Approach Using Multi-Kernel Ensemble Diffusion Models
von: Kim, Soowon, et al.
Veröffentlicht: (2024)
von: Kim, Soowon, et al.
Veröffentlicht: (2024)
Mind the Prompt: Prompting Strategies in Audio Generations for Improving Sound Classification
von: Ronchini, Francesca, et al.
Veröffentlicht: (2025)
von: Ronchini, Francesca, et al.
Veröffentlicht: (2025)
Lightweight Self-Supervised Detection of Fundamental Frequency and Accurate Probability of Voicing in Monophonic Music
von: Bitra, Venkat Suprabath, et al.
Veröffentlicht: (2026)
von: Bitra, Venkat Suprabath, et al.
Veröffentlicht: (2026)
Robust Generative Audio Quality Assessment: Disentangling Quality from Spurious Correlations
von: Huang, Kuan-Tang, et al.
Veröffentlicht: (2026)
von: Huang, Kuan-Tang, et al.
Veröffentlicht: (2026)
Construction and Evaluation of Mandarin Multimodal Emotional Speech Database
von: Ting, Zhu, et al.
Veröffentlicht: (2024)
von: Ting, Zhu, et al.
Veröffentlicht: (2024)
UniverSR: Unified and Versatile Audio Super-Resolution via Vocoder-Free Flow Matching
von: Choi, Woongjib, et al.
Veröffentlicht: (2025)
von: Choi, Woongjib, et al.
Veröffentlicht: (2025)
IndieFake Dataset: A Benchmark Dataset for Audio Deepfake Detection
von: Kumar, Abhay, et al.
Veröffentlicht: (2025)
von: Kumar, Abhay, et al.
Veröffentlicht: (2025)
JenGAN: Stacked Shifted Filters in GAN-Based Speech Synthesis
von: Cho, Hyunjae, et al.
Veröffentlicht: (2024)
von: Cho, Hyunjae, et al.
Veröffentlicht: (2024)
Toward Fully-End-to-End Listened Speech Decoding from EEG Signals
von: Lee, Jihwan, et al.
Veröffentlicht: (2024)
von: Lee, Jihwan, et al.
Veröffentlicht: (2024)
Automatic Speech Recognition using Advanced Deep Learning Approaches: A survey
von: Kheddar, Hamza, et al.
Veröffentlicht: (2024)
von: Kheddar, Hamza, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Multilingual Dataset Integration Strategies for Robust Audio Deepfake Detection: A SAFE Challenge System
von: Ali, Hashim, et al.
Veröffentlicht: (2025) -
Augmentation through Laundering Attacks for Audio Spoof Detection
von: Ali, Hashim, et al.
Veröffentlicht: (2024) -
Collecting, Curating, and Annotating Good Quality Speech deepfake dataset for Famous Figures: Process and Challenges
von: Ali, Hashim, et al.
Veröffentlicht: (2025) -
LJ-Spoof: A Generatively Varied Corpus for Audio Anti-Spoofing and Synthesis Source Tracing
von: Subramani, Surya, et al.
Veröffentlicht: (2026) -
Is Audio Spoof Detection Robust to Laundering Attacks?
von: Ali, Hashim, et al.
Veröffentlicht: (2024)