Speech DF Arena: A Leaderboard for Speech DeepFake Detection Models
Fuente:
arXiv
Saved in:
| Main Authors: | Dowerah, Sandipana, Kulkarni, Atharva, Kulkarni, Ajinkya, Tran, Hoan My, Kalda, Joonas, Fedorchenko, Artem, Fauve, Benoit, Lolive, Damien, Alumäe, Tanel, Doss, Matthew Magimai |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do Compact SSL Backbones Matter for Audio Deepfake Detection? A Controlled Study with RAPTOR
by: Kulkarni, Ajinkya, et al.
Published: (2026)
by: Kulkarni, Ajinkya, et al.
Published: (2026)
Unveiling Audio Deepfake Origins: A Deep Metric learning And Conformer Network Approach With Ensemble Fusion
by: Kulkarni, Ajinkya, et al.
Published: (2025)
by: Kulkarni, Ajinkya, et al.
Published: (2025)
TalTech Systems for the Interspeech 2025 ML-SUPERB 2.0 Challenge
by: Alumäe, Tanel, et al.
Published: (2025)
by: Alumäe, Tanel, et al.
Published: (2025)
Optimizing Estonian TV Subtitles with Semi-supervised Learning and LLMs
by: Fedorchenko, Artem, et al.
Published: (2025)
by: Fedorchenko, Artem, et al.
Published: (2025)
Unsupervised Rhythm and Voice Conversion of Dysarthric to Healthy Speech for ASR
by: Hajal, Karl El, et al.
Published: (2025)
by: Hajal, Karl El, et al.
Published: (2025)
kNN Retrieval for Simple and Effective Zero-Shot Multi-speaker Text-to-Speech
by: Hajal, Karl El, et al.
Published: (2024)
by: Hajal, Karl El, et al.
Published: (2024)
PixIT: Joint Training of Speaker Diarization and Speech Separation from Real-world Multi-speaker Recordings
by: Kalda, Joonas, et al.
Published: (2024)
by: Kalda, Joonas, et al.
Published: (2024)
Multilingual Hidden Prompt Injection Attacks on LLM-Based Academic Reviewing
by: Theocharopoulos, Panagiotis, et al.
Published: (2025)
by: Theocharopoulos, Panagiotis, et al.
Published: (2025)
Unveiling Biases while Embracing Sustainability: Assessing the Dual Challenges of Automatic Speech Recognition Systems
by: Kulkarni, Ajinkya, et al.
Published: (2025)
by: Kulkarni, Ajinkya, et al.
Published: (2025)
Comparison of End-to-end Speech Assessment Models for the NOCASA 2025 Challenge
by: Žavoronkov, Aleksei, et al.
Published: (2025)
by: Žavoronkov, Aleksei, et al.
Published: (2025)
On the Utility of Speech and Audio Foundation Models for Marmoset Call Analysis
by: Sarkar, Eklavya, et al.
Published: (2024)
by: Sarkar, Eklavya, et al.
Published: (2024)
TalTech-IRIT-LIS Speaker and Language Diarization Systems for DISPLACE 2024
by: Kalda, Joonas, et al.
Published: (2024)
by: Kalda, Joonas, et al.
Published: (2024)
Supplementary Information for: "Speech power spectra: a window into neural oscillations in Parkinson's disease"
by: Hovsepyan, Sevada, et al.
Published: (2025)
by: Hovsepyan, Sevada, et al.
Published: (2025)
Comparing Self-Supervised Learning Models Pre-Trained on Human Speech and Animal Vocalizations for Bioacoustics Processing
by: Sarkar, Eklavya, et al.
Published: (2025)
by: Sarkar, Eklavya, et al.
Published: (2025)
Children's Voice Privacy: First Steps And Emerging Challenges
by: Kulkarni, Ajinkya, et al.
Published: (2025)
by: Kulkarni, Ajinkya, et al.
Published: (2025)
Can DeepFake Speech be Reliably Detected?
by: Liu, Hongbin, et al.
Published: (2024)
by: Liu, Hongbin, et al.
Published: (2024)
Predicting Heart Activity from Speech using Data-driven and Knowledge-based features
by: Elbanna, Gasser, et al.
Published: (2024)
by: Elbanna, Gasser, et al.
Published: (2024)
Multi-Source Evidence Fusion for Audio Question Answering
by: Olev, Aivo, et al.
Published: (2026)
by: Olev, Aivo, et al.
Published: (2026)
Multi-level SSL Feature Gating for Audio Deepfake Detection
by: Tran, Hoan My, et al.
Published: (2025)
by: Tran, Hoan My, et al.
Published: (2025)
Assessment of Personality Dimensions Across Situations Using Conversational Speech
by: Zhang, Alice, et al.
Published: (2025)
by: Zhang, Alice, et al.
Published: (2025)
Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech
by: Hajal, Karl El, et al.
Published: (2025)
by: Hajal, Karl El, et al.
Published: (2025)
Estonian Native Large Language Model Benchmark
by: Lillepalu, Helena Grete, et al.
Published: (2025)
by: Lillepalu, Helena Grete, et al.
Published: (2025)
Finetuning End-to-End Models for Estonian Conversational Spoken Language Translation
by: Sildam, Tiia, et al.
Published: (2024)
by: Sildam, Tiia, et al.
Published: (2024)
SCDF: A Speaker Characteristics DeepFake Speech Dataset for Bias Analysis
by: Staněk, Vojtěch, et al.
Published: (2025)
by: Staněk, Vojtěch, et al.
Published: (2025)
Towards Leveraging Sequential Structure in Animal Vocalizations
by: Sarkar, Eklavya, et al.
Published: (2025)
by: Sarkar, Eklavya, et al.
Published: (2025)
Documentation of sediment core HE443/10-3
by: Henkel, Susann, et al.
Published: (2015)
by: Henkel, Susann, et al.
Published: (2015)
Toward using Speech to Sense Student Emotion in Remote Learning Environments
by: Vyas, Sargam, et al.
Published: (2026)
by: Vyas, Sargam, et al.
Published: (2026)
TriDF: Evaluating Perception, Detection, and Hallucination for Interpretable DeepFake Detection
by: Jiang-Lin, Jian-Yu, et al.
Published: (2025)
by: Jiang-Lin, Jian-Yu, et al.
Published: (2025)
DeepFake-Adapter: Dual-Level Adapter for DeepFake Detection
by: Shao, Rui, et al.
Published: (2023)
by: Shao, Rui, et al.
Published: (2023)
Code-Switching in End-to-End Automatic Speech Recognition: A Systematic Literature Review
by: Agro, Maha Tufail, et al.
Published: (2025)
by: Agro, Maha Tufail, et al.
Published: (2025)
Active Fake: DeepFake Camouflage
by: Sun, Pu, et al.
Published: (2024)
by: Sun, Pu, et al.
Published: (2024)
Celeb-DF++: A Large-scale Challenging Video DeepFake Benchmark for Generalizable Forensics
by: Li, Yuezun, et al.
Published: (2025)
by: Li, Yuezun, et al.
Published: (2025)
DeepFake-O-Meter v2.0: An Open Platform for DeepFake Detection
by: Ju, Yan, et al.
Published: (2024)
by: Ju, Yan, et al.
Published: (2024)
Physical Design: Methodologies and Developments
by: Kulkarni, Atharva M., et al.
Published: (2024)
by: Kulkarni, Atharva M., et al.
Published: (2024)
Personalized Speech Enhancement Without a Separate Speaker Embedding Model
by: Pärnamaa, Tanel, et al.
Published: (2024)
by: Pärnamaa, Tanel, et al.
Published: (2024)
Robust Sequential DeepFake Detection
by: Shao, Rui, et al.
Published: (2023)
by: Shao, Rui, et al.
Published: (2023)
SpeechColab Leaderboard: An Open-Source Platform for Automatic Speech Recognition Evaluation
by: Du, Jiayu, et al.
Published: (2024)
by: Du, Jiayu, et al.
Published: (2024)
Are audio DeepFake detection models polyglots?
by: Marek, Bartłomiej, et al.
Published: (2024)
by: Marek, Bartłomiej, et al.
Published: (2024)
Do DeepFake Attribution Models Generalize?
by: Baxavanakis, Spiros, et al.
Published: (2025)
by: Baxavanakis, Spiros, et al.
Published: (2025)
FakeParts: a New Family of AI-Generated DeepFakes
by: Liu, Ziyi, et al.
Published: (2025)
by: Liu, Ziyi, et al.
Published: (2025)
Similar Items
-
Do Compact SSL Backbones Matter for Audio Deepfake Detection? A Controlled Study with RAPTOR
by: Kulkarni, Ajinkya, et al.
Published: (2026) -
Unveiling Audio Deepfake Origins: A Deep Metric learning And Conformer Network Approach With Ensemble Fusion
by: Kulkarni, Ajinkya, et al.
Published: (2025) -
TalTech Systems for the Interspeech 2025 ML-SUPERB 2.0 Challenge
by: Alumäe, Tanel, et al.
Published: (2025) -
Optimizing Estonian TV Subtitles with Semi-supervised Learning and LLMs
by: Fedorchenko, Artem, et al.
Published: (2025) -
Unsupervised Rhythm and Voice Conversion of Dysarthric to Healthy Speech for ASR
by: Hajal, Karl El, et al.
Published: (2025)