A General Model for Deepfake Speech Detection: Diverse Bonafide Resources or Diverse AI-Based Generators
Fuente:
arXiv
Salvato in:
| Autori principali: | Pham, Lam, Vu, Khoi, Tran, Dat, Fischinger, David, Schindler, Alexander, Boyer, Martin, McLoughlin, Ian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Environmental Sound Deepfake Detection Using Deep-Learning Framework
di: Pham, Lam, et al.
Pubblicazione: (2026)
di: Pham, Lam, et al.
Pubblicazione: (2026)
DIN-CTS: Low-Complexity Depthwise-Inception Neural Network with Contrastive Training Strategy for Deepfake Speech Detection
di: Pham, Lam, et al.
Pubblicazione: (2025)
di: Pham, Lam, et al.
Pubblicazione: (2025)
A Comprehensive Survey with Critical Analysis for Deepfake Speech Detection
di: Pham, Lam, et al.
Pubblicazione: (2024)
di: Pham, Lam, et al.
Pubblicazione: (2024)
An Efficient Transfer Learning Method Based on Adapter with Local Attributes for Speech Emotion Recognition
di: Song, Haoyu, et al.
Pubblicazione: (2025)
di: Song, Haoyu, et al.
Pubblicazione: (2025)
The Impact of Frequency Bands on Acoustic Anomaly Detection of Machines using Deep Learning Based Model
di: Nguyen, Tin, et al.
Pubblicazione: (2024)
di: Nguyen, Tin, et al.
Pubblicazione: (2024)
Deepfake Audio Detection Using Spectrogram-based Feature and Ensemble of Deep Learning Models
di: Pham, Lam, et al.
Pubblicazione: (2024)
di: Pham, Lam, et al.
Pubblicazione: (2024)
Fine-Grained Frame Modeling in Multi-head Self-Attention for Speech Deepfake Detection
di: Phuong, Tuan Dat, et al.
Pubblicazione: (2026)
di: Phuong, Tuan Dat, et al.
Pubblicazione: (2026)
Pushing the Performance of Synthetic Speech Detection with Kolmogorov-Arnold Networks and Self-Supervised Learning Models
di: Phuong, Tuan Dat, et al.
Pubblicazione: (2025)
di: Phuong, Tuan Dat, et al.
Pubblicazione: (2025)
MAT-SED: A Masked Audio Transformer with Masked-Reconstruction Based Pre-training for Sound Event Detection
di: Cai, Pengfei, et al.
Pubblicazione: (2024)
di: Cai, Pengfei, et al.
Pubblicazione: (2024)
Prototype based Masked Audio Model for Self-Supervised Learning of Sound Event Detection
di: Cai, Pengfei, et al.
Pubblicazione: (2024)
di: Cai, Pengfei, et al.
Pubblicazione: (2024)
XLSR-Kanformer: A KAN-Intergrated model for Synthetic Speech Detection
di: Dat, Phuong Tuan, et al.
Pubblicazione: (2025)
di: Dat, Phuong Tuan, et al.
Pubblicazione: (2025)
CodecFlow: Efficient Bandwidth Extension via Conditional Flow Matching in Neural Codec Latent Space
di: Zhang, Bowen, et al.
Pubblicazione: (2026)
di: Zhang, Bowen, et al.
Pubblicazione: (2026)
Continuous Learning of Transformer-based Audio Deepfake Detection
di: Le, Tuan Duy Nguyen, et al.
Pubblicazione: (2024)
di: Le, Tuan Duy Nguyen, et al.
Pubblicazione: (2024)
Detect Any Sound: Open-Vocabulary Sound Event Detection with Multi-Modal Queries
di: Cai, Pengfei, et al.
Pubblicazione: (2025)
di: Cai, Pengfei, et al.
Pubblicazione: (2025)
MAGE: A Coarse-to-Fine Speech Enhancer with Masked Generative Model
di: Pham, The Hieu, et al.
Pubblicazione: (2025)
di: Pham, The Hieu, et al.
Pubblicazione: (2025)
From Sharpness to Better Generalization for Speech Deepfake Detection
di: Huang, Wen, et al.
Pubblicazione: (2025)
di: Huang, Wen, et al.
Pubblicazione: (2025)
Qwen vs. Gemma Integration with Whisper: A Comparative Study in Multilingual SpeechLLM Systems
di: Nguyen, Tuan, et al.
Pubblicazione: (2025)
di: Nguyen, Tuan, et al.
Pubblicazione: (2025)
How Well Do Current Speech Deepfake Detection Methods Generalize to the Real World?
di: Li, Daixian, et al.
Pubblicazione: (2026)
di: Li, Daixian, et al.
Pubblicazione: (2026)
Aud-Sur: An Audio Analyzer Assistant for Audio Surveillance Applications
di: Lam, Phat, et al.
Pubblicazione: (2025)
di: Lam, Phat, et al.
Pubblicazione: (2025)
A Toolchain for Comprehensive Audio/Video Analysis Using Deep Learning Based Multimodal Approach (A use case of riot or violent context detection)
di: Pham, Lam, et al.
Pubblicazione: (2024)
di: Pham, Lam, et al.
Pubblicazione: (2024)
Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment
di: Hoang, Long-Vu, et al.
Pubblicazione: (2025)
di: Hoang, Long-Vu, et al.
Pubblicazione: (2025)
MixFake: Benchmarking and Enhancing Audio Deepfake Detection in Diverse Real-world Mixed Audio
di: Li, Qingcao, et al.
Pubblicazione: (2026)
di: Li, Qingcao, et al.
Pubblicazione: (2026)
RTCFake: Speech Deepfake Detection in Real-Time Communication
di: Xue, Jun, et al.
Pubblicazione: (2026)
di: Xue, Jun, et al.
Pubblicazione: (2026)
RO-N3WS: Enhancing Generalization in Low-Resource ASR with Diverse Romanian Speech Benchmarks
di: Diaconu, Alexandra, et al.
Pubblicazione: (2026)
di: Diaconu, Alexandra, et al.
Pubblicazione: (2026)
Zero-Shot Text-to-Speech for Vietnamese
di: Vu, Thi, et al.
Pubblicazione: (2025)
di: Vu, Thi, et al.
Pubblicazione: (2025)
A Comparative Study on Proactive and Passive Detection of Deepfake Speech
di: Wu, Chia-Hua, et al.
Pubblicazione: (2025)
di: Wu, Chia-Hua, et al.
Pubblicazione: (2025)
Attention-based Mixture of Experts for Robust Speech Deepfake Detection
di: Negroni, Viola, et al.
Pubblicazione: (2025)
di: Negroni, Viola, et al.
Pubblicazione: (2025)
AsyncSwitch: Asynchronous Text-Speech Adaptation for Code-Switched ASR
di: Nguyen, Tuan, et al.
Pubblicazione: (2025)
di: Nguyen, Tuan, et al.
Pubblicazione: (2025)
Towards Fine-grained Temporal Perception: Post-Training Large Audio-Language Models with Audio-Side Time Prompt
di: Shi, Yanfeng, et al.
Pubblicazione: (2026)
di: Shi, Yanfeng, et al.
Pubblicazione: (2026)
SLIM: Style-Linguistics Mismatch Model for Generalized Audio Deepfake Detection
di: Zhu, Yi, et al.
Pubblicazione: (2024)
di: Zhu, Yi, et al.
Pubblicazione: (2024)
Multimodal Zero-Shot Framework for Deepfake Hate Speech Detection in Low-Resource Languages
di: Ranjan, Rishabh, et al.
Pubblicazione: (2025)
di: Ranjan, Rishabh, et al.
Pubblicazione: (2025)
Forensic Similarity for Speech Deepfakes
di: Negroni, Viola, et al.
Pubblicazione: (2025)
di: Negroni, Viola, et al.
Pubblicazione: (2025)
Unmasking Deepfakes: Leveraging Augmentations and Features Variability for Deepfake Speech Detection
di: Rimon, Inbal, et al.
Pubblicazione: (2025)
di: Rimon, Inbal, et al.
Pubblicazione: (2025)
Profiling the Voice: Speaker-Specific Phoneme Fingerprinting for Speech Deepfake Detection
di: Xue, Jun, et al.
Pubblicazione: (2026)
di: Xue, Jun, et al.
Pubblicazione: (2026)
From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview
di: Li, Yupei, et al.
Pubblicazione: (2024)
di: Li, Yupei, et al.
Pubblicazione: (2024)
Why Speech Deepfake Detectors Won't Generalize: The Limits of Detection in an Open World
di: Berisha, Visar, et al.
Pubblicazione: (2025)
di: Berisha, Visar, et al.
Pubblicazione: (2025)
Fake Speech Wild: Detecting Deepfake Speech on Social Media Platform
di: Xie, Yuankun, et al.
Pubblicazione: (2025)
di: Xie, Yuankun, et al.
Pubblicazione: (2025)
Towards Generalized Source Tracing for Codec-Based Deepfake Speech
di: Chen, Xuanjun, et al.
Pubblicazione: (2025)
di: Chen, Xuanjun, et al.
Pubblicazione: (2025)
Multi-Task Transformer for Explainable Speech Deepfake Detection via Formant Modeling
di: Negroni, Viola, et al.
Pubblicazione: (2026)
di: Negroni, Viola, et al.
Pubblicazione: (2026)
Toward Fine-Grained Speech Inpainting Forensics:A Dataset, Method, and Metric for Multi-Region Tampering Localization
di: Vu, Tung, et al.
Pubblicazione: (2026)
di: Vu, Tung, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Environmental Sound Deepfake Detection Using Deep-Learning Framework
di: Pham, Lam, et al.
Pubblicazione: (2026) -
DIN-CTS: Low-Complexity Depthwise-Inception Neural Network with Contrastive Training Strategy for Deepfake Speech Detection
di: Pham, Lam, et al.
Pubblicazione: (2025) -
A Comprehensive Survey with Critical Analysis for Deepfake Speech Detection
di: Pham, Lam, et al.
Pubblicazione: (2024) -
An Efficient Transfer Learning Method Based on Adapter with Local Attributes for Speech Emotion Recognition
di: Song, Haoyu, et al.
Pubblicazione: (2025) -
The Impact of Frequency Bands on Acoustic Anomaly Detection of Machines using Deep Learning Based Model
di: Nguyen, Tin, et al.
Pubblicazione: (2024)