Multi-Task Transformer for Explainable Speech Deepfake Detection via Formant Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Negroni, Viola, Cuccovillo, Luca, Bestagini, Paolo, Aichroth, Patrick, Tubaro, Stefano |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Source Verification for Speech Deepfakes
von: Negroni, Viola, et al.
Veröffentlicht: (2025)
von: Negroni, Viola, et al.
Veröffentlicht: (2025)
Freeze and Learn: Continual Learning with Selective Freezing for Speech Deepfake Detection
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
Attention-based Mixture of Experts for Robust Speech Deepfake Detection
von: Negroni, Viola, et al.
Veröffentlicht: (2025)
von: Negroni, Viola, et al.
Veröffentlicht: (2025)
Phoneme-Level Analysis for Person-of-Interest Speech Deepfake Detection
von: Salvi, Davide, et al.
Veröffentlicht: (2025)
von: Salvi, Davide, et al.
Veröffentlicht: (2025)
Forensic Similarity for Speech Deepfakes
von: Negroni, Viola, et al.
Veröffentlicht: (2025)
von: Negroni, Viola, et al.
Veröffentlicht: (2025)
Leveraging Mixture of Experts for Improved Speech Deepfake Detection
von: Negroni, Viola, et al.
Veröffentlicht: (2024)
von: Negroni, Viola, et al.
Veröffentlicht: (2024)
Analyzing the Impact of Splicing Artifacts in Partially Fake Speech Signals
von: Negroni, Viola, et al.
Veröffentlicht: (2024)
von: Negroni, Viola, et al.
Veröffentlicht: (2024)
Anomaly Detection and Localization for Speech Deepfakes via Feature Pyramid Matching
von: Coletta, Emma, et al.
Veröffentlicht: (2025)
von: Coletta, Emma, et al.
Veröffentlicht: (2025)
FakeMusicCaps: a Dataset for Detection and Attribution of Synthetic Music Generated via Text-to-Music Models
von: Comanducci, Luca, et al.
Veröffentlicht: (2024)
von: Comanducci, Luca, et al.
Veröffentlicht: (2024)
Comparative Analysis of ASR Methods for Speech Deepfake Detection
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
Listening Between the Lines: Synthetic Speech Detection Disregarding Verbal Content
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
Compression Robust Synthetic Speech Detection Using Patched Spectrogram Transformer
von: Yadav, Amit Kumar Singh, et al.
Veröffentlicht: (2024)
von: Yadav, Amit Kumar Singh, et al.
Veröffentlicht: (2024)
Neural Encoding Detection is Not All You Need for Synthetic Speech Detection
von: Cuccovillo, Luca, et al.
Veröffentlicht: (2026)
von: Cuccovillo, Luca, et al.
Veröffentlicht: (2026)
POLIPHONE: A Dataset for Smartphone Model Identification from Audio Recordings
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
Improving Robustness of Diffusion-Based Zero-Shot Speech Synthesis via Stable Formant Generation
von: Han, Changjin, et al.
Veröffentlicht: (2024)
von: Han, Changjin, et al.
Veröffentlicht: (2024)
Fine-Grained Frame Modeling in Multi-head Self-Attention for Speech Deepfake Detection
von: Phuong, Tuan Dat, et al.
Veröffentlicht: (2026)
von: Phuong, Tuan Dat, et al.
Veröffentlicht: (2026)
RTCFake: Speech Deepfake Detection in Real-Time Communication
von: Xue, Jun, et al.
Veröffentlicht: (2026)
von: Xue, Jun, et al.
Veröffentlicht: (2026)
A Comparative Study on Proactive and Passive Detection of Deepfake Speech
von: Wu, Chia-Hua, et al.
Veröffentlicht: (2025)
von: Wu, Chia-Hua, et al.
Veröffentlicht: (2025)
Towards Robust Speech Deepfake Detection via Human-Inspired Reasoning
von: Dvirniak, Artem, et al.
Veröffentlicht: (2026)
von: Dvirniak, Artem, et al.
Veröffentlicht: (2026)
Unmasking Deepfakes: Leveraging Augmentations and Features Variability for Deepfake Speech Detection
von: Rimon, Inbal, et al.
Veröffentlicht: (2025)
von: Rimon, Inbal, et al.
Veröffentlicht: (2025)
Profiling the Voice: Speaker-Specific Phoneme Fingerprinting for Speech Deepfake Detection
von: Xue, Jun, et al.
Veröffentlicht: (2026)
von: Xue, Jun, et al.
Veröffentlicht: (2026)
Visual and audio scene classification for detecting discrepancies in video: a baseline method and experimental protocol
von: Apostolidis, Konstantinos, et al.
Veröffentlicht: (2024)
von: Apostolidis, Konstantinos, et al.
Veröffentlicht: (2024)
Fake Speech Wild: Detecting Deepfake Speech on Social Media Platform
von: Xie, Yuankun, et al.
Veröffentlicht: (2025)
von: Xie, Yuankun, et al.
Veröffentlicht: (2025)
Generalizable Speech Deepfake Detection via Information Bottleneck Enhanced Adversarial Alignment
von: Huang, Pu, et al.
Veröffentlicht: (2025)
von: Huang, Pu, et al.
Veröffentlicht: (2025)
QAMO: Quality-aware Multi-centroid One-class Learning For Speech Deepfake Detection
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2025)
von: Truong, Duc-Tuan, et al.
Veröffentlicht: (2025)
Localizing Speech Deepfakes Beyond Transitions via Segment-Aware Learning
von: Mao, Yuchen, et al.
Veröffentlicht: (2026)
von: Mao, Yuchen, et al.
Veröffentlicht: (2026)
Spectrogram-Based Detection of Auto-Tuned Vocals in Music Recordings
von: Gohari, Mahyar, et al.
Veröffentlicht: (2024)
von: Gohari, Mahyar, et al.
Veröffentlicht: (2024)
SNAP: Speaker Nulling for Artifact Projection in Speech Deepfake Detection
von: Jung, Kyudan, et al.
Veröffentlicht: (2026)
von: Jung, Kyudan, et al.
Veröffentlicht: (2026)
A Data-Centric Approach to Generalizable Speech Deepfake Detection
von: Huang, Wen, et al.
Veröffentlicht: (2025)
von: Huang, Wen, et al.
Veröffentlicht: (2025)
HiFi-Glot: High-Fidelity Neural Formant Synthesis with Differentiable Resonant Filters
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
How Well Do Current Speech Deepfake Detection Methods Generalize to the Real World?
von: Li, Daixian, et al.
Veröffentlicht: (2026)
von: Li, Daixian, et al.
Veröffentlicht: (2026)
ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Evaluation Plan
von: Zhang, Xueping, et al.
Veröffentlicht: (2026)
von: Zhang, Xueping, et al.
Veröffentlicht: (2026)
PhonemeDF: A Synthetic Speech Dataset for Audio Deepfake Detection and Naturalness Evaluation
von: Nallaguntla, Vamshi, et al.
Veröffentlicht: (2026)
von: Nallaguntla, Vamshi, et al.
Veröffentlicht: (2026)
HQ-MPSD: A Multilingual Artifact-Controlled Benchmark for Partial Deepfake Speech Detection
von: Li, Menglu, et al.
Veröffentlicht: (2025)
von: Li, Menglu, et al.
Veröffentlicht: (2025)
From Sharpness to Better Generalization for Speech Deepfake Detection
von: Huang, Wen, et al.
Veröffentlicht: (2025)
von: Huang, Wen, et al.
Veröffentlicht: (2025)
When Synthetic Traces Hide Real Content: Analysis of Stable Diffusion Image Laundering
von: Mandelli, Sara, et al.
Veröffentlicht: (2024)
von: Mandelli, Sara, et al.
Veröffentlicht: (2024)
MoLEx: Mixture of LoRA Experts in Speech Self-Supervised Models for Audio Deepfake Detection
von: Pan, Zihan, et al.
Veröffentlicht: (2025)
von: Pan, Zihan, et al.
Veröffentlicht: (2025)
A Survey on Speech Deepfake Detection
von: Li, Menglu, et al.
Veröffentlicht: (2024)
von: Li, Menglu, et al.
Veröffentlicht: (2024)
DiffSSD: A Diffusion-Based Dataset For Speech Forensics
von: Bhagtani, Kratika, et al.
Veröffentlicht: (2024)
von: Bhagtani, Kratika, et al.
Veröffentlicht: (2024)
The Affective Bridge: Preserving Speech Representations while Enhancing Deepfake Detection vian emotional Constraints
von: Li, Yupei, et al.
Veröffentlicht: (2025)
von: Li, Yupei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Source Verification for Speech Deepfakes
von: Negroni, Viola, et al.
Veröffentlicht: (2025) -
Freeze and Learn: Continual Learning with Selective Freezing for Speech Deepfake Detection
von: Salvi, Davide, et al.
Veröffentlicht: (2024) -
Attention-based Mixture of Experts for Robust Speech Deepfake Detection
von: Negroni, Viola, et al.
Veröffentlicht: (2025) -
Phoneme-Level Analysis for Person-of-Interest Speech Deepfake Detection
von: Salvi, Davide, et al.
Veröffentlicht: (2025) -
Forensic Similarity for Speech Deepfakes
von: Negroni, Viola, et al.
Veröffentlicht: (2025)