Audio Language Model for Deepfake Detection Grounded in Acoustic Chain-of-Thought
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Runkun, Fang, Yixiong, Chang, Pengyu, Li, Yuante, Baali, Massa, Raj, Bhiksha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DELULU: Discriminative Embedding Learning Using Latent Units for Speaker-Aware Self-Trained Speech Foundational Model
von: Baali, Massa, et al.
Veröffentlicht: (2025)
von: Baali, Massa, et al.
Veröffentlicht: (2025)
What and When to Learn: CURriculum Ranking Loss for Large-Scale Speaker Verification
von: Baali, Massa, et al.
Veröffentlicht: (2026)
von: Baali, Massa, et al.
Veröffentlicht: (2026)
Improving Speaker Representations Using Contrastive Losses on Multi-scale Features
von: Dixit, Satvik, et al.
Veröffentlicht: (2024)
von: Dixit, Satvik, et al.
Veröffentlicht: (2024)
PDAF: A Phonetic Debiasing Attention Framework For Speaker Verification
von: Baali, Massa, et al.
Veröffentlicht: (2024)
von: Baali, Massa, et al.
Veröffentlicht: (2024)
SVeritas: Benchmark for Robust Speaker Verification under Diverse Conditions
von: Baali, Massa, et al.
Veröffentlicht: (2025)
von: Baali, Massa, et al.
Veröffentlicht: (2025)
CAARMA: Class Augmentation with Adversarial Mixup Regularization
von: Baali, Massa, et al.
Veröffentlicht: (2025)
von: Baali, Massa, et al.
Veröffentlicht: (2025)
CoLMbo: Speaker Language Model for Descriptive Profiling
von: Baali, Massa, et al.
Veröffentlicht: (2025)
von: Baali, Massa, et al.
Veröffentlicht: (2025)
Domain Adaptation for Contrastive Audio-Language Models
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
MACE: Leveraging Audio for Evaluating Audio Captioning Systems
von: Dixit, Satvik, et al.
Veröffentlicht: (2024)
von: Dixit, Satvik, et al.
Veröffentlicht: (2024)
Revisiting Acoustic Features for Robust ASR
von: Shah, Muhammad A., et al.
Veröffentlicht: (2024)
von: Shah, Muhammad A., et al.
Veröffentlicht: (2024)
Emotion and Acoustics Should Agree: Cross-Level Inconsistency Analysis for Audio Deepfake Detection
von: Zhang, Jinhua, et al.
Veröffentlicht: (2026)
von: Zhang, Jinhua, et al.
Veröffentlicht: (2026)
PAM: Prompting Audio-Language Models for Audio Quality Assessment
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
Thinking with Sound: Audio Chain-of-Thought Enables Multimodal Reasoning in Large Audio-Language Models
von: Xiong, Zhen, et al.
Veröffentlicht: (2025)
von: Xiong, Zhen, et al.
Veröffentlicht: (2025)
Towards Explicit Acoustic Evidence Perception in Audio LLMs for Speech Deepfake Detection
von: Guo, Xiaoxuan, et al.
Veröffentlicht: (2026)
von: Guo, Xiaoxuan, et al.
Veröffentlicht: (2026)
Diffusion Reconstruction towards Generalizable Audio Deepfake Detection
von: Cheng, Bo, et al.
Veröffentlicht: (2026)
von: Cheng, Bo, et al.
Veröffentlicht: (2026)
Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model
von: Ma, Ziyang, et al.
Veröffentlicht: (2025)
von: Ma, Ziyang, et al.
Veröffentlicht: (2025)
Continual Audio Deepfake Detection via Universal Adversarial Perturbation
von: Li, Wangjie, et al.
Veröffentlicht: (2025)
von: Li, Wangjie, et al.
Veröffentlicht: (2025)
Pitch Imperfect: Detecting Audio Deepfakes Through Acoustic Prosodic Analysis
von: Warren, Kevin, et al.
Veröffentlicht: (2025)
von: Warren, Kevin, et al.
Veröffentlicht: (2025)
Does Current Deepfake Audio Detection Model Effectively Detect ALM-based Deepfake Audio?
von: Xie, Yuankun, et al.
Veröffentlicht: (2024)
von: Xie, Yuankun, et al.
Veröffentlicht: (2024)
Audio-DeepThinker: Progressive Reasoning-Aware Reinforcement Learning for High-Quality Chain-of-Thought Emergence in Audio Language Models
von: He, Xiang, et al.
Veröffentlicht: (2026)
von: He, Xiang, et al.
Veröffentlicht: (2026)
Generalizable Detection of Audio Deepfakes
von: Lopez, Jose A., et al.
Veröffentlicht: (2025)
von: Lopez, Jose A., et al.
Veröffentlicht: (2025)
Detection of Deepfake Environmental Audio
von: Ouajdi, Hafsa, et al.
Veröffentlicht: (2024)
von: Ouajdi, Hafsa, et al.
Veröffentlicht: (2024)
Human Voice is Unique
von: Singh, Rita, et al.
Veröffentlicht: (2025)
von: Singh, Rita, et al.
Veröffentlicht: (2025)
Deciphering GunType Hierarchy through Acoustic Analysis of Gunshot Recordings
von: Shah, Ankit, et al.
Veröffentlicht: (2025)
von: Shah, Ankit, et al.
Veröffentlicht: (2025)
Audio Entailment: Assessing Deductive Reasoning for Audio Understanding
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
ALLM4ADD: Unlocking the Capabilities of Audio Large Language Models for Audio Deepfake Detection
von: Gu, Hao, et al.
Veröffentlicht: (2025)
von: Gu, Hao, et al.
Veröffentlicht: (2025)
ESPnet-Codec: Comprehensive Training and Evaluation of Neural Codecs for Audio, Music, and Speech
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
DFALLM: Achieving Generalizable Multitask Deepfake Detection by Optimizing Audio LLM Components
von: Li, Yupei, et al.
Veröffentlicht: (2025)
von: Li, Yupei, et al.
Veröffentlicht: (2025)
Efficient Autoregressive Audio Modeling via Next-Scale Prediction
von: Qiu, Kai, et al.
Veröffentlicht: (2024)
von: Qiu, Kai, et al.
Veröffentlicht: (2024)
Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection
von: Hao, Yunqi, et al.
Veröffentlicht: (2025)
von: Hao, Yunqi, et al.
Veröffentlicht: (2025)
MixFake: Benchmarking and Enhancing Audio Deepfake Detection in Diverse Real-world Mixed Audio
von: Li, Qingcao, et al.
Veröffentlicht: (2026)
von: Li, Qingcao, et al.
Veröffentlicht: (2026)
ERF-BA-TFD+: A Multimodal Model for Audio-Visual Deepfake Detection
von: Zhang, Xin, et al.
Veröffentlicht: (2025)
von: Zhang, Xin, et al.
Veröffentlicht: (2025)
Transferable Adversarial Attacks on Audio Deepfake Detection
von: Farooq, Muhammad Umar, et al.
Veröffentlicht: (2025)
von: Farooq, Muhammad Umar, et al.
Veröffentlicht: (2025)
How to Label Resynthesized Audio: The Dual Role of Neural Audio Codecs in Audio Deepfake Detection
von: Xiao, Yixuan, et al.
Veröffentlicht: (2026)
von: Xiao, Yixuan, et al.
Veröffentlicht: (2026)
TwinShift: Benchmarking Audio Deepfake Detection across Synthesizer and Speaker Shifts
von: Hong, Jiyoung, et al.
Veröffentlicht: (2025)
von: Hong, Jiyoung, et al.
Veröffentlicht: (2025)
Did You Hear That? Introducing AADG: A Framework for Generating Benchmark Data in Audio Anomaly Detection
von: Raghavan, Ksheeraja, et al.
Veröffentlicht: (2024)
von: Raghavan, Ksheeraja, et al.
Veröffentlicht: (2024)
Retrieval-Augmented Audio Deepfake Detection
von: Kang, Zuheng, et al.
Veröffentlicht: (2024)
von: Kang, Zuheng, et al.
Veröffentlicht: (2024)
Does Audio Deepfake Detection Generalize?
von: Müller, Nicolas M., et al.
Veröffentlicht: (2022)
von: Müller, Nicolas M., et al.
Veröffentlicht: (2022)
Comprehensive Layer-wise Analysis of SSL Models for Audio Deepfake Detection
von: Kheir, Yassine El, et al.
Veröffentlicht: (2025)
von: Kheir, Yassine El, et al.
Veröffentlicht: (2025)
Measuring the Robustness of Audio Deepfake Detectors
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DELULU: Discriminative Embedding Learning Using Latent Units for Speaker-Aware Self-Trained Speech Foundational Model
von: Baali, Massa, et al.
Veröffentlicht: (2025) -
What and When to Learn: CURriculum Ranking Loss for Large-Scale Speaker Verification
von: Baali, Massa, et al.
Veröffentlicht: (2026) -
Improving Speaker Representations Using Contrastive Losses on Multi-scale Features
von: Dixit, Satvik, et al.
Veröffentlicht: (2024) -
PDAF: A Phonetic Debiasing Attention Framework For Speaker Verification
von: Baali, Massa, et al.
Veröffentlicht: (2024) -
SVeritas: Benchmark for Robust Speaker Verification under Diverse Conditions
von: Baali, Massa, et al.
Veröffentlicht: (2025)