The Affective Bridge: Preserving Speech Representations while Enhancing Deepfake Detection vian emotional Constraints
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yupei, Lyu, Chenyang, Wang, Longyue, Luo, Weihua, Zhang, Kaifu, Schuller, Björn W. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Efficiency and Performance in Deepfake Audio Detection through Neuron-level Dropin & Neuroplasticity Mechanisms
von: Li, Yupei, et al.
Veröffentlicht: (2026)
von: Li, Yupei, et al.
Veröffentlicht: (2026)
From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview
von: Li, Yupei, et al.
Veröffentlicht: (2024)
von: Li, Yupei, et al.
Veröffentlicht: (2024)
LongSpeech: A Scalable Benchmark for Transcription, Translation and Understanding in Long Speech
von: Yang, Fei, et al.
Veröffentlicht: (2026)
von: Yang, Fei, et al.
Veröffentlicht: (2026)
DFALLM: Achieving Generalizable Multitask Deepfake Detection by Optimizing Audio LLM Components
von: Li, Yupei, et al.
Veröffentlicht: (2025)
von: Li, Yupei, et al.
Veröffentlicht: (2025)
Speech-XL: Towards Long-Form Speech Understanding in Large Speech Language Models
von: Sun, Haoqin, et al.
Veröffentlicht: (2026)
von: Sun, Haoqin, et al.
Veröffentlicht: (2026)
Marco-ASR: A Principled and Metric-Driven Framework for Fine-Tuning Large-Scale ASR Models for Domain Adaptation
von: Ni, Xuanfan, et al.
Veröffentlicht: (2025)
von: Ni, Xuanfan, et al.
Veröffentlicht: (2025)
GatedxLSTM: A Multimodal Affective Computing Approach for Emotion Recognition in Conversations
von: Li, Yupei, et al.
Veröffentlicht: (2025)
von: Li, Yupei, et al.
Veröffentlicht: (2025)
M6: Multi-generator, Multi-domain, Multi-lingual and cultural, Multi-genres, Multi-instrument Machine-Generated Music Detection Databases
von: Li, Yupei, et al.
Veröffentlicht: (2024)
von: Li, Yupei, et al.
Veröffentlicht: (2024)
Explainable Detection of Machine Generated Music and Early Systematic Evaluation
von: Li, Yupei, et al.
Veröffentlicht: (2024)
von: Li, Yupei, et al.
Veröffentlicht: (2024)
MECap-R1: Emotion-aware Policy with Reinforcement Learning for Multimodal Emotion Captioning
von: Sun, Haoqin, et al.
Veröffentlicht: (2025)
von: Sun, Haoqin, et al.
Veröffentlicht: (2025)
Quantifying Dimensional Independence in Speech: An Information-Theoretic Framework for Disentangled Representation Learning
von: Kashyap, Bipasha, et al.
Veröffentlicht: (2026)
von: Kashyap, Bipasha, et al.
Veröffentlicht: (2026)
Marco-Voice Technical Report
von: Tian, Fengping, et al.
Veröffentlicht: (2025)
von: Tian, Fengping, et al.
Veröffentlicht: (2025)
Charting 15 years of progress in deep learning for speech emotion recognition: A replication study
von: Triantafyllopoulos, Andreas, et al.
Veröffentlicht: (2025)
von: Triantafyllopoulos, Andreas, et al.
Veröffentlicht: (2025)
SmoothCLAP: Soft-Target Enhanced Contrastive Language\--Audio Pretraining for Affective Computing
von: Jing, Xin, et al.
Veröffentlicht: (2026)
von: Jing, Xin, et al.
Veröffentlicht: (2026)
Abusive Speech Detection in Indic Languages Using Acoustic Features
von: Spiesberger, Anika A., et al.
Veröffentlicht: (2024)
von: Spiesberger, Anika A., et al.
Veröffentlicht: (2024)
AffectSpeech: A Large-Scale Emotional Speech Dataset with Fine-Grained Textual Descriptions for Speech Emotion Captioning and Synthesis
von: Qi, Tianhua, et al.
Veröffentlicht: (2026)
von: Qi, Tianhua, et al.
Veröffentlicht: (2026)
Enhancing Emotional Text-to-Speech Controllability with Natural Language Guidance through Contrastive Learning and Diffusion Models
von: Jing, Xin, et al.
Veröffentlicht: (2024)
von: Jing, Xin, et al.
Veröffentlicht: (2024)
Representation Learning with Parameterised Quantum Circuits for Advancing Speech Emotion Recognition
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2025)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2025)
DOTA-ME-CS: Daily Oriented Text Audio-Mandarin English-Code Switching Dataset
von: Li, Yupei, et al.
Veröffentlicht: (2025)
von: Li, Yupei, et al.
Veröffentlicht: (2025)
EmoSURA: Towards Accurate Evaluation of Detailed and Long-Context Emotional Speech Captions
von: Jing, Xin, et al.
Veröffentlicht: (2026)
von: Jing, Xin, et al.
Veröffentlicht: (2026)
RTCFake: Speech Deepfake Detection in Real-Time Communication
von: Xue, Jun, et al.
Veröffentlicht: (2026)
von: Xue, Jun, et al.
Veröffentlicht: (2026)
ProsodyFM: Unsupervised Phrasing and Intonation Control for Intelligible Speech Synthesis
von: He, Xiangheng, et al.
Veröffentlicht: (2024)
von: He, Xiangheng, et al.
Veröffentlicht: (2024)
Wav2Small: Distilling Wav2Vec2 to 72K parameters for Low-Resource Speech emotion recognition
von: Kounadis-Bastian, Dionyssos, et al.
Veröffentlicht: (2024)
von: Kounadis-Bastian, Dionyssos, et al.
Veröffentlicht: (2024)
Profiling the Voice: Speaker-Specific Phoneme Fingerprinting for Speech Deepfake Detection
von: Xue, Jun, et al.
Veröffentlicht: (2026)
von: Xue, Jun, et al.
Veröffentlicht: (2026)
Enhancing Speech Emotion Recognition Through Differentiable Architecture Search
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2023)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2023)
ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Evaluation Plan
von: Zhang, Xueping, et al.
Veröffentlicht: (2026)
von: Zhang, Xueping, et al.
Veröffentlicht: (2026)
HAFFormer: A Hierarchical Attention-Free Framework for Alzheimer's Disease Detection From Spontaneous Speech
von: Dong, Zhongren, et al.
Veröffentlicht: (2024)
von: Dong, Zhongren, et al.
Veröffentlicht: (2024)
A Survey on Speech Deepfake Detection
von: Li, Menglu, et al.
Veröffentlicht: (2024)
von: Li, Menglu, et al.
Veröffentlicht: (2024)
HQ-MPSD: A Multilingual Artifact-Controlled Benchmark for Partial Deepfake Speech Detection
von: Li, Menglu, et al.
Veröffentlicht: (2025)
von: Li, Menglu, et al.
Veröffentlicht: (2025)
Intelligent Cardiac Auscultation for Murmur Detection via Parallel-Attentive Models with Uncertainty Estimation
von: Zhang, Zixing, et al.
Veröffentlicht: (2024)
von: Zhang, Zixing, et al.
Veröffentlicht: (2024)
Fake Speech Wild: Detecting Deepfake Speech on Social Media Platform
von: Xie, Yuankun, et al.
Veröffentlicht: (2025)
von: Xie, Yuankun, et al.
Veröffentlicht: (2025)
Can Large Language Models Aid in Annotating Speech Emotional Data? Uncovering New Frontiers
von: Latif, Siddique, et al.
Veröffentlicht: (2023)
von: Latif, Siddique, et al.
Veröffentlicht: (2023)
A Comparative Study on Proactive and Passive Detection of Deepfake Speech
von: Wu, Chia-Hua, et al.
Veröffentlicht: (2025)
von: Wu, Chia-Hua, et al.
Veröffentlicht: (2025)
Attention-based Mixture of Experts for Robust Speech Deepfake Detection
von: Negroni, Viola, et al.
Veröffentlicht: (2025)
von: Negroni, Viola, et al.
Veröffentlicht: (2025)
Generalizable Speech Deepfake Detection via Information Bottleneck Enhanced Adversarial Alignment
von: Huang, Pu, et al.
Veröffentlicht: (2025)
von: Huang, Pu, et al.
Veröffentlicht: (2025)
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2022)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2022)
Forensic Similarity for Speech Deepfakes
von: Negroni, Viola, et al.
Veröffentlicht: (2025)
von: Negroni, Viola, et al.
Veröffentlicht: (2025)
Unmasking Deepfakes: Leveraging Augmentations and Features Variability for Deepfake Speech Detection
von: Rimon, Inbal, et al.
Veröffentlicht: (2025)
von: Rimon, Inbal, et al.
Veröffentlicht: (2025)
SEA-Spoof: Bridging The Gap in Multilingual Audio Deepfake Detection for South-East Asian
von: Wu, Jinyang, et al.
Veröffentlicht: (2025)
von: Wu, Jinyang, et al.
Veröffentlicht: (2025)
How Well Do Current Speech Deepfake Detection Methods Generalize to the Real World?
von: Li, Daixian, et al.
Veröffentlicht: (2026)
von: Li, Daixian, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Enhancing Efficiency and Performance in Deepfake Audio Detection through Neuron-level Dropin & Neuroplasticity Mechanisms
von: Li, Yupei, et al.
Veröffentlicht: (2026) -
From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview
von: Li, Yupei, et al.
Veröffentlicht: (2024) -
LongSpeech: A Scalable Benchmark for Transcription, Translation and Understanding in Long Speech
von: Yang, Fei, et al.
Veröffentlicht: (2026) -
DFALLM: Achieving Generalizable Multitask Deepfake Detection by Optimizing Audio LLM Components
von: Li, Yupei, et al.
Veröffentlicht: (2025) -
Speech-XL: Towards Long-Form Speech Understanding in Large Speech Language Models
von: Sun, Haoqin, et al.
Veröffentlicht: (2026)