Can LLMs Help Localize Fake Words in Partially Fake Speech?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Lin, Thebaud, Thomas, Cai, Zexin, Khudanpur, Sanjeev, Povey, Daniel, García-Perera, Leibny Paola, Wiesner, Matthew, Andrews, Nicholas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Universal Speech Content Factorization
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2026)
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2026)
On Speaker Attribution with SURT
von: Raj, Desh, et al.
Veröffentlicht: (2024)
von: Raj, Desh, et al.
Veröffentlicht: (2024)
ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts
von: Garg, Ashi, et al.
Veröffentlicht: (2025)
von: Garg, Ashi, et al.
Veröffentlicht: (2025)
Integrated Spoofing-Robust Automatic Speaker Verification via a Three-Class Formulation and LLR
von: Tan, Kai, et al.
Veröffentlicht: (2026)
von: Tan, Kai, et al.
Veröffentlicht: (2026)
Rapidly Adapting to New Voice Spoofing: Few-Shot Detection of Synthesized Speech Under Distribution Shifts
von: Garg, Ashi, et al.
Veröffentlicht: (2025)
von: Garg, Ashi, et al.
Veröffentlicht: (2025)
HENT-SRT: Hierarchical Efficient Neural Transducer with Self-Distillation for Joint Speech Recognition and Translation
von: Hussein, Amir, et al.
Veröffentlicht: (2025)
von: Hussein, Amir, et al.
Veröffentlicht: (2025)
Robust Localization of Partially Fake Speech: Metrics and Out-of-Domain Evaluation
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2025)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2025)
Scalable Controllable Accented TTS
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2025)
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2025)
AntiDeepFake: AI for Deep Fake Speech Recognition
von: Togootogtokh, Enkhtogtokh, et al.
Veröffentlicht: (2024)
von: Togootogtokh, Enkhtogtokh, et al.
Veröffentlicht: (2024)
GenVC: Self-Supervised Zero-Shot Voice Conversion
von: Cai, Zexin, et al.
Veröffentlicht: (2025)
von: Cai, Zexin, et al.
Veröffentlicht: (2025)
Privacy versus Emotion Preservation Trade-offs in Emotion-Preserving Speaker Anonymization
von: Cai, Zexin, et al.
Veröffentlicht: (2024)
von: Cai, Zexin, et al.
Veröffentlicht: (2024)
HLTCOE JHU Submission to the Voice Privacy Challenge 2024
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2024)
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2024)
LlamaPartialSpoof: An LLM-Driven Fake Speech Dataset Simulating Disinformation Generation
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
Target Speaker ASR with Whisper
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
CASPER: A Large Scale Spontaneous Speech Dataset
von: Xiao, Cihan, et al.
Veröffentlicht: (2025)
von: Xiao, Cihan, et al.
Veröffentlicht: (2025)
DiCoW: Diarization-Conditioned Whisper for Target Speaker Automatic Speech Recognition
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
Modeling Overlapped Speech with Shuffles
von: Wiesner, Matthew, et al.
Veröffentlicht: (2026)
von: Wiesner, Matthew, et al.
Veröffentlicht: (2026)
Frequency-mix Knowledge Distillation for Fake Speech Detection
von: Fan, Cunhang, et al.
Veröffentlicht: (2024)
von: Fan, Cunhang, et al.
Veröffentlicht: (2024)
EmoFake: An Initial Dataset for Emotion Fake Audio Detection
von: Zhao, Yan, et al.
Veröffentlicht: (2022)
von: Zhao, Yan, et al.
Veröffentlicht: (2022)
Unsupervised Speech Enhancement using Data-defined Priors
von: Klement, Dominik, et al.
Veröffentlicht: (2025)
von: Klement, Dominik, et al.
Veröffentlicht: (2025)
Analyzing the Impact of Splicing Artifacts in Partially Fake Speech Signals
von: Negroni, Viola, et al.
Veröffentlicht: (2024)
von: Negroni, Viola, et al.
Veröffentlicht: (2024)
Spatial Reconstructed Local Attention Res2Net with F0 Subband for Fake Speech Detection
von: Fan, Cunhang, et al.
Veröffentlicht: (2023)
von: Fan, Cunhang, et al.
Veröffentlicht: (2023)
How Does Instrumental Music Help SingFake Detection?
von: Chen, Xuanjun, et al.
Veröffentlicht: (2025)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2025)
CodecFake+: A Large-Scale Neural Audio Codec-Based Deepfake Speech Dataset
von: Chen, Xuanjun, et al.
Veröffentlicht: (2025)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2025)
WeDefense: A Toolkit to Defend Against Fake Audio
von: Zhang, Lin, et al.
Veröffentlicht: (2026)
von: Zhang, Lin, et al.
Veröffentlicht: (2026)
F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching
von: Chen, Yushen, et al.
Veröffentlicht: (2024)
von: Chen, Yushen, et al.
Veröffentlicht: (2024)
Speech DF Arena: A Leaderboard for Speech DeepFake Detection Models
von: Dowerah, Sandipana, et al.
Veröffentlicht: (2025)
von: Dowerah, Sandipana, et al.
Veröffentlicht: (2025)
SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods
von: Huang, Wen, et al.
Veröffentlicht: (2025)
von: Huang, Wen, et al.
Veröffentlicht: (2025)
Benchmarking Fake Voice Detection in the Fake Voice Generation Arms Race
von: Mao, Xutao, et al.
Veröffentlicht: (2025)
von: Mao, Xutao, et al.
Veröffentlicht: (2025)
An RFP dataset for Real, Fake, and Partially fake audio detection
von: AlAli, Abdulazeez, et al.
Veröffentlicht: (2024)
von: AlAli, Abdulazeez, et al.
Veröffentlicht: (2024)
Are audio DeepFake detection models polyglots?
von: Marek, Bartłomiej, et al.
Veröffentlicht: (2024)
von: Marek, Bartłomiej, et al.
Veröffentlicht: (2024)
SceneFake: An Initial Dataset and Benchmarks for Scene Fake Audio Detection
von: Yi, Jiangyan, et al.
Veröffentlicht: (2022)
von: Yi, Jiangyan, et al.
Veröffentlicht: (2022)
CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
Generalized Fake Audio Detection via Deep Stable Learning
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
Adversarial Attacks and Defenses for Speech Recognition Systems
von: Żelasko, Piotr, et al.
Veröffentlicht: (2021)
von: Żelasko, Piotr, et al.
Veröffentlicht: (2021)
A Noval Feature via Color Quantisation for Fake Audio Detection
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyong, et al.
Veröffentlicht: (2024)
Room Impulse Responses help attackers to evade Deep Fake Detection
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
EchoFake: A Replay-Aware Dataset for Practical Speech Deepfake Detection
von: Zhang, Tong, et al.
Veröffentlicht: (2025)
von: Zhang, Tong, et al.
Veröffentlicht: (2025)
The CHiME-8 DASR Challenge for Generalizable and Array Agnostic Distant Automatic Speech Recognition and Diarization
von: Cornell, Samuele, et al.
Veröffentlicht: (2024)
von: Cornell, Samuele, et al.
Veröffentlicht: (2024)
SpatialEmb: Extract and Encode Spatial Information for 1-Stage Multi-channel Multi-speaker ASR on Arbitrary Microphone Arrays
von: Shao, Yiwen, et al.
Veröffentlicht: (2026)
von: Shao, Yiwen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Universal Speech Content Factorization
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2026) -
On Speaker Attribution with SURT
von: Raj, Desh, et al.
Veröffentlicht: (2024) -
ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts
von: Garg, Ashi, et al.
Veröffentlicht: (2025) -
Integrated Spoofing-Robust Automatic Speaker Verification via a Three-Class Formulation and LLR
von: Tan, Kai, et al.
Veröffentlicht: (2026) -
Rapidly Adapting to New Voice Spoofing: Few-Shot Detection of Synthesized Speech Under Distribution Shifts
von: Garg, Ashi, et al.
Veröffentlicht: (2025)