Rapidly Adapting to New Voice Spoofing: Few-Shot Detection of Synthesized Speech Under Distribution Shifts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Garg, Ashi, Cai, Zexin, Xinyuan, Henry Li, García-Perera, Leibny Paola, Duh, Kevin, Khudanpur, Sanjeev, Wiesner, Matthew, Andrews, Nicholas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GenVC: Self-Supervised Zero-Shot Voice Conversion
von: Cai, Zexin, et al.
Veröffentlicht: (2025)
von: Cai, Zexin, et al.
Veröffentlicht: (2025)
ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts
von: Garg, Ashi, et al.
Veröffentlicht: (2025)
von: Garg, Ashi, et al.
Veröffentlicht: (2025)
HLTCOE JHU Submission to the Voice Privacy Challenge 2024
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2024)
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2024)
Scalable Controllable Accented TTS
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2025)
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2025)
Privacy versus Emotion Preservation Trade-offs in Emotion-Preserving Speaker Anonymization
von: Cai, Zexin, et al.
Veröffentlicht: (2024)
von: Cai, Zexin, et al.
Veröffentlicht: (2024)
Universal Speech Content Factorization
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2026)
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2026)
Integrated Spoofing-Robust Automatic Speaker Verification via a Three-Class Formulation and LLR
von: Tan, Kai, et al.
Veröffentlicht: (2026)
von: Tan, Kai, et al.
Veröffentlicht: (2026)
Can LLMs Help Localize Fake Words in Partially Fake Speech?
von: Zhang, Lin, et al.
Veröffentlicht: (2026)
von: Zhang, Lin, et al.
Veröffentlicht: (2026)
On Speaker Attribution with SURT
von: Raj, Desh, et al.
Veröffentlicht: (2024)
von: Raj, Desh, et al.
Veröffentlicht: (2024)
HENT-SRT: Hierarchical Efficient Neural Transducer with Self-Distillation for Joint Speech Recognition and Translation
von: Hussein, Amir, et al.
Veröffentlicht: (2025)
von: Hussein, Amir, et al.
Veröffentlicht: (2025)
CASPER: A Large Scale Spontaneous Speech Dataset
von: Xiao, Cihan, et al.
Veröffentlicht: (2025)
von: Xiao, Cihan, et al.
Veröffentlicht: (2025)
Modeling Overlapped Speech with Shuffles
von: Wiesner, Matthew, et al.
Veröffentlicht: (2026)
von: Wiesner, Matthew, et al.
Veröffentlicht: (2026)
Improving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation
von: Huang, Ruizhe, et al.
Veröffentlicht: (2024)
von: Huang, Ruizhe, et al.
Veröffentlicht: (2024)
Target Speaker ASR with Whisper
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
Clean Label Attacks against SLU Systems
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2024)
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2024)
DiCoW: Diarization-Conditioned Whisper for Target Speaker Automatic Speech Recognition
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
Unsupervised Speech Enhancement using Data-defined Priors
von: Klement, Dominik, et al.
Veröffentlicht: (2025)
von: Klement, Dominik, et al.
Veröffentlicht: (2025)
Towards Explainable Spoofed Speech Attribution and Detection:a Probabilistic Approach for Characterizing Speech Synthesizer Components
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2025)
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2025)
Acoustic modeling for Overlapping Speech Recognition: JHU Chime-5 Challenge System
von: Manohar, Vimal, et al.
Veröffentlicht: (2024)
von: Manohar, Vimal, et al.
Veröffentlicht: (2024)
DiffAnon: Diffusion-based Prosody Control for Voice Anonymization
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2026)
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2026)
SE-DiCoW: Self-Enrolled Diarization-Conditioned Whisper
von: Polok, Alexander, et al.
Veröffentlicht: (2026)
von: Polok, Alexander, et al.
Veröffentlicht: (2026)
Detecting Spoof Voices in Asian Non-Native Speech: An Indonesian and Thai Case Study
von: Adila, Aulia, et al.
Veröffentlicht: (2024)
von: Adila, Aulia, et al.
Veröffentlicht: (2024)
Dimensionality-Aware Anomaly Detection in Learned Representations of Self-Supervised Speech Models
von: Arcos-Holzinger, Sandra, et al.
Veröffentlicht: (2026)
von: Arcos-Holzinger, Sandra, et al.
Veröffentlicht: (2026)
Building Corpora for Single-Channel Speech Separation Across Multiple Domains
von: Maciejewski, Matthew, et al.
Veröffentlicht: (2018)
von: Maciejewski, Matthew, et al.
Veröffentlicht: (2018)
Adversarial Attacks and Defenses for Speech Recognition Systems
von: Żelasko, Piotr, et al.
Veröffentlicht: (2021)
von: Żelasko, Piotr, et al.
Veröffentlicht: (2021)
Aligning Speech to Languages to Enhance Code-switching Speech Recognition
von: Liu, Hexin, et al.
Veröffentlicht: (2024)
von: Liu, Hexin, et al.
Veröffentlicht: (2024)
SpatialEmb: Extract and Encode Spatial Information for 1-Stage Multi-channel Multi-speaker ASR on Arbitrary Microphone Arrays
von: Shao, Yiwen, et al.
Veröffentlicht: (2026)
von: Shao, Yiwen, et al.
Veröffentlicht: (2026)
SAV-SE: Scene-aware Audio-Visual Speech Enhancement with Selective State Space Model
von: Qian, Xinyuan, et al.
Veröffentlicht: (2024)
von: Qian, Xinyuan, et al.
Veröffentlicht: (2024)
SF-Speech: Straightened Flow for Zero-Shot Voice Clone
von: Li, Xuyuan, et al.
Veröffentlicht: (2024)
von: Li, Xuyuan, et al.
Veröffentlicht: (2024)
VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
von: Anastassiou, Philip, et al.
Veröffentlicht: (2024)
von: Anastassiou, Philip, et al.
Veröffentlicht: (2024)
Where Does Speech Enhancement Adapt? Probing Study Under Controlled Degradation
von: Amar, Yair, et al.
Veröffentlicht: (2025)
von: Amar, Yair, et al.
Veröffentlicht: (2025)
Spoof Diarization: "What Spoofed When" in Partially Spoofed Audio
von: Zhang, Lin, et al.
Veröffentlicht: (2024)
von: Zhang, Lin, et al.
Veröffentlicht: (2024)
An Empirical Study on Channel Effects for Synthetic Voice Spoofing Countermeasure Systems
von: Zhang, You, et al.
Veröffentlicht: (2021)
von: Zhang, You, et al.
Veröffentlicht: (2021)
SpoofCeleb: Speech Deepfake Detection and SASV In The Wild
von: Jung, Jee-weon, et al.
Veröffentlicht: (2024)
von: Jung, Jee-weon, et al.
Veröffentlicht: (2024)
SwanVoice: Expressive Long-Form Zero-Shot Speech Synthesis for Both Monologue and Dialogue
von: Li, Ruiqi, et al.
Veröffentlicht: (2026)
von: Li, Ruiqi, et al.
Veröffentlicht: (2026)
The Universal Personalizer: Few-Shot Dysarthric Speech Recognition via Meta-Learning
von: Agarwal, Dhruuv, et al.
Veröffentlicht: (2025)
von: Agarwal, Dhruuv, et al.
Veröffentlicht: (2025)
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
von: Zhu, Han, et al.
Veröffentlicht: (2025)
von: Zhu, Han, et al.
Veröffentlicht: (2025)
Everyone-Can-Sing: Zero-Shot Singing Voice Synthesis and Conversion with Speech Reference
von: Dai, Shuqi, et al.
Veröffentlicht: (2025)
von: Dai, Shuqi, et al.
Veröffentlicht: (2025)
Few-Shot Speech Deepfake Detection Adaptation with Gaussian Processes
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
An Explainable Probabilistic Attribute Embedding Approach for Spoofed Speech Characterization
von: Chhibber, Manasi, et al.
Veröffentlicht: (2024)
von: Chhibber, Manasi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
GenVC: Self-Supervised Zero-Shot Voice Conversion
von: Cai, Zexin, et al.
Veröffentlicht: (2025) -
ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts
von: Garg, Ashi, et al.
Veröffentlicht: (2025) -
HLTCOE JHU Submission to the Voice Privacy Challenge 2024
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2024) -
Scalable Controllable Accented TTS
von: Xinyuan, Henry Li, et al.
Veröffentlicht: (2025) -
Privacy versus Emotion Preservation Trade-offs in Emotion-Preserving Speaker Anonymization
von: Cai, Zexin, et al.
Veröffentlicht: (2024)