A New Approach to Voice Authenticity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Müller, Nicolas M., Kawa, Piotr, Hu, Shen, Neu, Matthias, Williams, Jennifer, Sperl, Philip, Böttinger, Konstantin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Harder or Different? Understanding Generalization of Audio Deepfake Detection
von: Müller, Nicolas M., et al.
Veröffentlicht: (2024)
von: Müller, Nicolas M., et al.
Veröffentlicht: (2024)
DeePen: Penetration Testing for Audio Deepfake Detection
von: Müller, Nicolas, et al.
Veröffentlicht: (2025)
von: Müller, Nicolas, et al.
Veröffentlicht: (2025)
MLAAD: The Multi-Language Audio Anti-Spoofing Dataset
von: Müller, Nicolas M., et al.
Veröffentlicht: (2024)
von: Müller, Nicolas M., et al.
Veröffentlicht: (2024)
Replay Attacks Against Audio Deepfake Detection
von: Müller, Nicolas, et al.
Veröffentlicht: (2025)
von: Müller, Nicolas, et al.
Veröffentlicht: (2025)
Physics-Guided Deepfake Detection for Voice Authentication Systems
von: Mohammadi, Alireza, et al.
Veröffentlicht: (2025)
von: Mohammadi, Alireza, et al.
Veröffentlicht: (2025)
Are audio DeepFake detection models polyglots?
von: Marek, Bartłomiej, et al.
Veröffentlicht: (2024)
von: Marek, Bartłomiej, et al.
Veröffentlicht: (2024)
VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
von: Anastassiou, Philip, et al.
Veröffentlicht: (2024)
von: Anastassiou, Philip, et al.
Veröffentlicht: (2024)
Fast-VGAN: Lightweight Voice Conversion with Explicit Control of F0 and Duration Parameters
von: Abrassart, Mathilde, et al.
Veröffentlicht: (2025)
von: Abrassart, Mathilde, et al.
Veröffentlicht: (2025)
Does Audio Deepfake Detection Generalize?
von: Müller, Nicolas M., et al.
Veröffentlicht: (2022)
von: Müller, Nicolas M., et al.
Veröffentlicht: (2022)
Human Perception of Audio Deepfakes
von: Müller, Nicolas M., et al.
Veröffentlicht: (2021)
von: Müller, Nicolas M., et al.
Veröffentlicht: (2021)
VoiceGRPO: Modern MoE Transformers with Group Relative Policy Optimization GRPO for AI Voice Health Care Applications on Voice Pathology Detection
von: Togootogtokh, Enkhtogtokh, et al.
Veröffentlicht: (2025)
von: Togootogtokh, Enkhtogtokh, et al.
Veröffentlicht: (2025)
Voice Cloning: Comprehensive Survey
von: Azzuni, Hussam, et al.
Veröffentlicht: (2025)
von: Azzuni, Hussam, et al.
Veröffentlicht: (2025)
Voice Attribute Editing with Text Prompt
von: Sheng, Zhengyan, et al.
Veröffentlicht: (2024)
von: Sheng, Zhengyan, et al.
Veröffentlicht: (2024)
X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning
von: Xu, Rixi, et al.
Veröffentlicht: (2026)
von: Xu, Rixi, et al.
Veröffentlicht: (2026)
MFAAN: Unveiling Audio Deepfakes with a Multi-Feature Authenticity Network
von: Krishnan, Karthik Sivarama, et al.
Veröffentlicht: (2023)
von: Krishnan, Karthik Sivarama, et al.
Veröffentlicht: (2023)
Easy, Interpretable, Effective: openSMILE for voice deepfake detection
von: Pascu, Octavian, et al.
Veröffentlicht: (2024)
von: Pascu, Octavian, et al.
Veröffentlicht: (2024)
Echoes: A semantically-aligned music deepfake detection dataset
von: Pascu, Octavian, et al.
Veröffentlicht: (2026)
von: Pascu, Octavian, et al.
Veröffentlicht: (2026)
CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
von: Du, Zhihao, et al.
Veröffentlicht: (2024)
von: Du, Zhihao, et al.
Veröffentlicht: (2024)
SingFake: Singing Voice Deepfake Detection
von: Zang, Yongyi, et al.
Veröffentlicht: (2023)
von: Zang, Yongyi, et al.
Veröffentlicht: (2023)
Deepfake Detection of Singing Voices With Whisper Encodings
von: Sharma, Falguni, et al.
Veröffentlicht: (2025)
von: Sharma, Falguni, et al.
Veröffentlicht: (2025)
Who is Authentic Speaker
von: Huang, Qiang
Veröffentlicht: (2024)
von: Huang, Qiang
Veröffentlicht: (2024)
LTS-VoiceAgent: A Listen-Think-Speak Framework for Efficient Streaming Voice Interaction via Semantic Triggering and Incremental Reasoning
von: Zou, Wenhao, et al.
Veröffentlicht: (2026)
von: Zou, Wenhao, et al.
Veröffentlicht: (2026)
Voice Communication Analysis in Esports
von: Vinot, Aymeric, et al.
Veröffentlicht: (2024)
von: Vinot, Aymeric, et al.
Veröffentlicht: (2024)
ClapFM-EVC: High-Fidelity and Flexible Emotional Voice Conversion with Dual Control from Natural Language and Speech
von: Pan, Yu, et al.
Veröffentlicht: (2025)
von: Pan, Yu, et al.
Veröffentlicht: (2025)
Pronunciation Deviation Analysis Through Voice Cloning and Acoustic Comparison
von: Valdivia, Andrew, et al.
Veröffentlicht: (2025)
von: Valdivia, Andrew, et al.
Veröffentlicht: (2025)
MulliVC: Multi-lingual Voice Conversion With Cycle Consistency
von: Huang, Jiawei, et al.
Veröffentlicht: (2024)
von: Huang, Jiawei, et al.
Veröffentlicht: (2024)
The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan
von: Sheng, Zhengyan, et al.
Veröffentlicht: (2025)
von: Sheng, Zhengyan, et al.
Veröffentlicht: (2025)
Asynchronous Voice Anonymization Using Adversarial Perturbation On Speaker Embedding
von: Wang, Rui, et al.
Veröffentlicht: (2024)
von: Wang, Rui, et al.
Veröffentlicht: (2024)
RDSinger: Reference-based Diffusion Network for Singing Voice Synthesis
von: Sui, Kehan, et al.
Veröffentlicht: (2024)
von: Sui, Kehan, et al.
Veröffentlicht: (2024)
FunAudioLLM: Voice Understanding and Generation Foundation Models for Natural Interaction Between Humans and LLMs
von: An, Keyu, et al.
Veröffentlicht: (2024)
von: An, Keyu, et al.
Veröffentlicht: (2024)
SPA-SVC: Self-supervised Pitch Augmentation for Singing Voice Conversion
von: Bai, Bingsong, et al.
Veröffentlicht: (2024)
von: Bai, Bingsong, et al.
Veröffentlicht: (2024)
Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement
von: Zhang, Xueyao, et al.
Veröffentlicht: (2025)
von: Zhang, Xueyao, et al.
Veröffentlicht: (2025)
Improving Voice Quality in Speech Anonymization With Just Perception-Informed Losses
von: Ghosh, Suhita, et al.
Veröffentlicht: (2024)
von: Ghosh, Suhita, et al.
Veröffentlicht: (2024)
Controllable Singing Voice Synthesis using Phoneme-Level Energy Sequence
von: Ryu, Yerin, et al.
Veröffentlicht: (2025)
von: Ryu, Yerin, et al.
Veröffentlicht: (2025)
Spectral Mapping of Singing Voices: U-Net-Assisted Vocal Segmentation
von: Sorrenti, Adam
Veröffentlicht: (2024)
von: Sorrenti, Adam
Veröffentlicht: (2024)
LHQ-SVC: Lightweight and High Quality Singing Voice Conversion Modeling
von: Huang, Yubo, et al.
Veröffentlicht: (2024)
von: Huang, Yubo, et al.
Veröffentlicht: (2024)
A Real-Time Voice Activity Detection Based On Lightweight Neural
von: Jia, Jidong, et al.
Veröffentlicht: (2024)
von: Jia, Jidong, et al.
Veröffentlicht: (2024)
A Novel Labeled Human Voice Signal Dataset for Misbehavior Detection
von: Raza, Ali, et al.
Veröffentlicht: (2024)
von: Raza, Ali, et al.
Veröffentlicht: (2024)
EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion
von: Joglekar, Advait, et al.
Veröffentlicht: (2025)
von: Joglekar, Advait, et al.
Veröffentlicht: (2025)
Quantum-Inspired Audio Unlearning: Towards Privacy-Preserving Voice Biometrics
von: Pathak, Shreyansh, et al.
Veröffentlicht: (2025)
von: Pathak, Shreyansh, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Harder or Different? Understanding Generalization of Audio Deepfake Detection
von: Müller, Nicolas M., et al.
Veröffentlicht: (2024) -
DeePen: Penetration Testing for Audio Deepfake Detection
von: Müller, Nicolas, et al.
Veröffentlicht: (2025) -
MLAAD: The Multi-Language Audio Anti-Spoofing Dataset
von: Müller, Nicolas M., et al.
Veröffentlicht: (2024) -
Replay Attacks Against Audio Deepfake Detection
von: Müller, Nicolas, et al.
Veröffentlicht: (2025) -
Physics-Guided Deepfake Detection for Voice Authentication Systems
von: Mohammadi, Alireza, et al.
Veröffentlicht: (2025)