Self Voice Conversion as an Attack against Neural Audio Watermarking
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Özer, Yigitcan, Ge, Wanying, Zhang, Zhe, Wang, Xin, Yamagishi, Junichi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Comprehensive Real-World Assessment of Audio Watermarking Algorithms: Will They Survive Neural Codecs?
von: Özer, Yigitcan, et al.
Veröffentlicht: (2025)
von: Özer, Yigitcan, et al.
Veröffentlicht: (2025)
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
von: Tang, Jingjing, et al.
Veröffentlicht: (2025)
von: Tang, Jingjing, et al.
Veröffentlicht: (2025)
HarmonicAttack: An Adaptive Cross-Domain Audio Watermark Removal
von: Li, Kexin, et al.
Veröffentlicht: (2025)
von: Li, Kexin, et al.
Veröffentlicht: (2025)
VoxEffects: A Speech-Oriented Audio Effects Dataset and Benchmark
von: Zhang, Zhe, et al.
Veröffentlicht: (2026)
von: Zhang, Zhe, et al.
Veröffentlicht: (2026)
Latent-Mark: An Audio Watermark Robust to Neural Resynthesis
von: Chen, Yen-Shan, et al.
Veröffentlicht: (2026)
von: Chen, Yen-Shan, et al.
Veröffentlicht: (2026)
LENS-DF: Deepfake Detection and Temporal Localization for Long-Form Noisy Speech
von: Liu, Xuechen, et al.
Veröffentlicht: (2025)
von: Liu, Xuechen, et al.
Veröffentlicht: (2025)
Audio Codec Augmentation for Robust Collaborative Watermarking of Speech Synthesis
von: Juvela, Lauri, et al.
Veröffentlicht: (2024)
von: Juvela, Lauri, et al.
Veröffentlicht: (2024)
Spectral Masking and Interpolation Attack (SMIA): A Black-box Adversarial Attack against Voice Authentication and Anti-Spoofing Systems
von: Kamel, Kamel, et al.
Veröffentlicht: (2025)
von: Kamel, Kamel, et al.
Veröffentlicht: (2025)
Zero-Day Audio DeepFake Detection via Retrieval Augmentation and Profile Matching
von: Liu, Xuechen, et al.
Veröffentlicht: (2025)
von: Liu, Xuechen, et al.
Veröffentlicht: (2025)
Protecting Your Voice: Temporal-aware Robust Watermarking
von: Li, Yue, et al.
Veröffentlicht: (2025)
von: Li, Yue, et al.
Veröffentlicht: (2025)
Prosody-Adaptable Audio Codecs for Zero-Shot Voice Conversion via In-Context Learning
von: Zhao, Junchuan, et al.
Veröffentlicht: (2025)
von: Zhao, Junchuan, et al.
Veröffentlicht: (2025)
Proactive Detection of Voice Cloning with Localized Watermarking
von: Roman, Robin San, et al.
Veröffentlicht: (2024)
von: Roman, Robin San, et al.
Veröffentlicht: (2024)
AudioMotionBench: Evaluating Auditory Motion Perception in Audio LLMs
von: Sun, Zhe, et al.
Veröffentlicht: (2025)
von: Sun, Zhe, et al.
Veröffentlicht: (2025)
SAVe: Self-Supervised Audio-visual Deepfake Detection Exploiting Visual Artifacts and Audio-visual Misalignment
von: Shahzad, Sahibzada Adil, et al.
Veröffentlicht: (2026)
von: Shahzad, Sahibzada Adil, et al.
Veröffentlicht: (2026)
Quantifying Source Speaker Leakage in One-to-One Voice Conversion
von: Wellington, Scott, et al.
Veröffentlicht: (2025)
von: Wellington, Scott, et al.
Veröffentlicht: (2025)
Speech-Audio Compositional Attacks on Multimodal LLMs and Their Mitigation with SALMONN-Guard
von: Yang, Yudong, et al.
Veröffentlicht: (2025)
von: Yang, Yudong, et al.
Veröffentlicht: (2025)
Codec-Robust Attacks on Audio LLMs
von: Roh, Jaechul, et al.
Veröffentlicht: (2026)
von: Roh, Jaechul, et al.
Veröffentlicht: (2026)
SPA-SVC: Self-supervised Pitch Augmentation for Singing Voice Conversion
von: Bai, Bingsong, et al.
Veröffentlicht: (2024)
von: Bai, Bingsong, et al.
Veröffentlicht: (2024)
SelfVC: Voice Conversion With Iterative Refinement using Self Transformations
von: Neekhara, Paarth, et al.
Veröffentlicht: (2023)
von: Neekhara, Paarth, et al.
Veröffentlicht: (2023)
EmoAttack: Utilizing Emotional Voice Conversion for Speech Backdoor Attacks on Deep Speech Classification Models
von: Yao, Wenhan, et al.
Veröffentlicht: (2024)
von: Yao, Wenhan, et al.
Veröffentlicht: (2024)
AudioJailbreak: Jailbreak Attacks against End-to-End Large Audio-Language Models
von: Chen, Guangke, et al.
Veröffentlicht: (2025)
von: Chen, Guangke, et al.
Veröffentlicht: (2025)
Spoofing attack augmentation: can differently-trained attack models improve generalisation?
von: Ge, Wanying, et al.
Veröffentlicht: (2023)
von: Ge, Wanying, et al.
Veröffentlicht: (2023)
VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models
von: Chen, Yukun, et al.
Veröffentlicht: (2026)
von: Chen, Yukun, et al.
Veröffentlicht: (2026)
IntrinsicVoice: Empowering LLMs with Intrinsic Real-time Voice Interaction Abilities
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
The ISCSLP 2024 Conversational Voice Clone (CoVoC) Challenge: Tasks, Results and Findings
von: Xia, Kangxiang, et al.
Veröffentlicht: (2024)
von: Xia, Kangxiang, et al.
Veröffentlicht: (2024)
StyleStream: Real-Time Zero-Shot Voice Style Conversion
von: Liu, Yisi, et al.
Veröffentlicht: (2026)
von: Liu, Yisi, et al.
Veröffentlicht: (2026)
Defense Against Synthetic Speech: Real-Time Detection of RVC Voice Conversion Attacks
von: Chinchmalatpure, Prajwal, et al.
Veröffentlicht: (2025)
von: Chinchmalatpure, Prajwal, et al.
Veröffentlicht: (2025)
A Comparative Study on Proactive and Passive Detection of Deepfake Speech
von: Wu, Chia-Hua, et al.
Veröffentlicht: (2025)
von: Wu, Chia-Hua, et al.
Veröffentlicht: (2025)
LHQ-SVC: Lightweight and High Quality Singing Voice Conversion Modeling
von: Huang, Yubo, et al.
Veröffentlicht: (2024)
von: Huang, Yubo, et al.
Veröffentlicht: (2024)
ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
SegReConcat: A Data Augmentation Method for Voice Anonymization Attack
von: Arefeen, Ridwan, et al.
Veröffentlicht: (2025)
von: Arefeen, Ridwan, et al.
Veröffentlicht: (2025)
GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking
von: Wang, Yunqiang, et al.
Veröffentlicht: (2026)
von: Wang, Yunqiang, et al.
Veröffentlicht: (2026)
Generating Synthetic Doctor-Patient Conversations for Long-form Audio Summarization
von: Labrak, Yanis, et al.
Veröffentlicht: (2026)
von: Labrak, Yanis, et al.
Veröffentlicht: (2026)
A Preliminary Case Study on Long-Form In-the-Wild Audio Spoofing Detection
von: Liu, Xuechen, et al.
Veröffentlicht: (2024)
von: Liu, Xuechen, et al.
Veröffentlicht: (2024)
Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels
von: Weng, Yuzhe, et al.
Veröffentlicht: (2026)
von: Weng, Yuzhe, et al.
Veröffentlicht: (2026)
YingMusic-SVC: Real-World Robust Zero-Shot Singing Voice Conversion with Flow-GRPO and Singing-Specific Inductive Biases
von: Chen, Gongyu, et al.
Veröffentlicht: (2025)
von: Chen, Gongyu, et al.
Veröffentlicht: (2025)
Audio-to-Image Encoding for Improved Voice Characteristic Detection Using Deep Convolutional Neural Networks
von: Atif, Youness
Veröffentlicht: (2025)
von: Atif, Youness
Veröffentlicht: (2025)
Neural Multi-Speaker Voice Cloning for Nepali in Low-Resource Settings
von: Shrestha, Aayush M., et al.
Veröffentlicht: (2026)
von: Shrestha, Aayush M., et al.
Veröffentlicht: (2026)
Breaking Audio Large Language Models by Attacking Only the Encoder: A Universal Targeted Latent-Space Audio Attack
von: Ziv, Roee, et al.
Veröffentlicht: (2025)
von: Ziv, Roee, et al.
Veröffentlicht: (2025)
Deepfake Audio Detection Using Self-supervised Fusion Representations
von: Zaman, Khalid, et al.
Veröffentlicht: (2026)
von: Zaman, Khalid, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Comprehensive Real-World Assessment of Audio Watermarking Algorithms: Will They Survive Neural Codecs?
von: Özer, Yigitcan, et al.
Veröffentlicht: (2025) -
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
von: Tang, Jingjing, et al.
Veröffentlicht: (2025) -
HarmonicAttack: An Adaptive Cross-Domain Audio Watermark Removal
von: Li, Kexin, et al.
Veröffentlicht: (2025) -
VoxEffects: A Speech-Oriented Audio Effects Dataset and Benchmark
von: Zhang, Zhe, et al.
Veröffentlicht: (2026) -
Latent-Mark: An Audio Watermark Robust to Neural Resynthesis
von: Chen, Yen-Shan, et al.
Veröffentlicht: (2026)