Degrading Voice: A Comprehensive Overview of Robust Voice Conversion Through Input Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Xining, Wei, Zhihua, Wang, Rui, Hu, Haixiao, Chen, Yanxiang, Han, Meng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sok: Comprehensive Security Overview, Challenges, and Future Directions of Voice-Controlled Systems
by: Xu, Haozhe, et al.
Published: (2024)
by: Xu, Haozhe, et al.
Published: (2024)
Quantifying Source Speaker Leakage in One-to-One Voice Conversion
by: Wellington, Scott, et al.
Published: (2025)
by: Wellington, Scott, et al.
Published: (2025)
Benchmarking Fake Voice Detection in the Fake Voice Generation Arms Race
by: Mao, Xutao, et al.
Published: (2025)
by: Mao, Xutao, et al.
Published: (2025)
Your Microphone Array Retains Your Identity: A Robust Voice Liveness Detection System for Smart Speakers
by: Meng, Yan, et al.
Published: (2025)
by: Meng, Yan, et al.
Published: (2025)
SuperEar: Eavesdropping on Mobile Voice Calls via Stealthy Acoustic Metamaterials
by: Ning, Zhiyuan, et al.
Published: (2025)
by: Ning, Zhiyuan, et al.
Published: (2025)
A Practical Survey on Emerging Threats from AI-driven Voice Attacks: How Vulnerable are Commercial Voice Control Systems?
by: Wang, Yuanda, et al.
Published: (2023)
by: Wang, Yuanda, et al.
Published: (2023)
Evaluating Synthetic Command Attacks on Smart Voice Assistants
by: He, Zhengxian, et al.
Published: (2024)
by: He, Zhengxian, et al.
Published: (2024)
VoiceWukong: Benchmarking Deepfake Voice Detection
by: Yan, Ziwei, et al.
Published: (2024)
by: Yan, Ziwei, et al.
Published: (2024)
Cross-Technology Generalization in Synthesized Speech Detection: Evaluating AST Models with Modern Voice Generators
by: Ustinov, Andrew, et al.
Published: (2025)
by: Ustinov, Andrew, et al.
Published: (2025)
VoiceMark: Zero-Shot Voice Cloning-Resistant Watermarking Approach Leveraging Speaker-Specific Latents
by: Li, Haiyun, et al.
Published: (2025)
by: Li, Haiyun, et al.
Published: (2025)
VocalCrypt: Novel Active Defense Against Deepfake Voice Based on Masking Effect
by: Fei, Qingyuan, et al.
Published: (2025)
by: Fei, Qingyuan, et al.
Published: (2025)
Efficient Streaming Voice Steganalysis in Challenging Detection Scenarios
by: Zhou, Pengcheng, et al.
Published: (2024)
by: Zhou, Pengcheng, et al.
Published: (2024)
Exploiting Context-dependent Duration Features for Voice Anonymization Attack Systems
by: Tomashenko, Natalia, et al.
Published: (2025)
by: Tomashenko, Natalia, et al.
Published: (2025)
VoxMorph: Scalable Zero-shot Voice Identity Morphing via Disentangled Embeddings
by: Krishnamurthy, Bharath, et al.
Published: (2026)
by: Krishnamurthy, Bharath, et al.
Published: (2026)
Pitch Imperfect: Detecting Audio Deepfakes Through Acoustic Prosodic Analysis
by: Warren, Kevin, et al.
Published: (2025)
by: Warren, Kevin, et al.
Published: (2025)
SpeechVerifier: Robust Acoustic Fingerprint against Tampering Attacks via Watermarking
by: Yao, Lingfeng, et al.
Published: (2025)
by: Yao, Lingfeng, et al.
Published: (2025)
Parallel Stacked Aggregated Network for Voice Authentication in IoT-Enabled Smart Devices
by: Khan, Awais, et al.
Published: (2024)
by: Khan, Awais, et al.
Published: (2024)
Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation
by: Zhang, Kuiyuan, et al.
Published: (2024)
by: Zhang, Kuiyuan, et al.
Published: (2024)
Two Views, One Truth: Spectral and Self-Supervised Features Fusion for Robust Speech Deepfake Detection
by: Kheir, Yassine El, et al.
Published: (2025)
by: Kheir, Yassine El, et al.
Published: (2025)
Hybrid Audio Detection Using Fine-Tuned Audio Spectrogram Transformers: A Dataset-Driven Evaluation of Mixed AI-Human Speech
by: Huang, Kunyang, et al.
Published: (2025)
by: Huang, Kunyang, et al.
Published: (2025)
De-AntiFake: Rethinking the Protective Perturbations Against Voice Cloning Attacks
by: Fan, Wei, et al.
Published: (2025)
by: Fan, Wei, et al.
Published: (2025)
Synthetic Voices, Real Threats: Evaluating Large Text-to-Speech Models in Generating Harmful Audio
by: Chen, Guangke, et al.
Published: (2025)
by: Chen, Guangke, et al.
Published: (2025)
SongBsAb: A Dual Prevention Approach against Singing Voice Conversion based Illegal Song Covers
by: Chen, Guangke, et al.
Published: (2024)
by: Chen, Guangke, et al.
Published: (2024)
CLAD: Robust Audio Deepfake Detection Against Manipulation Attacks with Contrastive Learning
by: Wu, Haolin, et al.
Published: (2024)
by: Wu, Haolin, et al.
Published: (2024)
A Universal Identity Backdoor Attack against Speaker Verification based on Siamese Network
by: Zhao, Haodong, et al.
Published: (2023)
by: Zhao, Haodong, et al.
Published: (2023)
The Man Behind the Sound: Demystifying Audio Private Attribute Profiling via Multimodal Large Language Model Agents
by: Wang, Lixu, et al.
Published: (2025)
by: Wang, Lixu, et al.
Published: (2025)
Phoneme-Based Proactive Anti-Eavesdropping with Controlled Recording Privilege
by: Huang, Peng, et al.
Published: (2024)
by: Huang, Peng, et al.
Published: (2024)
Preset-Voice Matching for Privacy Regulated Speech-to-Speech Translation Systems
by: Platnick, Daniel, et al.
Published: (2024)
by: Platnick, Daniel, et al.
Published: (2024)
TriniMark: A Robust Generative Speech Watermarking Method for Trinity-Level Traceability
by: Li, Yue, et al.
Published: (2025)
by: Li, Yue, et al.
Published: (2025)
WaLi: Can Pressure Sensors in HVAC Systems Capture Human Speech?
by: Tamiti, Tarikul Islam, et al.
Published: (2025)
by: Tamiti, Tarikul Islam, et al.
Published: (2025)
LENS-DF: Deepfake Detection and Temporal Localization for Long-Form Noisy Speech
by: Liu, Xuechen, et al.
Published: (2025)
by: Liu, Xuechen, et al.
Published: (2025)
Why Speech Deepfake Detectors Won't Generalize: The Limits of Detection in an Open World
by: Berisha, Visar, et al.
Published: (2025)
by: Berisha, Visar, et al.
Published: (2025)
Spoofing attack augmentation: can differently-trained attack models improve generalisation?
by: Ge, Wanying, et al.
Published: (2023)
by: Ge, Wanying, et al.
Published: (2023)
Zero-Query Adversarial Attack on Black-box Automatic Speech Recognition Systems
by: Fang, Zheng, et al.
Published: (2024)
by: Fang, Zheng, et al.
Published: (2024)
DIN-CTS: Low-Complexity Depthwise-Inception Neural Network with Contrastive Training Strategy for Deepfake Speech Detection
by: Pham, Lam, et al.
Published: (2025)
by: Pham, Lam, et al.
Published: (2025)
Privacy in Speech Technology
by: Bäckström, Tom
Published: (2023)
by: Bäckström, Tom
Published: (2023)
Making Acoustic Side-Channel Attacks on Noisy Keyboards Viable with LLM-Assisted Spectrograms' "Typo" Correction
by: Ayati, Seyyed Ali, et al.
Published: (2025)
by: Ayati, Seyyed Ali, et al.
Published: (2025)
Lightweight Protection for Privacy in Offloaded Speech Understanding
by: Cai, Dongqi
Published: (2024)
by: Cai, Dongqi
Published: (2024)
An RFP dataset for Real, Fake, and Partially fake audio detection
by: AlAli, Abdulazeez, et al.
Published: (2024)
by: AlAli, Abdulazeez, et al.
Published: (2024)
A Preliminary Case Study on Long-Form In-the-Wild Audio Spoofing Detection
by: Liu, Xuechen, et al.
Published: (2024)
by: Liu, Xuechen, et al.
Published: (2024)
Similar Items
-
Sok: Comprehensive Security Overview, Challenges, and Future Directions of Voice-Controlled Systems
by: Xu, Haozhe, et al.
Published: (2024) -
Quantifying Source Speaker Leakage in One-to-One Voice Conversion
by: Wellington, Scott, et al.
Published: (2025) -
Benchmarking Fake Voice Detection in the Fake Voice Generation Arms Race
by: Mao, Xutao, et al.
Published: (2025) -
Your Microphone Array Retains Your Identity: A Robust Voice Liveness Detection System for Smart Speakers
by: Meng, Yan, et al.
Published: (2025) -
SuperEar: Eavesdropping on Mobile Voice Calls via Stealthy Acoustic Metamaterials
by: Ning, Zhiyuan, et al.
Published: (2025)