Latent Secret Spin: Keyed Orthogonal Rotations for Blind Speech Watermarking in Anisotropic Latent Spaces
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Coletta, Emma, Todisco, Massimiliano, Panariello, Michele, Faonio, Antonio, Evans, Nicholas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Preserving spoken content in voice anonymisation with character-level vocoder conditioning
von: Panariello, Michele, et al.
Veröffentlicht: (2024)
von: Panariello, Michele, et al.
Veröffentlicht: (2024)
Reference-free Adversarial Sex Obfuscation in Speech
von: Qu, Yangyang, et al.
Veröffentlicht: (2025)
von: Qu, Yangyang, et al.
Veröffentlicht: (2025)
Speaker anonymization using neural audio codec language models
von: Panariello, Michele, et al.
Veröffentlicht: (2023)
von: Panariello, Michele, et al.
Veröffentlicht: (2023)
MDD: a Mask Diffusion Detector to Protect Speaker Verification Systems from Adversarial Perturbations
von: Bai, Yibo, et al.
Veröffentlicht: (2025)
von: Bai, Yibo, et al.
Veröffentlicht: (2025)
The Risks and Detection of Overestimated Privacy Protection in Voice Anonymisation
von: Panariello, Michele, et al.
Veröffentlicht: (2025)
von: Panariello, Michele, et al.
Veröffentlicht: (2025)
Evaluating voice anonymisation using similarity rank disclosure
von: Chandra, Shilpa, et al.
Veröffentlicht: (2026)
von: Chandra, Shilpa, et al.
Veröffentlicht: (2026)
The VoicePrivacy 2022 Challenge: Progress and Perspectives in Voice Anonymisation
von: Panariello, Michele, et al.
Veröffentlicht: (2024)
von: Panariello, Michele, et al.
Veröffentlicht: (2024)
Explaining deep learning models for spoofing and deepfake detection with SHapley Additive exPlanations
von: Ge, Wanying, et al.
Veröffentlicht: (2021)
von: Ge, Wanying, et al.
Veröffentlicht: (2021)
Malacopula: adversarial automatic speaker verification attacks using a neural-based generalised Hammerstein model
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024)
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024)
Fair-Gate: Fairness-Aware Interpretable Risk Gating for Sex-Fair Voice Biometrics
von: Qu, Yangyang, et al.
Veröffentlicht: (2026)
von: Qu, Yangyang, et al.
Veröffentlicht: (2026)
Latent Filling: Latent Space Data Augmentation for Zero-shot Speech Synthesis
von: Bae, Jae-Sung, et al.
Veröffentlicht: (2023)
von: Bae, Jae-Sung, et al.
Veröffentlicht: (2023)
A Comparison of Differential Performance Metrics for the Evaluation of Automatic Speaker Verification Fairness
von: Chouchane, Oubaida, et al.
Veröffentlicht: (2024)
von: Chouchane, Oubaida, et al.
Veröffentlicht: (2024)
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2026)
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2026)
Latent Watermarking of Audio Generative Models
von: Roman, Robin San, et al.
Veröffentlicht: (2024)
von: Roman, Robin San, et al.
Veröffentlicht: (2024)
The VoicePrivacy 2024 Challenge Evaluation Plan
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2024)
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2024)
Voice Privacy Preservation with Multiple Random Orthogonal Secret Keys: Attack Resistance Analysis
von: Tanaka, Kohei, et al.
Veröffentlicht: (2025)
von: Tanaka, Kohei, et al.
Veröffentlicht: (2025)
Spoofing attack augmentation: can differently-trained attack models improve generalisation?
von: Ge, Wanying, et al.
Veröffentlicht: (2023)
von: Ge, Wanying, et al.
Veröffentlicht: (2023)
Blind Acoustic Parameter Estimation Through Task-Agnostic Embeddings Using Latent Approximations
von: Götz, Philipp, et al.
Veröffentlicht: (2024)
von: Götz, Philipp, et al.
Veröffentlicht: (2024)
Investigation of Speech and Noise Latent Representations in Single-channel VAE-based Speech Enhancement
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
Leveraging Discriminative Latent Representations for Conditioning GAN-Based Speech Enhancement
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2025)
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2025)
Semantic-VAE: Semantic-Alignment Latent Representation for Better Speech Synthesis
von: Niu, Zhikang, et al.
Veröffentlicht: (2025)
von: Niu, Zhikang, et al.
Veröffentlicht: (2025)
On the Relation Between Speech Quality and Quantized Latent Representations of Neural Codecs
von: Halimeh, Mhd Modar, et al.
Veröffentlicht: (2025)
von: Halimeh, Mhd Modar, et al.
Veröffentlicht: (2025)
LongCat-AudioDiT: High-Fidelity Diffusion Text-to-Speech in the Waveform Latent Space
von: Xin, Detai, et al.
Veröffentlicht: (2026)
von: Xin, Detai, et al.
Veröffentlicht: (2026)
Speech Synthesis From Continuous Features Using Per-Token Latent Diffusion
von: Turetzky, Arnon, et al.
Veröffentlicht: (2024)
von: Turetzky, Arnon, et al.
Veröffentlicht: (2024)
Impairments are Clustered in Latents of Deep Neural Network-based Speech Quality Models
von: Cumlin, Fredrik, et al.
Veröffentlicht: (2025)
von: Cumlin, Fredrik, et al.
Veröffentlicht: (2025)
Simulating Native Speaker Shadowing for Nonnative Speech Assessment with Latent Speech Representations
von: Geng, Haopeng, et al.
Veröffentlicht: (2024)
von: Geng, Haopeng, et al.
Veröffentlicht: (2024)
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis
von: Wu, Chun Yat, et al.
Veröffentlicht: (2025)
von: Wu, Chun Yat, et al.
Veröffentlicht: (2025)
FakeMark: Deepfake Speech Attribution With Watermarked Artifacts
von: Ge, Wanying, et al.
Veröffentlicht: (2025)
von: Ge, Wanying, et al.
Veröffentlicht: (2025)
SimpleSpeech: Towards Simple and Efficient Text-to-Speech with Scalar Latent Transformer Diffusion Models
von: Yang, Dongchao, et al.
Veröffentlicht: (2024)
von: Yang, Dongchao, et al.
Veröffentlicht: (2024)
Latent-Level Enhancement with Flow Matching for Robust Automatic Speech Recognition
von: Yang, Da-Hee, et al.
Veröffentlicht: (2026)
von: Yang, Da-Hee, et al.
Veröffentlicht: (2026)
DiffDSR: Dysarthric Speech Reconstruction Using Latent Diffusion Model
von: Chen, Xueyuan, et al.
Veröffentlicht: (2025)
von: Chen, Xueyuan, et al.
Veröffentlicht: (2025)
WAKE: Watermarking Audio with Key Enrichment
von: Xu, Yaoxun, et al.
Veröffentlicht: (2025)
von: Xu, Yaoxun, et al.
Veröffentlicht: (2025)
Single Channel Blind Dereverberation of Speech Signals
von: Nigam, Dhruv
Veröffentlicht: (2025)
von: Nigam, Dhruv
Veröffentlicht: (2025)
Efficient Speech Language Modeling via Energy Distance in Continuous Latent Space
von: Ma, Zhengrui, et al.
Veröffentlicht: (2025)
von: Ma, Zhengrui, et al.
Veröffentlicht: (2025)
Conditional Latent Diffusion-Based Speech Enhancement Via Dual Context Learning
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
Deep Audio Watermarks are Shallow: Limitations of Post-Hoc Watermarking Techniques for Speech
von: O'Reilly, Patrick, et al.
Veröffentlicht: (2025)
von: O'Reilly, Patrick, et al.
Veröffentlicht: (2025)
TraceableSpeech: Towards Proactively Traceable Text-to-Speech with Watermarking
von: Zhou, Junzuo, et al.
Veröffentlicht: (2024)
von: Zhou, Junzuo, et al.
Veröffentlicht: (2024)
Generalizable Audio Deepfake Detection via Latent Space Refinement and Augmentation
von: Huang, Wen, et al.
Veröffentlicht: (2025)
von: Huang, Wen, et al.
Veröffentlicht: (2025)
DiTSE: High-Fidelity Generative Speech Enhancement via Latent Diffusion Transformers
von: Guimarães, Heitor R., et al.
Veröffentlicht: (2025)
von: Guimarães, Heitor R., et al.
Veröffentlicht: (2025)
VQalAttent: a Transparent Speech Generation Pipeline based on Transformer-learned VQ-VAE Latent Space
von: Rodriguez, Armani, et al.
Veröffentlicht: (2024)
von: Rodriguez, Armani, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Preserving spoken content in voice anonymisation with character-level vocoder conditioning
von: Panariello, Michele, et al.
Veröffentlicht: (2024) -
Reference-free Adversarial Sex Obfuscation in Speech
von: Qu, Yangyang, et al.
Veröffentlicht: (2025) -
Speaker anonymization using neural audio codec language models
von: Panariello, Michele, et al.
Veröffentlicht: (2023) -
MDD: a Mask Diffusion Detector to Protect Speaker Verification Systems from Adversarial Perturbations
von: Bai, Yibo, et al.
Veröffentlicht: (2025) -
The Risks and Detection of Overestimated Privacy Protection in Voice Anonymisation
von: Panariello, Michele, et al.
Veröffentlicht: (2025)