Quantum-Inspired Audio Unlearning: Towards Privacy-Preserving Voice Biometrics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pathak, Shreyansh, Shreshtha, Sonu, Singh, Richa, Vatsa, Mayank |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SynHate: Detecting Hate Speech in Synthetic Deepfake Audio
von: Ranjan, Rishabh, et al.
Veröffentlicht: (2025)
von: Ranjan, Rishabh, et al.
Veröffentlicht: (2025)
Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection?
von: Dutta, Bikash, et al.
Veröffentlicht: (2025)
von: Dutta, Bikash, et al.
Veröffentlicht: (2025)
Multimodal Zero-Shot Framework for Deepfake Hate Speech Detection in Low-Resource Languages
von: Ranjan, Rishabh, et al.
Veröffentlicht: (2025)
von: Ranjan, Rishabh, et al.
Veröffentlicht: (2025)
LAPS-Diff: A Diffusion-Based Framework for Singing Voice Synthesis With Language Aware Prosody-Style Guided Learning
von: Dhar, Sandipan, et al.
Veröffentlicht: (2025)
von: Dhar, Sandipan, et al.
Veröffentlicht: (2025)
Do Not Mimic My Voice: Speaker Identity Unlearning for Zero-Shot Text-to-Speech
von: Kim, Taesoo, et al.
Veröffentlicht: (2025)
von: Kim, Taesoo, et al.
Veröffentlicht: (2025)
VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
von: Anastassiou, Philip, et al.
Veröffentlicht: (2024)
von: Anastassiou, Philip, et al.
Veröffentlicht: (2024)
Towards Privacy-Preserving Audio Classification Systems
von: Chhaglani, Bhawana, et al.
Veröffentlicht: (2024)
von: Chhaglani, Bhawana, et al.
Veröffentlicht: (2024)
Prosody-Adaptable Audio Codecs for Zero-Shot Voice Conversion via In-Context Learning
von: Zhao, Junchuan, et al.
Veröffentlicht: (2025)
von: Zhao, Junchuan, et al.
Veröffentlicht: (2025)
RBA-FE: A Robust Brain-Inspired Audio Feature Extractor for Depression Diagnosis
von: Wu, Yu-Xuan, et al.
Veröffentlicht: (2025)
von: Wu, Yu-Xuan, et al.
Veröffentlicht: (2025)
Audio-to-Image Encoding for Improved Voice Characteristic Detection Using Deep Convolutional Neural Networks
von: Atif, Youness
Veröffentlicht: (2025)
von: Atif, Youness
Veröffentlicht: (2025)
Towards Controllable Audio Texture Morphing
von: Gupta, Chitralekha, et al.
Veröffentlicht: (2023)
von: Gupta, Chitralekha, et al.
Veröffentlicht: (2023)
FunAudioLLM: Voice Understanding and Generation Foundation Models for Natural Interaction Between Humans and LLMs
von: An, Keyu, et al.
Veröffentlicht: (2024)
von: An, Keyu, et al.
Veröffentlicht: (2024)
WavShape: Information-Theoretic Speech Representation Learning for Fair and Privacy-Aware Audio Processing
von: Baser, Oguzhan, et al.
Veröffentlicht: (2025)
von: Baser, Oguzhan, et al.
Veröffentlicht: (2025)
EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion
von: Joglekar, Advait, et al.
Veröffentlicht: (2025)
von: Joglekar, Advait, et al.
Veröffentlicht: (2025)
FusionAudio-1.2M: Towards Fine-grained Audio Captioning with Multimodal Contextual Fusion
von: Chen, Shunian, et al.
Veröffentlicht: (2025)
von: Chen, Shunian, et al.
Veröffentlicht: (2025)
Towards Better Disentanglement in Non-Autoregressive Zero-Shot Expressive Voice Conversion
von: Akti, Seymanur, et al.
Veröffentlicht: (2025)
von: Akti, Seymanur, et al.
Veröffentlicht: (2025)
Tuning Music Education: AI-Powered Personalization in Learning Music
von: Sanganeria, Mayank, et al.
Veröffentlicht: (2024)
von: Sanganeria, Mayank, et al.
Veröffentlicht: (2024)
TTMBA: Towards Text To Multiple Sources Binaural Audio Generation
von: He, Yuxuan, et al.
Veröffentlicht: (2025)
von: He, Yuxuan, et al.
Veröffentlicht: (2025)
Towards Attention-based Contrastive Learning for Audio Spoof Detection
von: Goel, Chirag, et al.
Veröffentlicht: (2024)
von: Goel, Chirag, et al.
Veröffentlicht: (2024)
SafeEar: Content Privacy-Preserving Audio Deepfake Detection
von: Li, Xinfeng, et al.
Veröffentlicht: (2024)
von: Li, Xinfeng, et al.
Veröffentlicht: (2024)
SoulX-Singer: Towards High-Quality Zero-Shot Singing Voice Synthesis
von: Qian, Jiale, et al.
Veröffentlicht: (2026)
von: Qian, Jiale, et al.
Veröffentlicht: (2026)
Towards Leveraging Contrastively Pretrained Neural Audio Embeddings for Recommender Tasks
von: Grötschla, Florian, et al.
Veröffentlicht: (2024)
von: Grötschla, Florian, et al.
Veröffentlicht: (2024)
Quantum-Inspired Genetic Algorithm for Robust Source Separation in Smart City Acoustics
von: Quan, Minh K., et al.
Veröffentlicht: (2025)
von: Quan, Minh K., et al.
Veröffentlicht: (2025)
CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training
von: Du, Zhihao, et al.
Veröffentlicht: (2025)
von: Du, Zhihao, et al.
Veröffentlicht: (2025)
LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging
von: Singh, Shubhr, et al.
Veröffentlicht: (2025)
von: Singh, Shubhr, et al.
Veröffentlicht: (2025)
R2-SVC: Towards Real-World Robust and Expressive Zero-shot Singing Voice Conversion
von: Zheng, Junjie, et al.
Veröffentlicht: (2025)
von: Zheng, Junjie, et al.
Veröffentlicht: (2025)
HQ-SVC: Towards High-Quality Zero-Shot Singing Voice Conversion in Low-Resource Scenarios
von: Bai, Bingsong, et al.
Veröffentlicht: (2025)
von: Bai, Bingsong, et al.
Veröffentlicht: (2025)
Audio Mamba: Pretrained Audio State Space Model For Audio Tagging
von: Lin, Jiaju, et al.
Veröffentlicht: (2024)
von: Lin, Jiaju, et al.
Veröffentlicht: (2024)
Audio Deepfake Detection in the Age of Advanced Text-to-Speech models
von: Singh, Robin, et al.
Veröffentlicht: (2026)
von: Singh, Robin, et al.
Veröffentlicht: (2026)
The Sounds of Home: A Speech-Removed Residential Audio Dataset for Sound Event Detection
von: Bibbó, Gabriel, et al.
Veröffentlicht: (2024)
von: Bibbó, Gabriel, et al.
Veröffentlicht: (2024)
Audio Atlas: Visualizing and Exploring Audio Datasets
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2024)
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2024)
VoiceGRPO: Modern MoE Transformers with Group Relative Policy Optimization GRPO for AI Voice Health Care Applications on Voice Pathology Detection
von: Togootogtokh, Enkhtogtokh, et al.
Veröffentlicht: (2025)
von: Togootogtokh, Enkhtogtokh, et al.
Veröffentlicht: (2025)
Voice Cloning: Comprehensive Survey
von: Azzuni, Hussam, et al.
Veröffentlicht: (2025)
von: Azzuni, Hussam, et al.
Veröffentlicht: (2025)
Do Music Source Separation Models Preserve Spatial Information in Binaural Audio?
von: Namballa, Richa, et al.
Veröffentlicht: (2025)
von: Namballa, Richa, et al.
Veröffentlicht: (2025)
Representation-Regularized Convolutional Audio Transformer for Audio Understanding
von: Han, Bing, et al.
Veröffentlicht: (2026)
von: Han, Bing, et al.
Veröffentlicht: (2026)
Audio Spatially-Guided Fusion for Audio-Visual Navigation
von: Zhou, Xinyu, et al.
Veröffentlicht: (2026)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2026)
Voice EHR: Introducing Multimodal Audio Data for Health
von: Anibal, James, et al.
Veröffentlicht: (2024)
von: Anibal, James, et al.
Veröffentlicht: (2024)
EnCLAP: Combining Neural Audio Codec and Audio-Text Joint Embedding for Automated Audio Captioning
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2024)
AudioTurbo: Fast Text-to-Audio Generation with Rectified Diffusion
von: Zhao, Junqi, et al.
Veröffentlicht: (2025)
von: Zhao, Junqi, et al.
Veröffentlicht: (2025)
DreamAudio: Customized Text-to-Audio Generation with Diffusion Models
von: Yuan, Yi, et al.
Veröffentlicht: (2025)
von: Yuan, Yi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SynHate: Detecting Hate Speech in Synthetic Deepfake Audio
von: Ranjan, Rishabh, et al.
Veröffentlicht: (2025) -
Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection?
von: Dutta, Bikash, et al.
Veröffentlicht: (2025) -
Multimodal Zero-Shot Framework for Deepfake Hate Speech Detection in Low-Resource Languages
von: Ranjan, Rishabh, et al.
Veröffentlicht: (2025) -
LAPS-Diff: A Diffusion-Based Framework for Singing Voice Synthesis With Language Aware Prosody-Style Guided Learning
von: Dhar, Sandipan, et al.
Veröffentlicht: (2025) -
Do Not Mimic My Voice: Speaker Identity Unlearning for Zero-Shot Text-to-Speech
von: Kim, Taesoo, et al.
Veröffentlicht: (2025)