TVTSyn: Content-Synchronous Time-Varying Timbre for Streaming Voice Conversion and Anonymization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Quamer, Waris, Tseng, Mu-Ruei, Nasrallah, Ghady, Gutierrez-Osuna, Ricardo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PHONOS: PHOnetic Neutralization for Online Streaming Applications
von: Quamer, Waris, et al.
Veröffentlicht: (2026)
von: Quamer, Waris, et al.
Veröffentlicht: (2026)
DarkStream: real-time speech anonymization with low latency
von: Quamer, Waris, et al.
Veröffentlicht: (2025)
von: Quamer, Waris, et al.
Veröffentlicht: (2025)
Disentangling segmental and prosodic factors to non-native speech comprehensibility
von: Quamer, Waris, et al.
Veröffentlicht: (2024)
von: Quamer, Waris, et al.
Veröffentlicht: (2024)
End-to-end streaming model for low-latency speech anonymization
von: Quamer, Waris, et al.
Veröffentlicht: (2024)
von: Quamer, Waris, et al.
Veröffentlicht: (2024)
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2026)
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2026)
Zero-Shot Voice Conversion via Content-Aware Timbre Ensemble and Conditional Flow Matching
von: Pan, Yu, et al.
Veröffentlicht: (2024)
von: Pan, Yu, et al.
Veröffentlicht: (2024)
StreamVoice+: Evolving into End-to-end Streaming Zero-shot Voice Conversion
von: Wang, Zhichao, et al.
Veröffentlicht: (2024)
von: Wang, Zhichao, et al.
Veröffentlicht: (2024)
ControlVC: Zero-Shot Voice Conversion with Time-Varying Controls on Pitch and Speed
von: Chen, Meiying, et al.
Veröffentlicht: (2022)
von: Chen, Meiying, et al.
Veröffentlicht: (2022)
Analysis of Speech Temporal Dynamics in the Context of Speaker Verification and Voice Anonymization
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2024)
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2024)
Takin-VC: Expressive Zero-Shot Voice Conversion via Adaptive Hybrid Content Encoding and Enhanced Timbre Modeling
von: Yang, Yuguang, et al.
Veröffentlicht: (2024)
von: Yang, Yuguang, et al.
Veröffentlicht: (2024)
QvTAD: Differential Relative Attribute Learning for Voice Timbre Attribute Detection
von: Wu, Zhiyu, et al.
Veröffentlicht: (2025)
von: Wu, Zhiyu, et al.
Veröffentlicht: (2025)
Adversarial Multi-Task Learning for Disentangling Timbre and Pitch in Singing Voice Synthesis
von: Kim, Tae-Woo, et al.
Veröffentlicht: (2022)
von: Kim, Tae-Woo, et al.
Veröffentlicht: (2022)
On the Impact of Voice Anonymization on Speech Diagnostic Applications: a Case Study on COVID-19 Detection
von: Zhu, Yi, et al.
Veröffentlicht: (2023)
von: Zhu, Yi, et al.
Veröffentlicht: (2023)
DualVC 2: Dynamic Masked Convolution for Unified Streaming and Non-Streaming Voice Conversion
von: Ning, Ziqian, et al.
Veröffentlicht: (2023)
von: Ning, Ziqian, et al.
Veröffentlicht: (2023)
StreamVoice: Streamable Context-Aware Language Modeling for Real-time Zero-Shot Voice Conversion
von: Wang, Zhichao, et al.
Veröffentlicht: (2024)
von: Wang, Zhichao, et al.
Veröffentlicht: (2024)
USM-VC: Mitigating Timbre Leakage with Universal Semantic Mapping Residual Block for Voice Conversion
von: Li, Na, et al.
Veröffentlicht: (2025)
von: Li, Na, et al.
Veröffentlicht: (2025)
AdaptVC: High Quality Voice Conversion with Adaptive Learning
von: Kim, Jaehun, et al.
Veröffentlicht: (2025)
von: Kim, Jaehun, et al.
Veröffentlicht: (2025)
StreamVC: Real-Time Low-Latency Voice Conversion
von: Yang, Yang, et al.
Veröffentlicht: (2024)
von: Yang, Yang, et al.
Veröffentlicht: (2024)
MeanVC: Lightweight and Streaming Zero-Shot Voice Conversion via Mean Flows
von: Ma, Guobin, et al.
Veröffentlicht: (2025)
von: Ma, Guobin, et al.
Veröffentlicht: (2025)
Stepback: Enhanced Disentanglement for Voice Conversion via Multi-Task Learning
von: Yang, Qian, et al.
Veröffentlicht: (2025)
von: Yang, Qian, et al.
Veröffentlicht: (2025)
Streaming Non-Autoregressive Model for Accent Conversion and Pronunciation Improvement
von: Nguyen, Tuan-Nam, et al.
Veröffentlicht: (2025)
von: Nguyen, Tuan-Nam, et al.
Veröffentlicht: (2025)
SynthVC: Leveraging Synthetic Data for End-to-End Low Latency Streaming Voice Conversion
von: Guo, Zhao, et al.
Veröffentlicht: (2025)
von: Guo, Zhao, et al.
Veröffentlicht: (2025)
Attacking Voice Anonymization Systems with Augmented Feature and Speaker Identity Difference
von: Zhang, Yanzhe, et al.
Veröffentlicht: (2024)
von: Zhang, Yanzhe, et al.
Veröffentlicht: (2024)
The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan
von: Sheng, Zhengyan, et al.
Veröffentlicht: (2025)
von: Sheng, Zhengyan, et al.
Veröffentlicht: (2025)
Voice Conversion Improves Cross-Domain Robustness for Spoken Arabic Dialect Identification
von: Abdullah, Badr M., et al.
Veröffentlicht: (2025)
von: Abdullah, Badr M., et al.
Veröffentlicht: (2025)
Voice Conversion for Lombard Speaking Style with Implicit and Explicit Acoustic Feature Conditioning
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2025)
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2025)
Conan: A Chunkwise Online Network for Zero-Shot Adaptive Voice Conversion
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
Noro: Noise-Robust One-shot Voice Conversion with Hidden Speaker Representation Learning
von: He, Haorui, et al.
Veröffentlicht: (2024)
von: He, Haorui, et al.
Veröffentlicht: (2024)
SEF-VC: Speaker Embedding Free Zero-Shot Voice Conversion with Cross Attention
von: Li, Junjie, et al.
Veröffentlicht: (2023)
von: Li, Junjie, et al.
Veröffentlicht: (2023)
Probing the Feasibility of Multilingual Speaker Anonymization
von: Meyer, Sarina, et al.
Veröffentlicht: (2024)
von: Meyer, Sarina, et al.
Veröffentlicht: (2024)
A Benchmark for Multi-speaker Anonymization
von: Miao, Xiaoxiao, et al.
Veröffentlicht: (2024)
von: Miao, Xiaoxiao, et al.
Veröffentlicht: (2024)
Timbre Difference Capturing in Anomalous Sound Detection
von: Nishida, Tomoya, et al.
Veröffentlicht: (2024)
von: Nishida, Tomoya, et al.
Veröffentlicht: (2024)
Exploiting Context-dependent Duration Features for Voice Anonymization Attack Systems
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2025)
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2025)
Spatial Voice Conversion: Voice Conversion Preserving Spatial Information and Non-target Signals
von: Seki, Kentaro, et al.
Veröffentlicht: (2024)
von: Seki, Kentaro, et al.
Veröffentlicht: (2024)
Towards Inclusive ASR: Investigating Voice Conversion for Dysarthric Speech Recognition in Low-Resource Languages
von: Li, Chin-Jou, et al.
Veröffentlicht: (2025)
von: Li, Chin-Jou, et al.
Veröffentlicht: (2025)
Custom Data Augmentation for low resource ASR using Bark and Retrieval-Based Voice Conversion
von: Kamble, Anand, et al.
Veröffentlicht: (2023)
von: Kamble, Anand, et al.
Veröffentlicht: (2023)
Auden-Voice: General-Purpose Voice Encoder for Speech and Language Understanding
von: Huo, Mingyue, et al.
Veröffentlicht: (2025)
von: Huo, Mingyue, et al.
Veröffentlicht: (2025)
Assessing the Alignment of Audio Representations with Timbre Similarity Ratings
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
DiariST: Streaming Speech Translation with Speaker Diarization
von: Yang, Mu, et al.
Veröffentlicht: (2023)
von: Yang, Mu, et al.
Veröffentlicht: (2023)
Neural Concatenative Singing Voice Conversion: Rethinking Concatenation-Based Approach for One-Shot Singing Voice Conversion
von: Sha, Binzhu, et al.
Veröffentlicht: (2023)
von: Sha, Binzhu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
PHONOS: PHOnetic Neutralization for Online Streaming Applications
von: Quamer, Waris, et al.
Veröffentlicht: (2026) -
DarkStream: real-time speech anonymization with low latency
von: Quamer, Waris, et al.
Veröffentlicht: (2025) -
Disentangling segmental and prosodic factors to non-native speech comprehensibility
von: Quamer, Waris, et al.
Veröffentlicht: (2024) -
End-to-end streaming model for low-latency speech anonymization
von: Quamer, Waris, et al.
Veröffentlicht: (2024) -
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2026)