SPA-SVC: Self-supervised Pitch Augmentation for Singing Voice Conversion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bai, Bingsong, Wang, Fengping, Gao, Yingming, Li, Ya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HQ-SVC: Towards High-Quality Zero-Shot Singing Voice Conversion in Low-Resource Scenarios
von: Bai, Bingsong, et al.
Veröffentlicht: (2025)
von: Bai, Bingsong, et al.
Veröffentlicht: (2025)
LHQ-SVC: Lightweight and High Quality Singing Voice Conversion Modeling
von: Huang, Yubo, et al.
Veröffentlicht: (2024)
von: Huang, Yubo, et al.
Veröffentlicht: (2024)
R2-SVC: Towards Real-World Robust and Expressive Zero-shot Singing Voice Conversion
von: Zheng, Junjie, et al.
Veröffentlicht: (2025)
von: Zheng, Junjie, et al.
Veröffentlicht: (2025)
VibE-SVC: Vibrato Extraction with High-frequency F0 Contour for Singing Voice Conversion
von: Choi, Joon-Seung, et al.
Veröffentlicht: (2025)
von: Choi, Joon-Seung, et al.
Veröffentlicht: (2025)
CoMoSVC: Consistency Model-based Singing Voice Conversion
von: Lu, Yiwen, et al.
Veröffentlicht: (2024)
von: Lu, Yiwen, et al.
Veröffentlicht: (2024)
FreeSVC: Towards Zero-shot Multilingual Singing Voice Conversion
von: Ferreira, Alef Iury Siqueira, et al.
Veröffentlicht: (2025)
von: Ferreira, Alef Iury Siqueira, et al.
Veröffentlicht: (2025)
SYKI-SVC: Advancing Singing Voice Conversion with Post-Processing Innovations and an Open-Source Professional Testset
von: Zhou, Yiquan, et al.
Veröffentlicht: (2025)
von: Zhou, Yiquan, et al.
Veröffentlicht: (2025)
kNN-SVC: Robust Zero-Shot Singing Voice Conversion with Additive Synthesis and Concatenation Smoothness Optimization
von: Shao, Keren, et al.
Veröffentlicht: (2025)
von: Shao, Keren, et al.
Veröffentlicht: (2025)
SingFake: Singing Voice Deepfake Detection
von: Zang, Yongyi, et al.
Veröffentlicht: (2023)
von: Zang, Yongyi, et al.
Veröffentlicht: (2023)
CONTUNER: Singing Voice Beautifying with Pitch and Expressiveness Condition
von: Wang, Jianzong, et al.
Veröffentlicht: (2024)
von: Wang, Jianzong, et al.
Veröffentlicht: (2024)
RobustSVC: HuBERT-based Melody Extractor and Adversarial Learning for Robust Singing Voice Conversion
von: Chen, Wei, et al.
Veröffentlicht: (2024)
von: Chen, Wei, et al.
Veröffentlicht: (2024)
LDM-SVC: Latent Diffusion Model Based Zero-Shot Any-to-Any Singing Voice Conversion with Singer Guidance
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
Self-Supervised Singing Voice Pre-Training towards Speech-to-Singing Conversion
von: Li, Ruiqi, et al.
Veröffentlicht: (2024)
von: Li, Ruiqi, et al.
Veröffentlicht: (2024)
LCM-SVC: Latent Diffusion Model Based Singing Voice Conversion with Inference Acceleration via Latent Consistency Distillation
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
von: Chen, Shihao, et al.
Veröffentlicht: (2024)
Deepfake Detection of Singing Voices With Whisper Encodings
von: Sharma, Falguni, et al.
Veröffentlicht: (2025)
von: Sharma, Falguni, et al.
Veröffentlicht: (2025)
DiTSinger: Scaling Singing Voice Synthesis with Diffusion Transformer and Implicit Alignment
von: Du, Zongcai, et al.
Veröffentlicht: (2025)
von: Du, Zongcai, et al.
Veröffentlicht: (2025)
RDSinger: Reference-based Diffusion Network for Singing Voice Synthesis
von: Sui, Kehan, et al.
Veröffentlicht: (2024)
von: Sui, Kehan, et al.
Veröffentlicht: (2024)
Neural Concatenative Singing Voice Conversion: Rethinking Concatenation-Based Approach for One-Shot Singing Voice Conversion
von: Sha, Binzhu, et al.
Veröffentlicht: (2023)
von: Sha, Binzhu, et al.
Veröffentlicht: (2023)
Spectral Mapping of Singing Voices: U-Net-Assisted Vocal Segmentation
von: Sorrenti, Adam
Veröffentlicht: (2024)
von: Sorrenti, Adam
Veröffentlicht: (2024)
Controllable Singing Voice Synthesis using Phoneme-Level Energy Sequence
von: Ryu, Yerin, et al.
Veröffentlicht: (2025)
von: Ryu, Yerin, et al.
Veröffentlicht: (2025)
SelfVC: Voice Conversion With Iterative Refinement using Self Transformations
von: Neekhara, Paarth, et al.
Veröffentlicht: (2023)
von: Neekhara, Paarth, et al.
Veröffentlicht: (2023)
Adversarial Multi-Task Learning for Disentangling Timbre and Pitch in Singing Voice Synthesis
von: Kim, Tae-Woo, et al.
Veröffentlicht: (2022)
von: Kim, Tae-Woo, et al.
Veröffentlicht: (2022)
SoulX-Singer: Towards High-Quality Zero-Shot Singing Voice Synthesis
von: Qian, Jiale, et al.
Veröffentlicht: (2026)
von: Qian, Jiale, et al.
Veröffentlicht: (2026)
Speech Foundation Model Ensembles for the Controlled Singing Voice Deepfake Detection (CtrSVDD) Challenge 2024
von: Guragain, Anmol, et al.
Veröffentlicht: (2024)
von: Guragain, Anmol, et al.
Veröffentlicht: (2024)
A Mamba-based Network for Semi-supervised Singing Melody Extraction Using Confidence Binary Regularization
von: He, Xiaoliang, et al.
Veröffentlicht: (2025)
von: He, Xiaoliang, et al.
Veröffentlicht: (2025)
Mitigating Latent Mismatch in cVAE-Based Singing Voice Synthesis via Flow Matching
von: Yun, Minhyeok, et al.
Veröffentlicht: (2026)
von: Yun, Minhyeok, et al.
Veröffentlicht: (2026)
Everyone-Can-Sing: Zero-Shot Singing Voice Synthesis and Conversion with Speech Reference
von: Dai, Shuqi, et al.
Veröffentlicht: (2025)
von: Dai, Shuqi, et al.
Veröffentlicht: (2025)
PESTO: Real-Time Pitch Estimation with Self-supervised Transposition-equivariant Objective
von: Riou, Alain, et al.
Veröffentlicht: (2025)
von: Riou, Alain, et al.
Veröffentlicht: (2025)
DJCM: A Deep Joint Cascade Model for Singing Voice Separation and Vocal Pitch Estimation
von: Wei, Haojie, et al.
Veröffentlicht: (2024)
von: Wei, Haojie, et al.
Veröffentlicht: (2024)
RRPO: Robust Reward Policy Optimization for LLM-based Emotional TTS
von: Wang, Cong, et al.
Veröffentlicht: (2025)
von: Wang, Cong, et al.
Veröffentlicht: (2025)
SingMOS-Pro: An Comprehensive Benchmark for Singing Quality Assessment
von: Tang, Yuxun, et al.
Veröffentlicht: (2025)
von: Tang, Yuxun, et al.
Veröffentlicht: (2025)
Multi-Loss Learning for Speech Emotion Recognition with Energy-Adaptive Mixup and Frame-Level Attention
von: Wang, Cong, et al.
Veröffentlicht: (2025)
von: Wang, Cong, et al.
Veröffentlicht: (2025)
Auffusion: Leveraging the Power of Diffusion and Large Language Models for Text-to-Audio Generation
von: Xue, Jinlong, et al.
Veröffentlicht: (2024)
von: Xue, Jinlong, et al.
Veröffentlicht: (2024)
LAPS-Diff: A Diffusion-Based Framework for Singing Voice Synthesis With Language Aware Prosody-Style Guided Learning
von: Dhar, Sandipan, et al.
Veröffentlicht: (2025)
von: Dhar, Sandipan, et al.
Veröffentlicht: (2025)
USM-VC: Mitigating Timbre Leakage with Universal Semantic Mapping Residual Block for Voice Conversion
von: Li, Na, et al.
Veröffentlicht: (2025)
von: Li, Na, et al.
Veröffentlicht: (2025)
vec2wav 2.0: Advancing Voice Conversion via Discrete Token Vocoders
von: Guo, Yiwei, et al.
Veröffentlicht: (2024)
von: Guo, Yiwei, et al.
Veröffentlicht: (2024)
Prompt-Singer: Controllable Singing-Voice-Synthesis with Natural Language Prompt
von: Wang, Yongqi, et al.
Veröffentlicht: (2024)
von: Wang, Yongqi, et al.
Veröffentlicht: (2024)
Retrieval Augmented Generation in Prompt-based Text-to-Speech Synthesis with Context-Aware Contrastive Language-Audio Pretraining
von: Xue, Jinlong, et al.
Veröffentlicht: (2024)
von: Xue, Jinlong, et al.
Veröffentlicht: (2024)
SVDD Challenge 2024: A Singing Voice Deepfake Detection Challenge Evaluation Plan
von: Zhang, You, et al.
Veröffentlicht: (2024)
von: Zhang, You, et al.
Veröffentlicht: (2024)
Prosody-Adaptable Audio Codecs for Zero-Shot Voice Conversion via In-Context Learning
von: Zhao, Junchuan, et al.
Veröffentlicht: (2025)
von: Zhao, Junchuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HQ-SVC: Towards High-Quality Zero-Shot Singing Voice Conversion in Low-Resource Scenarios
von: Bai, Bingsong, et al.
Veröffentlicht: (2025) -
LHQ-SVC: Lightweight and High Quality Singing Voice Conversion Modeling
von: Huang, Yubo, et al.
Veröffentlicht: (2024) -
R2-SVC: Towards Real-World Robust and Expressive Zero-shot Singing Voice Conversion
von: Zheng, Junjie, et al.
Veröffentlicht: (2025) -
VibE-SVC: Vibrato Extraction with High-frequency F0 Contour for Singing Voice Conversion
von: Choi, Joon-Seung, et al.
Veröffentlicht: (2025) -
CoMoSVC: Consistency Model-based Singing Voice Conversion
von: Lu, Yiwen, et al.
Veröffentlicht: (2024)